Viewing profile — bcherry
bcherry
HN member- Joined
- Mon, Sep 20, 2010, 9:53 PM UTC
- HN karma
- 353
- Public activity
- 111 items
- HN profile
- View on Hacker News ↗
About bcherry
No profile information was provided.
Recent public activity
-
comment
Comment #47420920
this really reminds me of the "worst volume control" from reddit https://uxdesign.cc/the-worst-volume-control-ui-in-the-world...
-
comment
Comment #47148244
you mention voice ai in the announcement but I wonder how this works in practice. most voice AI systems are bound not by full response latency but just by time-to-first-non-reasoni…
-
comment
Comment #46735117
This isn't really the author's point, but I think one effect of AI and the forthcoming robotics revolution will be the unrolling of a lot of consolidated supply chains for all sort…
-
comment
Comment #45843907
wow thanks for leaving this comment - i now realize two things: 1. the farmer's almanac i thought of when i saw the title and even read the article is not going anywhere 2. i have …
-
comment
Comment #45727186
they'd have to be extra careful with cpython, it's got a lot of include
-
comment
Comment #44831150
yeah i think they shot themselves in the foot a bit here by creating the o series. the truth is that GPT-5 _is_ a huge step forward, for the "GPT-x" models. The current GPT-x model…
-
comment
Comment #44032017
"The sculpture is already complete within the marble block, before I start my work. It is already there, I just have to chisel away the superfluous material." - Michelangelo
-
comment
Comment #42936179
Chat is a great UX _around_ development tools. Imagine having a pair programmer and never being allowed to speak to them. You could only communicate by taking over the keyboard and…
-
comment
Comment #42009634
a little glossed over, but they do point out that most important improvement o1 has over gpt-4o is not it's "correct" score improving from 38% to 42% but actually it's "not attempt…
-
comment
Comment #41997250
disagree - good products meet their users where they are and bury complexity under the hood. i can't imagine trying to use a calendar app (or any app really) that refuses to operat…
-
comment
Comment #41927789
It's kind of interesting because I think most people implementing RAG aren't even thinking about tokenization at all. They're thinking about embeddings: 1. chunk the corpus of data…
-
comment
Comment #41725490
hey sorry about that - ran into a snag with the API but we got it back online an hour ago! hope you get another chance to take a look! reply
-
comment
Comment #41725489
hey sorry about that - ran into a snag with the API but we got it back online an hour ago! hope you get another chance to take a look!
-
comment
Comment #41717409
this one was actually so much fun I built it into the defaults https://playground.livekit.io/?preset=doom
-
comment
Comment #41716617
It sure can https://playground.livekit.io/?preset=0tfwypgx7&instructions...
-
story
Show HN: Speech-to-speech playground for OpenAI's new Realtime API
Hi there - Ben from LiveKit here! If you’re curious about OpenAI’s brand-new Realtime API and speech-to-speech model, check out this hosted playground and play with the model yours…
- story
-
comment
Comment #41715725
yes and you can use it in text-text mode if you want. a key benefit is for turn-based usages (where you have running back and forth between user and assistant) you only need to sen…
-
comment
Comment #41715003
correct - you should also be able to save a lot by skipping their built-in VAD and doing turn detection (if you need it) locally to avoid paying for silent inputs.
-
comment
Comment #41714101
keep in mind that this is just v1 of the realtime api. they'll add realtime vision/video down the road which can also have wide applications beyond synchronous communication.
-
comment
Comment #41714081
yes it transcribes inputs automatically, but not in realtime. outputs are sent in text + audio but you'll get the text very quickly and audio a bit slower, and of course the audio …
-
comment
Comment #41712262
No, it's the same thing as ChatGPT advanced voice. Full speech-to-speech model.
-
comment
Comment #41373583
LiveKit, Cartesia, Deepgram, and Vercel
-
comment
Comment #40121026
there's a difference between "click bait" (a misleading title specifically crafted to drive instinctual interest but which is not an accurate summation of the content) and titles a…
-
comment
Comment #39568532
I wonder if they're allowed at SFO?