Live data from Hacker News

Viewing profile — bcherry

bcherry

HN member
Joined
Mon, Sep 20, 2010, 9:53 PM UTC
HN karma
353
Public activity
111 items

About bcherry

No profile information was provided.

Recent public activity

  1. comment
    Comment #47420920

    this really reminds me of the "worst volume control" from reddit https://uxdesign.cc/the-worst-volume-control-ui-in-the-world...

  2. comment
    Comment #47148244

    you mention voice ai in the announcement but I wonder how this works in practice. most voice AI systems are bound not by full response latency but just by time-to-first-non-reasoni…

  3. comment
    Comment #46735117

    This isn't really the author's point, but I think one effect of AI and the forthcoming robotics revolution will be the unrolling of a lot of consolidated supply chains for all sort…

  4. comment
    Comment #45843907

    wow thanks for leaving this comment - i now realize two things: 1. the farmer's almanac i thought of when i saw the title and even read the article is not going anywhere 2. i have …

  5. comment
    Comment #45727186

    they'd have to be extra careful with cpython, it's got a lot of include

  6. comment
    Comment #44831150

    yeah i think they shot themselves in the foot a bit here by creating the o series. the truth is that GPT-5 _is_ a huge step forward, for the "GPT-x" models. The current GPT-x model…

  7. comment
    Comment #44032017

    "The sculpture is already complete within the marble block, before I start my work. It is already there, I just have to chisel away the superfluous material." - Michelangelo

  8. comment
    Comment #42936179

    Chat is a great UX _around_ development tools. Imagine having a pair programmer and never being allowed to speak to them. You could only communicate by taking over the keyboard and…

  9. comment
    Comment #42009634

    a little glossed over, but they do point out that most important improvement o1 has over gpt-4o is not it's "correct" score improving from 38% to 42% but actually it's "not attempt…

  10. comment
    Comment #41997250

    disagree - good products meet their users where they are and bury complexity under the hood. i can't imagine trying to use a calendar app (or any app really) that refuses to operat…

  11. comment
    Comment #41927789

    It's kind of interesting because I think most people implementing RAG aren't even thinking about tokenization at all. They're thinking about embeddings: 1. chunk the corpus of data…

  12. comment
    Comment #41725490

    hey sorry about that - ran into a snag with the API but we got it back online an hour ago! hope you get another chance to take a look! reply

  13. comment
    Comment #41725489

    hey sorry about that - ran into a snag with the API but we got it back online an hour ago! hope you get another chance to take a look!

  14. comment
    Comment #41717409

    this one was actually so much fun I built it into the defaults https://playground.livekit.io/?preset=doom

  15. comment
    Comment #41716617

    It sure can https://playground.livekit.io/?preset=0tfwypgx7&instructions...

  16. story
    Show HN: Speech-to-speech playground for OpenAI's new Realtime API

    Hi there - Ben from LiveKit here! If you’re curious about OpenAI’s brand-new Realtime API and speech-to-speech model, check out this hosted playground and play with the model yours…

  17. story
  18. comment
    Comment #41715725

    yes and you can use it in text-text mode if you want. a key benefit is for turn-based usages (where you have running back and forth between user and assistant) you only need to sen…

  19. comment
    Comment #41715003

    correct - you should also be able to save a lot by skipping their built-in VAD and doing turn detection (if you need it) locally to avoid paying for silent inputs.

  20. comment
    Comment #41714101

    keep in mind that this is just v1 of the realtime api. they'll add realtime vision/video down the road which can also have wide applications beyond synchronous communication.

  21. comment
    Comment #41714081

    yes it transcribes inputs automatically, but not in realtime. outputs are sent in text + audio but you'll get the text very quickly and audio a bit slower, and of course the audio …

  22. comment
    Comment #41712262

    No, it's the same thing as ChatGPT advanced voice. Full speech-to-speech model.

  23. comment
    Comment #41373583

    LiveKit, Cartesia, Deepgram, and Vercel

  24. comment
    Comment #40121026

    there's a difference between "click bait" (a misleading title specifically crafted to drive instinctual interest but which is not an accurate summation of the content) and titles a…

  25. comment
    Comment #39568532

    I wonder if they're allowed at SFO?