Live data from Hacker News

Viewing profile — andrew-w

andrew-w

HN member
Joined
Fri, Dec 15, 2023, 2:09 PM UTC
HN karma
32
Public activity
57 items

About andrew-w

No profile information was provided.

Recent public activity

  1. comment
    Comment #46801663

    We have not released the weights, but it is fully available to use in your websites or applications. I can see how our wording there could be misconstrued -- sorry about that.

  2. comment
    Comment #46797033

    So glad you enjoyed it! We've been able to significantly reduce those text hallucinations with a few tricks, but it seems they haven't been fully squashed. The /imagine command onl…

  3. comment
    Comment #46796960

    Not something we had thought to do tbh, but would definitely enhance the experience. And, should be reasonable to do. Thanks!

  4. comment
    Comment #46796655

    We have not released the weights, but it is fully available to use in your websites or applications. I can see how our wording there could be misconstrued -- sorry about that. You …

  5. comment
    Comment #46796587

    Haha, I kind of get that reaction. Convincing the world "this was hard to do" is generally not easy. Re: user uploads, we're operating in good faith at the moment (no built-in IP m…

  6. comment
    Comment #46796361

    Thank you! Impressive demo with OVA. Still feels very snappy, even fully local. It will be interesting to see how video plays out in that regard. I think we're still at least a yea…

  7. comment
    Comment #46790728

    I wonder how it would come across with the right voice. We're focused on building out the video layer tech, but at the end of the day, the voice is also pretty important for a posi…

  8. comment
    Comment #46790709

    Thanks for the feedback. The current avatars use a STT-LLM-TTS pipeline (rather than true speech-to-speech), which limits nuanced understanding of pronunciations. Speech-to-speech …

  9. comment
    Comment #46789216

    This isn't natively supported -- we are continuously streaming frames throughout the conversation session that are generated in real-time. If you were building your own conversatio…

  10. comment
    Comment #46789092

    Thanks! And sorry! I can see how our wording there could be misconstrued. With a real-time model, the streaming infrastructure matters almost as much as the weights themselves. It …

  11. comment
    Comment #46788124

    Yep, the model is running on Hopper architecture. Anything less was not sufficient in our experiments.

  12. comment
    Comment #43841316

    Thanks for trying it out! character.ai has put their model behind a waitlist, so it's hard to compare. As far as I can tell, they don't appear to make any specific claims about spe…

  13. comment
    Comment #43807956

    thanks for trying us out!

  14. comment
    Comment #43796689

    Just added a signup at the bottom of the technical report: https://lemonslice.com/live/technical-report

  15. comment
    Comment #43796244

    glad to bring a little joy into the world :)

  16. comment
    Comment #43795401

    It works with any style of character! Check out the embedded videos in our tech report. Peachy and the toilet are my favorite. https://lemonslice.com/live/technical-report

  17. comment
    Comment #43794965

    Thanks! We think we can cut down the latency to <2s which should make it feel even more natural.

  18. comment
    Comment #43794929

    Thanks! What kind of use case are you thinking about?

  19. comment
    Comment #43794900

    It's something we are considering. What use cases do you have in mind?

  20. comment
    Comment #43794856

    We've been very inspired by interactive character experiences powered by traditional VFX + puppetry (turtle talk with crush is a favorite). I think that sort of interactive enterta…

  21. comment
    Comment #43794671

    Thanks for the feedback. This is definitely a demo where every piece matters for maximizing the enjoyment factor. We spent the most effort on optimizing video quality and latency, …

  22. comment
    Comment #43794006

    Not relying on facial keypoints means we can animate a wide range of non-humanoid characters. My favorite is talking to the Doge meme.

  23. comment
    Comment #43793924

    One way this differs is in the model architecture. Our approach relies on a single pass of a diffusion transformer (DiT), whereas Live Portrait relies on intermediate representatio…

  24. comment
    Comment #43793709

    I spent about 2 hours recording videos with different characters. Of course, the one I made as a joke for myself and never intended to share was the most enjoyable to watch :)

  25. comment
    Comment #43787824

    Just added as a public character :)