Viewing profile — andrew-w
andrew-w
HN member- Joined
- Fri, Dec 15, 2023, 2:09 PM UTC
- HN karma
- 32
- Public activity
- 57 items
- HN profile
- View on Hacker News ↗
About andrew-w
No profile information was provided.
Recent public activity
-
comment
Comment #46801663
We have not released the weights, but it is fully available to use in your websites or applications. I can see how our wording there could be misconstrued -- sorry about that.
-
comment
Comment #46797033
So glad you enjoyed it! We've been able to significantly reduce those text hallucinations with a few tricks, but it seems they haven't been fully squashed. The /imagine command onl…
-
comment
Comment #46796960
Not something we had thought to do tbh, but would definitely enhance the experience. And, should be reasonable to do. Thanks!
-
comment
Comment #46796655
We have not released the weights, but it is fully available to use in your websites or applications. I can see how our wording there could be misconstrued -- sorry about that. You …
-
comment
Comment #46796587
Haha, I kind of get that reaction. Convincing the world "this was hard to do" is generally not easy. Re: user uploads, we're operating in good faith at the moment (no built-in IP m…
-
comment
Comment #46796361
Thank you! Impressive demo with OVA. Still feels very snappy, even fully local. It will be interesting to see how video plays out in that regard. I think we're still at least a yea…
-
comment
Comment #46790728
I wonder how it would come across with the right voice. We're focused on building out the video layer tech, but at the end of the day, the voice is also pretty important for a posi…
-
comment
Comment #46790709
Thanks for the feedback. The current avatars use a STT-LLM-TTS pipeline (rather than true speech-to-speech), which limits nuanced understanding of pronunciations. Speech-to-speech …
-
comment
Comment #46789216
This isn't natively supported -- we are continuously streaming frames throughout the conversation session that are generated in real-time. If you were building your own conversatio…
-
comment
Comment #46789092
Thanks! And sorry! I can see how our wording there could be misconstrued. With a real-time model, the streaming infrastructure matters almost as much as the weights themselves. It …
-
comment
Comment #46788124
Yep, the model is running on Hopper architecture. Anything less was not sufficient in our experiments.
-
comment
Comment #43841316
Thanks for trying it out! character.ai has put their model behind a waitlist, so it's hard to compare. As far as I can tell, they don't appear to make any specific claims about spe…
-
comment
Comment #43807956
thanks for trying us out!
-
comment
Comment #43796689
Just added a signup at the bottom of the technical report: https://lemonslice.com/live/technical-report
-
comment
Comment #43796244
glad to bring a little joy into the world :)
-
comment
Comment #43795401
It works with any style of character! Check out the embedded videos in our tech report. Peachy and the toilet are my favorite. https://lemonslice.com/live/technical-report
-
comment
Comment #43794965
Thanks! We think we can cut down the latency to <2s which should make it feel even more natural.
-
comment
Comment #43794929
Thanks! What kind of use case are you thinking about?
-
comment
Comment #43794900
It's something we are considering. What use cases do you have in mind?
-
comment
Comment #43794856
We've been very inspired by interactive character experiences powered by traditional VFX + puppetry (turtle talk with crush is a favorite). I think that sort of interactive enterta…
-
comment
Comment #43794671
Thanks for the feedback. This is definitely a demo where every piece matters for maximizing the enjoyment factor. We spent the most effort on optimizing video quality and latency, …
-
comment
Comment #43794006
Not relying on facial keypoints means we can animate a wide range of non-humanoid characters. My favorite is talking to the Doge meme.
-
comment
Comment #43793924
One way this differs is in the model architecture. Our approach relies on a single pass of a diffusion transformer (DiT), whereas Live Portrait relies on intermediate representatio…
-
comment
Comment #43793709
I spent about 2 hours recording videos with different characters. Of course, the one I made as a joke for myself and never intended to share was the most enjoyable to watch :)
-
comment
Comment #43787824
Just added as a public character :)