Live data from Hacker News

Show HN: Infinity – Realistic AI characters that can speak

news.ycombinator.com

151–160 of 320 posts

Re: Show HN: Infinity – Realistic AI characters that can speak

#151

Earlier quoted context omitted.

Very cool! If we release an API, you could use it across the different Ragdoll experiences you're creating. I agree personalized character experiences are going to be a huge thing. FYI we plan to allow users to save their own characters (an image + voice combo) soon

> If we release an API, you could use it Absolutely, especially if the pricing makes sense! Would be very nice to just focus on the creative suite which is the real product, and less on the AI infra of hosting models, vector dbs, and paying for GPU. Curious if you're using providers for models or self-hosting?

We use Modal for cloud compute and autoscaling. The model is our own.

Re: Show HN: Infinity – Realistic AI characters that can speak

#152
post #116

Tried my hardest to push this into the uncanny valley. I did, but it was pretty hard. Seems robust. https://6ammc3n5zzf5ljnz.public.blob.vercel-storage.com/inf2...

Not robust enough to work against a sketch https://6ammc3n5zzf5ljnz.public.blob.vercel-storage.com/inf2... though perhaps it rebelled against the message

I had difficulty getting my lemming to speak. After selecting several alternatives, I tried one with a more defined, open mouth, which required multiple attempts but mostly worked. Additional iterations on the same image can produce different results.

Re: Show HN: Infinity – Realistic AI characters that can speak

#153
I am actively working in this area from a wrapper application perspective. In general, tools that generate video are not sufficient on their own. They are likely to be used as part of some larger video-production workflow.

One drawback of tools like runway (and midjourney) is the lack of an API allowing integration into products. I would love to re-sell your service to my clients as part of a larger offering. Is this something you plan to offer?

The examples are very promising by the way.

Re: Show HN: Infinity – Realistic AI characters that can speak

#154

can we choose our own voices?

The web app does allow you to upload any audio, but in order to use your voice, you would need to either record a sample for each video or clone your voice with a 3rd party TTS provider. We would like to make it easier to do all that within our site - hopefully soon!

Re: Show HN: Infinity – Realistic AI characters that can speak

#155

Earlier quoted context omitted.

> If we release an API, you could use it Absolutely, especially if the pricing makes sense! Would be very nice to just focus on the creative suite which is the real product, and less on the AI infra of hosting models, vector dbs, and paying for GPU. Curious if you're using providers for models or self-hosting?

We use Modal for cloud compute and autoscaling. The model is our own.

Amazing, great to hear it :)

Re: Show HN: Infinity – Realistic AI characters that can speak

#157

As soon as I saw the "Gnome" face option I gnew exactly what I gneeded to do: https://6ammc3n5zzf5ljnz.public.blob.vercel-storage.com/inf2... EDIT: looks like the model doesn't like Duke Nukem: https://6ammc3n5zzf5ljnz.public.blob.vercel-storage.com/inf2... Cropping out his pistol only made it worse lol: https://6ammc3n5zzf5ljnz.public.blob.vercel-storage.com/inf2... A different image works a little bit better, thoug…

[deleted]

Re: Show HN: Infinity – Realistic AI characters that can speak

#159

[flagged]

We are big fans of Hedra. Do you know if they've publicly commented on their model architecture? As far as we know, our particular choice of an end-to-end diffusion + transformer is novel. We don't know what Hedra is doing. It could be the approach EMO has taken ( https://humanaigc.github.io/emote-portrait-alive/ ) or VASA ( https://www.microsoft.com/en-us/research/project/vasa-1/ ) or Loopy Avatar ( https://loopyava…

Michael from Hedra, your choice is not novel :)

Re: Show HN: Infinity – Realistic AI characters that can speak

#160
post #158

Awesome, any plans for an API and, if so, how soon?

No plans at the moment, but there seems to be a decent amount of interest here. Our main focus has been making the model as good as it can be, since there are still many failure modes. What kind of application would you use it for?
Post reply on HN