The video shows scenarios of people listening to pdfs of pretty dense material (e.g., computer science, bio mechanics). Does anyone here actually have positive results doing this? It seems to me listening to anything that's even remotely complex with the intent of learning it just isn't something that's feasible.
ElevenReader
31–40 of 183 posts
Re: ElevenReader
#32This is definitely the future, I'm worried about the electric slip and slide world we're heading into though, where everything is completely spoonfed and consumptive. I can't help but think we're heading back into animalism.
Could you expand upon this? Any milestones towards that which we should be mindful of?
Re: ElevenReader
#33Is this streaming server-side audio or is the TTS running locally on device ? Can it work offline ?
you could build your local TTS using kokoro browser though — https://huggingface.co/spaces/webml-community/kokoro-webgpu
Re: ElevenReader
#34You can get pretty close with open source software: https://claudio.uk/posts/audiblez-v4.html
Re: ElevenReader
#35E. g. saying "1963" when the actual year in the text was 1967. Yeah, the voices sound very realistic. But I'm not sure how useful that is if you can't trust the spoken words.
Does anyone know if it got better in the last weeks?
Re: ElevenReader
#36You can get pretty close with open source software: https://claudio.uk/posts/audiblez-v4.html
How does it hold up on long stuff? I use Elevenlabs Studio daily and once things start to get into the chapters long, the voice can really start to go off the rails. It'd say they've solved a lot of this over the past 2/3 months, but it does still happen on long stuff.
Re: ElevenReader
#37The video shows scenarios of people listening to pdfs of pretty dense material (e.g., computer science, bio mechanics). Does anyone here actually have positive results doing this? It seems to me listening to anything that's even remotely complex with the intent of learning it just isn't something that's feasible.
Re: ElevenReader
#38I would never trust the company that acquired Omnivore only to sunset it with 2 weeks notice to retrieve data. Companies won't stop pulling this garbage unless we stop supporting them.
Re: ElevenReader
#39I would never trust the company that acquired Omnivore only to sunset it with 2 weeks notice to retrieve data. Companies won't stop pulling this garbage unless we stop supporting them.
You can fight back by supporting and advocating for open source foundation text to speech models. XTTS, GptSoVits, Tortoise, Zonos, etc. Open source models drive proprietary foundation models' margin to zero. The only reason elevenlabs became a unicorn was their margin. If they became a commodity, they'd find themselves in a deep pit.
Re: ElevenReader
#40The video shows scenarios of people listening to pdfs of pretty dense material (e.g., computer science, bio mechanics). Does anyone here actually have positive results doing this? It seems to me listening to anything that's even remotely complex with the intent of learning it just isn't something that's feasible.