The Spanish example (Miguel) is really bad.
Bark – Text-prompted generative audio model
11–20 of 99 posts
Re: Bark – Text-prompted generative audio model
#12Re: Bark – Text-prompted generative audio model
#13Well i can see it becoming sexy soon
Soon? Is the model filtered/censored?
>Bark has the capability to fully clone voices - including tone, pitch, emotion and prosody. The model also attempts to preserve music, ambient noise, etc. from input audio. However, to mitigate misuse of this technology, we limit the audio history prompts to a limited set of Suno-provided, fully synthetic options to choose from for each language.
It's not immediately clear how the audio history prompts are created.
Re: Bark – Text-prompted generative audio model
#14Re: Bark – Text-prompted generative audio model
#15It seems to be easily reproducible if I specify a non-existing speaker?
audio_array = generate_audio(text_prompt, 'en_speaker_3')
Re: Bark – Text-prompted generative audio model
#16Any idea what the training data for this is? Looking at the model, it looks like it is literally just copy-paste from Karpathy's nanoGPT, so the training data is what's most interesting. Pretty amazing anyway.
Re: Bark – Text-prompted generative audio model
#17Very cool. Side note: bark-gpt.com is already taken for a dog translator: "The world’s first AI powered, real-time communications tool between humans and their furry best friends."[0] I only know this because my law firm partner's name is Bark, and I wanted to automate some legal work and name the software "Bark GPT" after him. [0] https://www.bark-gpt.com/
I want to see the training set for this
Re: Bark – Text-prompted generative audio model
#18Earlier quoted context omitted.
Soon? Is the model filtered/censored?
From the readme: >Bark has the capability to fully clone voices - including tone, pitch, emotion and prosody. The model also attempts to preserve music, ambient noise, etc. from input audio. However, to mitigate misuse of this technology, we limit the audio history prompts to a limited set of Suno-provided, fully synthetic options to choose from for each language. It's not immediately clear how the audio history prom…
Re: Bark – Text-prompted generative audio model
#19Re: Bark – Text-prompted generative audio model
#20Earlier quoted context omitted.
Soon? Is the model filtered/censored?
From the readme: >Bark has the capability to fully clone voices - including tone, pitch, emotion and prosody. The model also attempts to preserve music, ambient noise, etc. from input audio. However, to mitigate misuse of this technology, we limit the audio history prompts to a limited set of Suno-provided, fully synthetic options to choose from for each language. It's not immediately clear how the audio history prom…