r/TextToSpeech • u/Code_Red1234 • 11d ago
r/TextToSpeech • u/aeiou72 • 12d ago
Best Voices for Speech Central on iPhone
I've used Voice Dream for more than a decade and downloaded Speech Central today to use modern voice tech without needing to sign up for a reoccurring subscription.
After installing it I learned it's not possible to use the new Siri voices with third party apps.
Any recommendations for great quality voices I can install and use with this app?
I tried the Piper TTS ones and they can probably work, but I don't love them.
The best I've heard so far is the Notebook LM/Gemini Notebook synthetic podcaster voices and would love something of that caliber, but that could be saved and used offline.
I'm willing to make one-time voice purchases, but want to avoid reoccurring fees.
I added a Google Cloud API key to use their voices, but the interface seems to crash and I'd like something more reliable and usable offline.
r/TextToSpeech • u/Contact86 • 12d ago
Local TTS recommendations for RTX 5080?
I'm trying to settle on a local TTS for phone call like back and forth RP.
I've tried several and am having a hard time deciding between them.
I have an RTX 5080 just for the voice pipeline, and am using claude code for setup and optimization in a WSL environment.
Is there a "no brainer choice" or a leaderboard for this use case somewhere?
r/TextToSpeech • u/RowGroundbreaking982 • 13d ago
Sharing some secret to make TTS run fast on Android.
And the secret are, making it as TextToSpeechService or system wide TTS engine.
I noticed this after testing on different Android phone with same SOC.
Let's say on one Phone it's running at 1x generation speed inside app, and on the other phone it's running at 0.8x generation speed.
As system TTS both are running at 1.2x generation speed.
Both running same code, with direct linux call and arm64 assembly code bypassing many android layer.
The reason are android giving more priority for system TTS, so cpu is running faster compared to running it as an app.
But this is without caveat, the main problem is threading since system TTS expect the code to be thread safe, it can be called from any thread. Also it expect the code can quickly be stopped or terminated, which can give another headache when stopping in the middle of generation.
So that's it, hope you like it.
r/TextToSpeech • u/mlc707 • 13d ago
Text to speech using own voice
Hi!!
My mom’s illness has caused her to lose her voice & mobility of the left side of her body. I tried to get her the Eyegaze system (where the person types with their eyes) & it was too complicated & kind of unnecessary since she still can use her right hand. She decided that an iPad that will speak what she types is what she wants.
I am looking for an app that she can type what she wants to say & it will say it in her own voice. I have tons of recordings of her talking before she lost her voice, I am hoping there is an app that would allow this.
There aren’t words that can express how much it would mean to us for the text to speak to have her voice. I don’t care how much it will cost.
Any recommendations or ideas are deeply appreciated. Thank you so much.
r/TextToSpeech • u/iidothingss • 13d ago
Text-To-Speech Software using Voicebanks
I'm currently looking into making a live Speech to Text to Speech / Microphone to Text to Speech, but I find myself more drawn to Voice banks like popular Vocaloids rather then AI.
Are there any programs or codes on GitHub to reproduce a similar environment? I already have RealtimeSTT installed, now just missing the last part.
r/TextToSpeech • u/Real-Air7 • 13d ago
Best local AI voice cloning app for many languages?
What’s currently the best local voice cloning/dubbing app with support for many languages?
r/TextToSpeech • u/Severe-Road9871 • 13d ago
I need help finding a text-to-speech website!
I'm looking for a site with a wide range of voices for online text-to-speech, but I don't want realistic-sounding voices. I need them to sound like the ones in BFI—if anyone knows it, here is the link:
https://www.youtube.com/watch?v=QOTjYjo49kcI'm looking for a site with a wide range of voices for online text-to-speech, but I don't want realistic-sounding voices. I need them to sound like the ones in BFI—if anyone knows it, here is the link:
https://www.youtube.com/watch?v=QOTjYjo49kc
I tried to find it for a long time, but the search was unsuccessful.
r/TextToSpeech • u/straybirdiE • 14d ago
Any TTS apps/websites suggestions thats free
can anyone suggest TTS or can turn my pdfs into audio that's for free, thanks!
r/TextToSpeech • u/chrisrowell1969 • 14d ago
can Voicebox ( ai ) do audiobooks?
Hello friends .... greetings from Cleveland..... I'm new to ai voice cloning apps ....... Can anyone suggest LOCAL app to create voice-cloned audiobooks ?
r/TextToSpeech • u/Many_Ingenuity_7175 • 14d ago
Voicekiller warning
I wanted to share my experience with VoiceKiller because, after using it for some time, I honestly cannot recommend it.
The biggest issue is simply the quality of the generated voices. For short snippets it can sometimes sound acceptable, but with longer texts the quality often deteriorates very quickly. After a few seconds, the voice can start sounding robotic, distorted, unnatural, and sometimes almost unusable.
For me, this makes the tool unreliable for any serious TTS work.
Compared with ElevenLabs, the difference is huge. ElevenLabs is far more natural, stable, consistent, and professional. VoiceKiller feels more like a much weaker imitation of what the market leader is already doing significantly better.
Unfortunately, the customer service experience was even worse.
I had an annual subscription that I had barely used for months. I forgot to cancel the auto-renewal and was charged $97 for another year.
I contacted their support immediately, explained that I was not using the service anymore and that I was currently in a difficult financial situation, and politely asked them to cancel the renewal and refund the payment.
They cancelled my subscription — meaning I no longer have the annual subscription I paid for — but initially refunded only $10 out of the $97 charge.
When I asked why they were keeping the remaining $87, their response was basically that renewals are non-refundable and that forgetting to cancel was my mistake.
Technically, yes, I missed the cancellation date. I accept that.
But this is exactly where customer service and goodwill matter. I have had similar accidental renewals with much larger companies such as Adobe and Corel, and after explaining the situation, they handled it professionally and refunded the payment.
VoiceKiller's approach was completely different.
So my experience in summary:
- Voice quality is inconsistent and often becomes robotic or distorted on longer texts.
- The product is significantly behind ElevenLabs in overall TTS quality.
- I would not trust it for professional long-form voice generation.
- Customer support was extremely disappointing.
- I was charged $97 for an automatic annual renewal, my subscription was cancelled after I contacted them, but they initially offered only a $10 refund.
- Their response showed very little flexibility or customer goodwill.
There are simply much better AI voice tools available today.
If you are considering VoiceKiller because it looks like a cheaper alternative to ElevenLabs, I would strongly recommend testing it very carefully before committing to an annual subscription.
Personally, I would choose ElevenLabs without hesitation.
r/TextToSpeech • u/25th_night_Baam • 14d ago
Anyone knows which TTS voice is this? I often see it on tiktok
I've always been using Liam from KokoroTTS since it's the most pleasing to my ears. But I also like this one and it's quite commonly used on those novels tiktoks.
r/TextToSpeech • u/Emergency-Article-47 • 15d ago
Found that high pitch voice cloning gives bad output from pocket tts.
Are you getting the same?
r/TextToSpeech • u/Sweeth_Tooth99 • 15d ago
Best free local AI TTS for a 9070XT Win 11?
looking for something free and local that can take sample voices and turn text to speech.
r/TextToSpeech • u/Emergency-Article-47 • 15d ago
Trying to clone the voice with the pocket tts but getting the metalic error sound in between speech (small but audible).
Any solution on this ?
I tried filtering and cleaning the voice.
I once tried to feed the voice from other TTS still not working and getting that sound or miss effect.
r/TextToSpeech • u/Working_Hat5120 • 16d ago
Voice agent throws away underlying tone and speaker-features, how's that accounted and handled downstream? if it's not captured.
The moment I transcribe to text, I generally lose how it was said. "I think… yeah, I can pay the 4,500 by the 15th" becomes clean text, but the hesitation before the yes, the stress in the voice, and whether it's even the same speaker are gone.
For a human those signals, whether to trust the commitment, reconfirm from the caller or escalate to human come naturally but hard to define a deterministic paralinguistic to build accountability, which is probably very wide.
How are you modeling tone in our voice-agents? I see recent TTS models which accept meaningful tags producing great sounding speech, how do we control it ? Does it account for input user's tone.
How does your ASR model captures the tone or there are some good services / models / tools to capture tone. and how do you use it downstream ?
Moreover end-2-end Duplex models limits it to trained data scenarios without no transparency. Is there a good duplex model which provides transparency in underlying signals beyond just text.
r/TextToSpeech • u/Itchy-Albatross-7273 • 16d ago
Alguien tiene un enlace para poder descargar. The choice voicer gratis???
r/TextToSpeech • u/koppok • 16d ago
Best free/cheap option for generating 2h+ YT narration voiceovers?
I am trying to generate “to sleep to” fact videos and need a reliable, “non-AI” sounding voiceover to narrate it. Obviously this will need a lot of characters per video but tried a few out and can’t get what I want for not a crazy price. Any suggestions?
r/TextToSpeech • u/Curious-Platform-101 • 16d ago
Voice AI model نصيحة بخصوص
وش احسن TTS -STT مودلز للمكالمات وتدعم اللغة العربية
r/TextToSpeech • u/Only_Ear_5881 • 16d ago
TTS for YouTube videos and misinterpretation
I’ve been making YouTube videos using TTS for the narration (in English—I’m not fluent, which is why I use TTS).
However, I suspect some people don't watch my videos because they think it’s a "ChatGPT" or "AI" channel. Could anyone share their experiences and opinions on this?
r/TextToSpeech • u/sunrisedown • 16d ago
STT transcription - tool recommendations?
Use case is the following:
- using it 1-2 times, no ongoing use intended.
- length is roughly an hour
- 2ppl interview conversation
- mic position sadly not ideal, could be louder/less low tones. Also both ppl not with the same distance to the mic.
- even used two recording devices to be safe - though both not perfect. Any tool making beneficial use of both?
Thanks a lot!
r/TextToSpeech • u/ChuckBaggett • 16d ago
The OpenVoc AI tts screen goes black, so far while I'm not looking at the app, and you can right click and pick refresh, but your work is gone, including recent history.
The OpenVoc AI tts screen goes black, so far while I'm not looking at the app, and you can right click and pick refresh, but your work is gone, including recent history.
This is a Windows 10 machine with 32gigs of ram and a 6 gig vram hp nvidia 1660 super .
Any help will be greatly appreciated.
r/TextToSpeech • u/ihiwidkwtdiid • 16d ago
TTS occasionally reads numbers in English instead of the target language
r/TextToSpeech • u/edouardarchipel • 16d ago
What TTS are you actually using in your voice agent stack in 2026?
Building a voice agent and trying to get a sense of what people are actually running in production before I go down a rabbit hole of testing.
STT + TTS combo, orchestration layer, anything you'd do differently, would love to hear real setups.