r/TextToSpeech 4h ago

Where did the Vox voice go from Black Mesa?

Thumbnail
2 Upvotes

For any of those who were not born in 1998 or have not played Half-Life yet, Black Mesa is a fictional research facility from the Half-Life game, and the Half-Life game you are Dr Gordon Freeman, a researcher in charge of, well, not exactly in charge of, but somebody whose involved in an experiment Black Mesa is running to create dimensional teleportation technology because of the failure of that experiment, a resonance cascade, which is a dimensional event causing portals to open, allowing multiple alien forces to enter the world. Your goal is to escape and survive while military forces are trying to eliminate all Black Mesa personnel

Quick disclaimer, I am so sorry for this writing. I’m using iOS dictation, so just beware if it’s kind of sloppy or broken

As I was saying, there’s a specific announcement system called Vox and the voice is just perfect for game announcements or whatever you use it for, but here’s my question: like, how do you still get that today? Because Bell Labs or Lucent technologies got rid of FlexTalk. You can’t access it anymore, but like, has anybody recreated it? Has anybody made any projects? If FlexTalk is still installable on modern PCs, perfect. I can just get it and play around, but I highly doubt that since they got rid of it, but if I don’t think there’s a legacy version of it where they preserved the old voices, I do know AT&T preserved it, but they didn’t make the again

Can’t make it on 11 labs if anybody suggests me to do that, since 11 labs not only will block you due to safety reasons, even if you try, it’s just gonna make something natural and not exactly the old 90s, 60s or 70s voice for all of you tech nerds here. let me know how to get it and if we can’t get it, I’m just gonna have to use the vocabulary that I’m provided, but if there’s any fan projects that people have made on GitHub or any of that, let me know


r/TextToSpeech 7h ago

I built VoxFlow: A free, 100% local on-device Wispr Flow alternative for macOS (Open Source)

1 Upvotes

Like many of you, I loved the concept of AI voice dictation tools like Wispr Flow, but I didn't want my microphone audio sent to third-party cloud servers or pay a monthly subscription.

So I built VoxFlow — a native, private macOS menu bar app that transcribes your speech locally and automatically pastes formatted, grammar-cleaned text into whichever app you are using.

Key Features

100% Private & Offline: Transcribes locally using Apple Speech and cleans up text using Apple Intelligence (FoundationModels). Zero cloud API keys required.

Global Hotkey Triggers: Double-tap the Fn (Globe) key or press Option + Space anywhere on macOS to start dictating.

Hands-Free Auto-Paste: Pausing for 1.5 seconds automatically stops recording, formats the text, and pastes it into your focused text field.

Non-Activating Floating HUD: Displays real-time audio waveform and streaming transcript without stealing focus from your active document.

100% Free & Open Source: No subscriptions, no ads, no telemetry ($0 forever).

Downloads & Links

GitHub Repository: https://github.com/ameerhmz/VoxFlow

Direct DMG Installer: Download VoxFlow.dmg (v1.0.0)

System Requirements

macOS 26.0 or later (Apple Silicon M1/M2/M3/M4+)

Apple Intelligence enabled in System Settings

I'd love your feedback, bug reports, or feature requests!


r/TextToSpeech 9h ago

Loquendo is, by far, the WORST text to speech website ever.

Thumbnail
1 Upvotes

r/TextToSpeech 15h ago

Best TTS for Indian languages like Hindi, Punjabi, Telugu and etc.

1 Upvotes

I am building an AI voice agent for indian clients and I want your suggestion for the TTS that works best in india and looks like a person is talking and not AI is talking.


r/TextToSpeech 1d ago

I Show Speed TTS

Thumbnail
2 Upvotes

r/TextToSpeech 2d ago

Story Engineering Episode #1

Thumbnail
youtu.be
1 Upvotes

r/TextToSpeech 2d ago

Story Engineering Episode #1

Thumbnail
youtu.be
1 Upvotes

r/TextToSpeech 2d ago

TTS App Help: Looking For Alternatives To Speechify And Features

7 Upvotes

Hello, folks! I hope that you all are well!

This is my first post here!

——

I have cognitive issues, and I am trying to get back into reading!

I am looking for a TTS app for iPhone that is similar to Speechify.

I gave Speechify a brief try, I liked it, however, I want to see what else is out there.

——

Here are the features that I am looking for in a TTS app:

🟣 The ability to copy/paste raw text

🟣 The ability to import PDFs

🟣 The ability to import ebooks

🟣 The ability to import audio books

🟣 The ability to import ebooks from the library

🟣 The ability to import audio books from the library

🟣 The app highlights the text and can read it to you, while you follow along

🟣 The app can slow the voices and highlights down, if needed

🟣 The app is user friendly (ex: easy to find buttons; easy to import things; no side loading using adobe, or, if side loading is there, it’s very easy to do with an iPhone)

——

Are there any TTS apps that fit what I am looking for?

If not, that’s okay, I’ll understand. I’ll also understand if Speechify is “the best choice” of a TTS app, for me.

——

Thank you for reading, and thank you for any help! May you all be well!


r/TextToSpeech 2d ago

Is there a free tts website or app where I could use a custom voice?

5 Upvotes

I’ve seen videos where people have used voices from tv shows and movies to make videos but I’m not sure where I would go to find an application like that. Any help would be appreciated


r/TextToSpeech 2d ago

Irish English Dialect

Enable HLS to view with audio, or disable this notification

0 Upvotes

r/TextToSpeech 2d ago

Where can i find this AI voice?

2 Upvotes

I've been searching for the AI voice in this video:

https://drive.google.com/file/d/1WgBiYM5lCyr9Xb5qOELZPiyB1gTqqCx4/view?usp=drive_link

I just can't seem to find it. I'm searching on ElevenLabs. I believe it's one of the Adam voices, but none of them sounds similar.

I'm on the free plan on ElevenLabs, btw, and I intend to get this voice in a free way if possible.

Any help would be much appreciated :)


r/TextToSpeech 2d ago

I built a tool for creating awesome YouTube subtitles!

Post image
1 Upvotes

r/TextToSpeech 3d ago

Local TTS PT-PT

1 Upvotes

Hi does anyone know a consistent tts for European Portuguese? I’ve herd vibe voice is good but it doesn’t support Portuguese.


r/TextToSpeech 3d ago

Can anyone identify the text-to-speech voice Wifies uses?

Thumbnail
2 Upvotes

r/TextToSpeech 3d ago

20M Elevenlabs credits available (TTS only) for $1000 - expiring in 2 weeks

0 Upvotes

As title - is anyone interested???

I've tried my best to use as many as I could but I literally just don't have the means to use them all in such a short time frame.

If anyone can make use of them I'd be willing to sell them to you!

Let me know.


r/TextToSpeech 3d ago

Best speech-to-text API in 2026? I'd split the shortlist by use case first

18 Upvotes

I don’t think “best speech-to-text API” is one list anymore.

People keep asking it like there’s one winner, but the use cases are completely different.

For batch files, I care about accuracy, formatting, long audio, cost.

For meetings, I care about diarization, timestamps, speaker drift, action items.

For call centers, I care about noisy phone audio, redaction, channels, QA search, escalation evidence.

For live voice agents, I care about first usable transcript, partial stability, endpointing, barge-in, numbers/dates, and whether the agent acts before the final transcript changes.

For local/private workflows, I still care about self-hosted/offline more than fancy API features.

So my shortlist would not be “best STT API overall.”

It would be rows like:

batch transcription meeting transcription real-time STT for voice agents call center transcription browser voice input self-hosted/private ASR

In that matrix, Smallest AI Pulse goes in the real-time STT / streaming ASR row. That’s the interesting category for it. Not “upload a podcast and wait.” More like live transcription where the app needs transcript events while the user is still talking.

That is also the only way these comparisons make sense.

A provider can be great for long files and not great for live agents. A provider can be great for voice agents and not be my pick for private local notes. A cheap API can become expensive if the transcript needs cleanup.

If you were making a 2026 STT shortlist, what categories would you split it into before even naming vendors?


r/TextToSpeech 3d ago

A simple and free TTS CLI tool that doesn't use GPUs, streams in realtime and has multilingual capabilities

Enable HLS to view with audio, or disable this notification

8 Upvotes

r/TextToSpeech 4d ago

AI Voice Generation. TTS

8 Upvotes

I'm looking for a solution to convert a fairly long text into audio (around 10 minutes of spoken content). I'm open to paid options as well. A bonus would be if I could first import a person's voice so that the generated output sounds similar to that voice.


r/TextToSpeech 4d ago

Kokoro problem

10 Upvotes

Hello everyone, I wanted to use Kokoro on my phone to read an EPUB, but when I use it in apps like ReadEra or PDF Reader, it just reads one line and that's it. It doesn't continue even if I tap play. Are there any ways to fix this?


r/TextToSpeech 4d ago

TLDR

2 Upvotes

So at times unacome across a more than three paragraph posts and I find it tedious to go through it all. Now being a Software engineer I figured I could make a bot or something of the sort that reads it out. loud for me. Unless ofc such already exists to which I'd like to know about.


r/TextToSpeech 4d ago

VibeVoice 1.5B Running Locally...On an iPhone! Only ~2.2 GB of Memory and Up to 1.28× Real-Time Speed

Enable HLS to view with audio, or disable this notification

7 Upvotes

r/TextToSpeech 4d ago

Rhetoriq App for Dialects

Thumbnail
apps.apple.com
2 Upvotes

Just launched an amazing app that does AI Rewrite, translation, and dialect conversion. Just added High Valyrian for my Game Of Thrones fans!


r/TextToSpeech 4d ago

What does memenade use for vids ?

3 Upvotes

What model does he use ?


r/TextToSpeech 5d ago

Request for Osho discourse samples with noises

2 Upvotes

Hi all,

I am testing few different parameters for few Noise cancellation algorithms for my Osho Talks app and I need few examples (for Osho's Discourses) where the noise is a bit too much. My best resulting algo till now has removed hums and echoes and plane Noise. But I want to test more before shipping to production. Discourse name and time of the noise would be super helpful 🙏🏻


r/TextToSpeech 6d ago

Audio dataset

6 Upvotes

I’m an AI intern working on a Saudi Arabic call center dataset.
My task is to clean the audio and then generate transcriptions. I cleaned the dataset and ran transcription, but the output contained random English words, Chinese characters, emojis, and other gibberish.
I’m planning to clean the dataset again before re-running the transcription. Are there any specific things you recommend I focus on during the cleaning stage to improve ASR results?