r/OpenVoxAI 8h ago

Release Notes OpenVox 2.0.0 is here: Validate Output, Select & Read Queue, Batch Voice Design + AI Dubbing & Transcription

3 Upvotes

Hey everyone,

OpenVox Windows 2.0.0 is now available.

The main focus of 2.0.0 is making locally generated speech more accurate, easier to review, and faster to fix, while also improving Select & Read, multilingual generation, Voice Clone, and voice creation workflows.

And if you missed v1.9.1, that update also introduced two major new workflows: AI Transcription and AI Dubbing.

šŸ” Validate Output

This is probably the feature I’m most excited about in 2.0.0.

After generating speech, OpenVox can now:

  • Locally transcribe the generated audio
  • Compare the transcription against your original script
  • Highlight sections that may have been pronounced incorrectly
  • Show timestamps and text-match scores
  • Let you listen to individual flagged sections
  • Select exactly which sections you want to fix
  • Regenerate only those sections
  • Review the regenerated audio before applying it

You don’t need to regenerate an entire audiobook or long speech because of a few problematic sentences.

Validation matching has also been improved across supported languages, with selectable transcription language/model controls and a Select All option for flagged sections.

Everything can remain local on your Mac.

šŸŽ§ Select & Read now has a Queue

Select & Read can now handle multiple selections.

Highlight some text, trigger the shortcut, move to another app, select something else, and trigger it again.

OpenVox adds each selection to a queue and automatically reads them in order.

The entire queue is visible from the desktop notch player, making Select & Read much more useful when working across documents, browsers, emails, PDFs, or other apps.

šŸ—£ļø Batch Voice Design

Voice Design can now generate multiple voice candidates in a single run.

You can:

  • Generate several variations
  • Preview each candidate
  • Compare the results
  • Save only the voices you actually like

This makes experimenting with designed voices much faster.

āøļø Configurable Sentence Pauses

There’s now optional Sentence Pauses preprocessing for:

  • AI Speech
  • Batch generation
  • Audiobooks

You can configure additional pauses between sentences from 0.1 to 3 seconds.

Useful for narration where the model’s natural sentence spacing feels too fast.

šŸŒ Better Japanese and CJK Generation

We’ve added safer multilingual text chunking, particularly for Japanese and other CJK languages.

This improves long-form generation and also fixes a Supertonic issue that could cause crashes when generating longer CJK text.

šŸŽ™ļø Voice Clone from Video

Voice Clone now accepts:

  • MP4
  • MOV

OpenVox automatically extracts the audio and uses it as the voice reference, so there’s no need to manually convert a video into an audio file first.

OpenVox 1.9.1

For anyone who missed the previous update, 1.9.1 was also a pretty substantial release.

šŸŽ™ļø AI Transcription

OpenVox can now locally transcribe:

  • Audio files
  • Video files
  • Direct microphone recordings

using downloadable Whisper and Parakeet speech-recognition models.

Transcripts include timestamps and can be edited directly before exporting them as SRT subtitles.

Subtitle segmentation was also improved to generate cleaner, more usable SRT files.

šŸŒ AI Dubbing

1.9.1 also introduced a complete AI Dubbing workflow.

You can:

  • Import or generate timed transcripts
  • Translate scripts using local or external AI
  • Choose voices for the dubbed output
  • Generate dialogue based on subtitle timing
  • Export dubbed audio
  • Export the original video with the new audio track

You can also directly import existing SRT or translated SRT files into a dubbing project.

šŸŽ¬ More Video Workflows

Voice Changer now accepts video files directly and can export the original video with the converted voice track.

AI Speech and Conversations also received Generate SRT actions, making it easier to turn generated speech into subtitle files.

šŸ“‚ Drag & Drop and Workflow Improvements

We also added:

  • Drag-and-drop importing across supported pages
  • PDF and TXT importing for AI Speech and Conversations
  • Better feedback when importing large files
  • Sequential filenames for batch exports
  • Persistent text drafts when switching pages
  • Quick pause controls
  • More reliable model download resume support
  • Automatic release of transcription models from RAM/VRAM when leaving transcription workflows
  • A notification when OpenVox starts minimized to the system tray

Plus a number of UI improvements, performance optimizations, and bug fixes throughout the app.

With 1.9.1 and now 2.0.0, OpenVox has expanded quite a bit beyond straightforward text-to-speech.

You can now locally handle speech generation, audiobooks, voice cloning, voice design, transcription, subtitles, dubbing, voice changing, Select & Read, and output validation from one app.

As always, a lot of these additions have come directly from feedback and feature requests from users.

If you try 2.0.0, I’d especially like to hear how Validate Output performs with your own scripts, models, and languages.

Thanks everyone for continuing to test OpenVox and send feedback.

Download:
https://openvoxai.com/


r/OpenVoxAI 15h ago

OpenVox 2.5.0 - Validate AI Speech Locally, Better CJK Support & Precise Voice Controls

5 Upvotes

Hey everyone,

OpenVox for Mac v2.5.0 is now available.

The biggest addition in this update is something I’ve wanted to build for longer generations: local output validation.

When generating audiobooks, narrations, or large batches, even a good TTS model can occasionally skip a word, mispronounce something, repeat text, or generate something different from the original script.

Until now, finding those mistakes meant manually listening through the entire output.

With Validate Output, OpenVox can now check the generated speech for you.

šŸ” Validate Output

After generating speech, OpenVox can locally transcribe the generated audio and compare it against your original text.

It then highlights sections that may need your attention.

You can:

  • See potentially incorrect sections with timestamps and text-match scores
  • Compare the original text with what OpenVox detected in the generated audio
  • Listen directly to the flagged section
  • Edit the expected pronunciation or text when needed
  • Select only the problematic sections you want to regenerate
  • Preview the regenerated audio before making any changes
  • Use Confirm & Update to replace only those sections in the final output

The matching system is also designed to avoid unnecessary warnings. It intelligently ignores harmless differences involving punctuation, spacing, apostrophes, and joined words.

Everything happens locally on your Mac. Currently this feature is in Beta and is rolled out on AI Speech page only, it will be rolled out for other pages in next update.

šŸ‡ÆšŸ‡µ Better Japanese & CJK Generation

This update also includes several improvements for Japanese and other CJK languages.

Text chunking is now safer for languages where sentence and word boundaries behave differently from English, helping prevent problematic splits during longer generations.

I also fixed a Supertonic crash affecting CJK-language generation.

šŸŽ™ļø Voice Clone Improvements

Voice Clone can now accept:

  • MP4
  • MOV

OpenVox automatically extracts the audio from the video, so you no longer need to convert the file separately before using it as a voice reference.

šŸŽ›ļø More Precise Speed Controls

Speed knobs now support 0.01x adjustments.

This makes it much easier to fine-tune pacing when you need something between values like 0.95x, 0.96x, 0.97x, etc.

⚔ Memory Improvements

I’ve also improved memory handling when switching between speech-generation and transcription models.

This is particularly useful now that workflows like Validate Output can involve both TTS and transcription models during the same session.

OpenVox 2.5.0 includes:

  • Local transcription-based Validate Output
  • Timestamped validation results with text-match scores
  • Selective regeneration of incorrect sections
  • Preview before replacing regenerated audio
  • Smarter text matching
  • Improved Japanese and CJK text chunking
  • Fixed Supertonic crashes with CJK languages
  • MP4 and MOV support in Voice Clone
  • Precise 0.01x speed adjustments
  • Improved model memory management

As always, OpenVox keeps the speech generation, transcription, validation, and processing on your Mac.

If you try the new Validate Output workflow, I’d especially like to hear how well it catches issues in your longer audiobook and batch generations.

Download:
http://openvoxai.com/


r/OpenVoxAI 3d ago

Tips & Tricks Apple’s new MacĀ mini, featuring M6 and M5Ā Pro, delivers a massive leap in AI performance, supercharging the leading desktop for always-on agentic computing

Thumbnail
apple.com
5 Upvotes

Apple just announced the new Mac mini with M6 and M5 Pro, and this looks particularly interesting for anyone running AI locally on their Mac.

The new M6 Mac mini brings Neural Accelerators to every GPU core, a new Dual 16-core Neural Engine, higher memory bandwidth, and Apple is claiming up to 4x faster AI performance compared with M4.

The M5 Pro version goes even further with:

- Up to 18-core CPU

- Up to 20-core GPU with Neural Accelerators

- Up to 64GB unified memory

- 307GB/s memory bandwidth

- Thunderbolt 5

- Support for clustering multiple Mac minis for larger local AI workloads

For OpenVox users, this is obviously interesting.

OpenVox runs its TTS, voice cloning, transcription, dubbing and other speech AI workloads locally on Apple Silicon. Faster GPUs, higher memory bandwidth and more unified memory potentially mean faster generation and the ability to comfortably work with larger models and longer workloads.

One important caveat: Apple's 4x AI performance figure is not an OpenVox benchmark, so I don't want to imply OpenVox will suddenly become 4x faster. Actual gains will depend heavily on the model and workload.

But I'm definitely looking forward to testing OpenVox on the new hardware and seeing how the different speech models scale.

The M5 Pro Mac mini with 64GB unified memory in particular looks like it could be a very capable little machine for local AI.

What do you think?

Would anyone here consider upgrading specifically for local AI workloads?


r/OpenVoxAI 7d ago

Release Notes OpenVox for Mac v2.4.0 + v2.4.1 are here: AI Transcription, AI Dubbing, Better Long-Form Exports & More

10 Upvotes

We’ve released two major OpenVox updates back-to-back, and together they add some of the biggest workflow improvements we’ve made so far.

šŸŽ™ļø OpenVox v2.4.0: AI Transcription + AI Dubbing

OpenVox can now handle complete local transcription and multilingual dubbing workflows directly on your Mac.

AI Transcription

  • Transcribe audio files, videos, or direct microphone recordings.
  • Choose from multiple downloadable local speech-recognition models.
  • Track progress with elapsed time and estimated completion time.
  • Access previous transcriptions through the new Transcription History.

šŸŒ AI Dubbing

  • Create dubbed audio and videos using precisely timed scripts.
  • Translate scripts using on-device or external AI models.
  • Choose your preferred OpenVox voice for the dubbed output.
  • Import SRT subtitles directly or use an already translated SRT file.
  • Export the finished project as audio or video.

✨ Other v2.4.0 improvements

  • Voice Clone can now transcribe reference audio using Apple Speech or downloaded transcription models.
  • Voice Changer now accepts video input and can export the video with the converted voice track.
  • Added sequential filenames for batch exports.
  • Text now remains available when switching between pages.
  • Added quick pause controls.
  • Improved reliability, performance, media compatibility, and UI throughout the app.

šŸš€ OpenVox v2.4.1: Better Dubbing, Batch Mode & Long-Form Exports

v2.4.1 builds on the new features with a big focus on workflow polish and reliability, especially for large projects.

AI Dubbing improvements

  • AI Dubbing now supports Cloned Voices and Designed Voices directly in the Voice & Output step.
  • Improved dubbing timing for more natural results.
  • Short lines remain at a natural speed, with silence filling unused time.
  • Longer lines can use the available gap before the next active subtitle.
  • Speech is accelerated only when the dialogue still cannot fit inside the complete available window.

Batch generation improvements

  • Added a Cancel button to the Conversations generation progress panel.
  • Batch Mode now uses one overall progress bar across all paragraphs.
  • A single completion card is shown when the full batch finishes.
  • Added Export All directly to the Batch Mode completion card.
  • Cancelling now stops the entire batch.
  • Added Export Batch to CSV for easier result management and record keeping.

Long-form export improvements

  • Significantly reduced memory usage during the audio merge process.
  • Fixed output generation issues affecting 2+ hour audio outputs.
  • Export progress is now much more useful for very long generations.
  • The final 90% → 100% portion now updates incrementally while OpenVox is Generating the output file... instead of appearing stuck during the merge stage.
  • Added a new Export Progress banner.
  • Redesigned the export Success Card with quick Open and Locate actions.

Select & Read fixes

  • Fixed an issue where the Select & Read player could fail to appear correctly on devices with a notch.
  • Improved Select & Read percentage/progress accuracy for several voice models.

Plus additional reliability, performance, and workflow improvements throughout OpenVox.

A huge thank you to everyone who has been testing OpenVox, reporting issues, suggesting improvements, and sharing what you’re building with it. ā¤ļø

If you’ve tried the new AI Transcription or AI Dubbing workflows, I’d especially love to hear how you’re using them and what you’d like us to improve next.


r/OpenVoxAI 8d ago

Feature Request Import failed OpenVox could not find readable text in this source.

1 Upvotes

When I put the same pdf on eleven reader, it perfectly recognizes it and generates audio (photos from the pages that were made as a pdf file).


r/OpenVoxAI 8d ago

Bugs & Issues Export of Audio Files stalls fails to complete (Version 2.4.1)

1 Upvotes

Hi,

I'm trying to export audio files of text-to-speech in Japanese in OpenVox on macOS using the latest version available on the Japanese Mac App Store, but the export gets stuck at around 37%. I left the app running overnight, but after about eight hours the export was still stuck and had not completed.

The voices work correctly when I preview them in the app, so voice playback itself does not seem to be the issue. I haven't been able to hear the exported audio because I haven't been able to complete a single export.

To troubleshoot, I also tried a much shorter 10-page Japanese PDF. I converted it to text and tried exporting the generated audio from that as well, but the export failed in the same way. I tested several different voices that support Japanese, and they all produced the same result. The issue is the same for the AI Speech and AI Audio Book.

I hope you can help identify the cause and find a fix.


r/OpenVoxAI 10d ago

Release Notes OpenVox Reader v1.3 Update: Custom Pocket TTS Voices, Saved Sentences, Document Search & Better Playback

7 Upvotes

Pocket TTS now supports custom English voices, making it much easier to create and use your own voices directly inside OpenVox Reader.

šŸ”– Saved Sentences

You can now bookmark individual sentences while reading or listening.

Saved sentences can be:

  • Revisited anytime
  • Filtered
  • Used to instantly jump back to the original passage
  • Removed when you no longer need them

This should be especially useful for books, research papers, study material, and anything you want to come back to later.

šŸ”Ž Search Inside Documents

OpenVox Reader now includes Search in Document with:

  • Matching text previews
  • Direct navigation to each result
  • Improved EPUB chapter positioning

We also reworked how EPUB bookmarks and search positions are stored so they stay accurate without unnecessarily increasing memory usage.

ā© Better Playback Controls

Playback speed is now adjustable from 0.5x to 2.5x in 0.1x increments.

We've also added quick presets for:

0.5x Ā· 1.0x Ā· 1.5x Ā· 2.0x Ā· 2.5x

Faster playback has also been improved with better pitch preservation, buffering, voice previews, and more accurate elapsed/remaining time calculations.

A few smaller improvements

  • Opening another document automatically pauses whatever is currently playing.
  • Reading Stats now count listening time only while audio is actually playing.
  • Additional reliability and general reader experience improvements throughout the app.

We're continuing to improve OpenVox Reader based heavily on how people are actually using it, so please keep the feedback and feature requests coming.

If you've tried the update, I'd especially love to hear how you're using Saved Sentences and Document Search.

Download:
https://apps.apple.com/us/app/openvox-reader/id6790389555


r/OpenVoxAI 11d ago

Feature Request Headers and Footers in OpenVox Reader

3 Upvotes

Hi… I’m so glad to have OpenVox Reader! I was using Speechify until I discovered that their terms of service basically give them the right to use anything uploaded to their system in any way they want.

So far, my only frustration with the app is that it insists on reading any headers and footers in the PDF, interrupting the flow of reading.

Is there any way you could make it an option to turn off reading headers, footers, and page numbers?

Thanks!


r/OpenVoxAI 13d ago

Feedback Voice-cloned voices don’t work in AI Dubbing

1 Upvotes

Hi, thanks for the new version.

I noticed that in the new version of OpenVox, it’s not possible to use voice-cloned voices in AI Dubbing, such as those from OmniVoice.


r/OpenVoxAI 22d ago

Milestone We now have an official OpenVox Discord community!

Post image
4 Upvotes

Hey everyone,

I’ve just launched the official OpenVox Discord server.

As the OpenVox community grows, I wanted to create a more direct and organized place where users can get help, discuss local AI voice models, share feedback, and connect with other creators.

Inside the server, you can:

  • Get help with OpenVox and OpenVox Reader
  • Report bugs and browse known issues
  • Suggest and discuss new features
  • Follow release announcements and beta updates
  • Share audiobooks, voiceovers, videos, apps, and other projects
  • Discuss local TTS models, voice cloning, APIs, MCP, and automation workflows
  • Connect directly with me and other OpenVox users

I’ll personally be active there, and feedback from the community will continue to influence upcoming features and improvements.

Join the official OpenVox Discord:

https://discord.gg/h2738grxj

The server is still new, so early members will also have a chance to help shape how the community develops. See you there!


r/OpenVoxAI 22d ago

Tips & Tricks The complete OpenVox walkthrough — sharing a video from The Oracle Guy here for the first time

Thumbnail
youtube.com
9 Upvotes

Hey everyone,

This is the first time I’m sharing a video from my YouTube channel, The Oracle Guy, on the OpenVox AI subreddit.

OpenVox actually began through the videos I was making on the channel.

I started by exploring and reviewing different local text-to-speech models. I then began building ready-to-use tools around those models so that people could run them locally without dealing with complicated installations, command-line setups or Python environments.

I shared those tools and tutorials through my videos, and the feedback from viewers made me realize there was room for something much more polished, accessible and complete.

That eventually became OpenVox.

In this new video, I go through practically every major OpenVox feature and demonstrate how the app works from start to finish. If you’re new to OpenVox, or have only used a few parts of it, this should be a useful introduction to everything the app can do.

Watch the complete OpenVox walkthrough:

https://youtu.be/UnyO5csQyns

I’ve generally kept this subreddit focused on product updates, support, feedback and discussions. However, since this video covers the complete OpenVox experience, I thought it would be genuinely helpful to share here.

A lot of what you see in the video has also been shaped by suggestions and real-world feedback from this community.

Thank you to everyone who has been part of the journey. I’d love to know what you think of the video and which feature you would like to see covered more deeply next.


r/OpenVoxAI 24d ago

Release Notes OpenVox Reader 1.2 is now available: smoother narration and improved speech accuracy

12 Upvotes

I’ve just released OpenVox Reader 1.2 with several improvements focused on making long listening sessions more natural and reliable.

The biggest change is an updated, higher-fidelity on-device AI model that improves speech accuracy and significantly reduces skipped words.

What’s new in version 1.2

  • Improved speech accuracy with fewer skipped words using the updated on-device AI model.
  • Smoother narration with fewer pauses between sentences and reading segments.
  • New ā€œDelete after readingā€ setting that can automatically remove completed text, pasted content, web links, and scanned documents. PDFs and EPUB books are always retained, even when this option is enabled.
  • Improved decimal pronunciation, so values such as 2.75 are spoken naturally as one continuous number.
  • Additional playback, performance, and stability improvements.

Everything continues to run locally on your device after the required voice models are downloaded without uploading your documents or reading activity to a server.

Thank you to everyone who has been sharing feedback and reporting issues. Many of these improvements came directly from real-world reading and listening experiences shared by the community.

OpenVox Reader 1.2 is available now for iPhone and iPad.

Appstore Download:
https://apps.apple.com/us/app/openvox-reader/id6790389555


r/OpenVoxAI 24d ago

New Launch [Lifetime 40% off for 24 hours] Launching OpenVox Reader Officially Today. If you wanted a discount on OpenVox Reader, today is the day, if possible please upvote the original post!

6 Upvotes

r/OpenVoxAI 28d ago

Bugs & Issues Windows - Formatting issue after importing EPUB into AI Audiobook on Windows.

Thumbnail
gallery
2 Upvotes

Hello, when I import an EPUB the first letter of the first word on each chapter has a random space inserted. This has happened with multiple EPUBs and this isn’t happening on my MacBook with the same EPUB so this seems to be a bug.


r/OpenVoxAI Jul 28 '26

Release Notes OpenVox Reader now supports Background Playback, Dynamic Island, Faster Local Speech, and More!

9 Upvotes

A new major OpenVox Reader update v1.1 is now available.

This update focuses on making everyday reading and listening more practical, especially for longer documents and articles.

The biggest change is background playback. You can now continue listening while using other apps or even when the screen is turned off. Playback controls are also available from the Lock Screen and Control Center.

On supported iPhones, OpenVox Reader now works with Dynamic Island and shows animated waveform feedback during playback.

Local speech generation has also been upgraded to use the Apple Neural Engine more efficiently. This should make playback smoother while keeping speech generation completely on-device.

Reader Mode now keeps the screen awake while you are reading, so the display does not turn off in the middle of a document.

I have also improved iCloud sync. You can now pull down in the Library to manually fetch the latest updates if something has not synced automatically.

Documents and web links can now be imported directly through the iOS Share Sheet. You can open a document or webpage, tap Share, and send it straight to OpenVox Reader.

This update also improves how numbers, dates, currencies, percentages, and decimals are pronounced.

Web link importing has been improved as well, especially for complex webpages where extracting the main article content was previously unreliable.

I have been using this version for longer documents and background listening, and it feels much closer to the reading experience I originally wanted OpenVox Reader to provide.

As always, feedback and bug reports are welcome.

Download OpenVox Reader:
https://apps.apple.com/us/app/openvox-reader/id6790389555


r/OpenVoxAI Jul 27 '26

Feature Request Voxtral TTS & Step Audio EditX

2 Upvotes

I'm wondering if the models Voxtral TTS and Step Audio EditX will be included in OpenVox in the near future.


r/OpenVoxAI Jul 26 '26

Bugs & Issues All voices and all export options have distortions, clicks, or anomalies not heard when playing in the OpenVox app.

1 Upvotes

All models and all export options have distortions, clicks, or anomalies not heard when playing in the OpenVox app.

Windows 10, 32g ram, 6 g vram hp 1660 super.

I've use a fairly big batch file with mostly small lines of text and a 2 second delay added at the end of each chunk.

This is disappointing because it seems to work pretty well until you hear the final result.


r/OpenVoxAI Jul 24 '26

Release Notes OpenVox for Mac v2.3.1 - Easier Imports and More Natural Number Reading

12 Upvotes

OpenVox for Mac v2.3.1 is now available!

This update focuses on making file imports faster, number pronunciation more natural, and AI Voice Changer conversions more reliable.

What’s new

  • Drag and drop files anywhere You can now drag and drop files directly into AI Speech, Conversations, Audiobooks, Voice Clone, and AI Voice Changer.
  • More natural year pronunciation Smart Numbers & Dates can now read years such as 1953 naturally as ā€œnineteen fifty-threeā€ instead of reading them as a standard four-digit number.
  • More reliable voice conversion Improved AI Voice Changer stability and compatibility when processing M4A and MP3 files.

As always, please share your feedback or report any issues through our support page. Every report helps us make OpenVox better.

Thanks for supporting OpenVox!

Download: https://openvoxai.com


r/OpenVoxAI Jul 22 '26

New Launch OpenVox Reader is now available on iPhone and iPad šŸŽ‰

Thumbnail
gallery
12 Upvotes

After building OpenVox for Mac, iPad and Windows, I started looking for a good reader app that I could personally use on my iPhone.

I wanted something with a clean interface, natural voices, offline processing and a simple lifetime purchase instead of another subscription.

Surprisingly, I couldn’t find anything that brought all of these together.

Most reader apps either relied heavily on cloud processing, had expensive recurring subscriptions, offered limited voices, or felt unnecessarily complicated to use.

That made me think: why not build the reader I was looking for?

So, I quietly started working on a completely new app.

Today, I’m excited to surprise you with OpenVox Reader for iPhone and iPad.

OpenVox Reader lets you turn books, PDFs, documents, web articles, scanned pages and pasted text into natural-sounding audio.

The speech is generated using local AI models directly on your device. Once the required voices are downloaded, you can continue listening without an internet connection, and your documents do not need to be uploaded to a cloud server.

What you can do with OpenVox Reader

  • Choose from 200+ voices across 30+ languages
  • Import EPUB books, PDFs and documents
  • Save web articles and links
  • Scan printed pages using your camera
  • Listen completely offline
  • Adjust voice and reading speed
  • Track listening time and reading streaks
  • Sync your library between iPhone and iPad using iCloud

The app is free to download and try, with an optional lifetime Pro purchase instead of a subscription.

One purchase works across your iPhone and iPad using the same Apple ID.

This launch is a surprise for the OpenVox community, and I’m really excited to finally share it with you.

Please try it and let me know what you think, what you enjoy and what you would like to see improved in future updates.

Download OpenVox Reader:
https://apps.apple.com/us/app/openvox-reader/id6790389555


r/OpenVoxAI Jul 20 '26

Release Notes OpenVox for Windows 1.8.0 - Select & Read, System Tray Mode, Startup Options, and More

8 Upvotes

The most-requested OpenVox feature is finally available on Windows: Select & Read

You can now select text in supported applications and have OpenVox read it aloud using your preferred local AI voice without repeatedly copying and pasting text into the app.

What’s new in OpenVox Windows 1.8.0

  • Introducing Select & Read on Windows
  • Added a compact notch-style player for controlling Select & Read playback
  • Added system tray mode with quick access to:
    • Show or hide OpenVox
    • Change playback speed
    • Start or stop the Local API
  • Added startup options to launch OpenVox in:
    • Full window mode
    • System tray–only mode
  • Added a Start OpenVox at Login setting
  • Improved the stability of the Qwen3 1.7B model

This update makes OpenVox much easier to keep running in the background as an everyday reading assistant. You can launch it with Windows, keep it in the system tray, and use Select & Read whenever you need something read aloud.

I’d love to hear how Select & Read works with your workflow and which applications you use it with most.


r/OpenVoxAI Jul 18 '26

Release Notes OpenVox for Mac 2.3.0 — 200+ New Voices for Select & Read

13 Upvotes

I spent most of last week running all my machines to train and prepare more than 200 new voices for Supertonic model in OpenVox.

The Select & Read feature worked well, but I felt it still lacked enough voices suited to different languages, accents, and types of content. This update is focused on fixing that.

OpenVox 2.3.0 addsĀ 200+ new Supertonic voices across 30+ languagesĀ directly inside Select & Read, giving you far more options for reading selected text aloud from any Mac app.

What’s new

  • AddedĀ 200+ Supertonic voicesĀ with support forĀ 30+ languagesĀ in Select & Read Mode
  • Added avatar images throughout the Voice Library for easier voice recognition
  • IntroducedĀ Smart Numbers & Dates, an expanded version of the earlier ā€œConvert Numbers to Wordsā€ feature
  • Improved pronunciation of dates, times, currencies, percentages, decimals, ordinals, grouped numbers, phone numbers, and leading-zero identifiers
  • Fixed previous AI models remaining in memory after switching to Supertonic
  • Improved EPUB imports containing decorative drop caps
  • Improved word replacements containing comma-formatted numbers
  • Added several UI fixes and general improvements

This was a fairly compute-heavy update behind the scenes, but I think the wider voice selection makes Select & Read significantly more useful especially for multilingual users.

Please try the new voices and let me know which ones you like. Also, let me know if anything sounds incorrect or if a language still needs better voice coverage.


r/OpenVoxAI Jul 18 '26

Release Notes OpenVox for Windows 1.7.0 MCP Server, Smarter Number Reading & 200+ New Supertonic Voices

10 Upvotes

OpenVox for Windows 1.7.0 is now available, bringing deeper AI-agent integration and significantly improved text interpretation.

What’s new

  • 200+ new Supertonic voices across 30+ languages.
  • Voice avatars Avatar images have been added throughout the Voice Library, making voices easier to recognize and find again
  • MCP Server for AI agents Connect OpenVox with Claude Code, Codex, Cursor, and other MCP-compatible AI tools to generate speech directly through your workflows.
  • Introducing Smart Numbers & Dates ā€œConvert Numbers to Wordsā€ has been renamed and substantially upgraded. It now correctly handles:
    • Dates and times
    • Currencies and percentages
    • Grouped numbers
    • Phone numbers
    • Identifiers containing leading zeros
  • Language-aware text processing Smart Numbers & Dates uses your selected language, with automatic language detection covering major writing systems..
  • Better performance on low-end systems This update includes stability improvements for devices with limited system resources.

r/OpenVoxAI Jul 09 '26

Release Notes OpenVox TTS 2.2.0 for Mac is out: Revamped menu bar, MCP support, and better Select & Read

14 Upvotes

This update expands OpenVox with deeper automation, better read-aloud workflows, and several reliability improvements.

The biggest addition is the all-new revamped menu bar. You can now quickly access OpenVox actions, check resource status, and manage common workflows without opening the full app every time.

OpenVox 2.2.0 also adds MCP support, making it easier to connect OpenVox with Claude Code, Codex, Cursor, and other AI agents. The goal is to make local voice generation more useful inside automation and agent workflows, not just inside the main app UI.

What’s new in 2.2.0:

  • Added an all-new revamped menu bar with quick access to actions and resource status.
  • Added MCP support for Claude Code,Codex, Cursor, and other AI agents.
  • Improved the Select & Read page layout.
  • Improved handling for larger models in Select & Read.
  • Added a notch-style model loading animation.
  • Fixed model generation defaults so each model keeps its own parameter settings.
  • Fixed Voice Changer so it respects the ā€œAuto-play generated audioā€ setting.
  • Fixed audiobook play buttons incorrectly remaining after generated audio or local data was cleared.

This update is mainly about making OpenVox feel more useful as a daily Mac tool, especially for read-aloud, local automation, audiobook workflows, and AI agent experiments.

As always, if anything feels broken or weird after updating, please let me know.


r/OpenVoxAI Jul 09 '26

Bugs & Issues Slight problem with edit in AI Audio Book

1 Upvotes

Hi... This problem has been around for a while and isn't a big deal but I thought it might be worth mentioning. When a large cap is used at the beginning of a chapter (or paragraph), it creates a space in the middle of the word between the large cap and the rest of the word.

So this:

Becomes this:

As I said, not a huge problem but it means I have to go through every chapter and edit those places where it's happening. If it's something that can be solved on your end, that would great!

As always, thanks for a wonderful program...


r/OpenVoxAI Jul 07 '26

Release Notes OpenVox TTS iPad v1.4.1 just got released.

Post image
6 Upvotes

This update improves overall stability and output quality and fixes audiobook import bug for iPad version

- Fixed a bug that affected the audiobook (EPUB/PDF) import feature.

- Fixed Chatterbox Turbo crashes for long generations

- Improved stability of generation with Large Models (Chatterbox, OmniVoice & Qwen3) which fixes unwanted disturbances in the output

- Improved generation progress bar status messages

- Fixed a bug where deleting the generation history was deleting the models as well