创意与媒体Reddit 原帖

OpenVox 1.6.4 is out: local AI TTS, better audiobooks, M4B export, Conversations, and Local API

完整上下文原始内容 · Reddit

Hey everyone, I posted OpenVox here about a month ago when it was around version 1.4. Since then I’ve been shipping updates pretty aggressively, and the app is now at v1.6.4. Just wanted to come back and say thanks, because a lot of the improvements were directly because of feedback I got from r/macapps People pointed out rough edges, asked for better audiobook support, asked for automation/API use cases, reported bugs, and also gave suggestions that I probably would not have prioritized this quickly on my own. The biggest area that changed is the audiobook workflow. It is honestly much more polished now. Since 1.4, I’ve added / improved: Local API for using OpenVox with agents, automations, and external tools Conversations mode for multi-speaker scripts, interviews, skits, podcasts, and dialogue Settings are now easier to access from the top title bar Audiobook export is much faster and more reliable M4A and M4B audiobook export support M4B support with better audiobook compatibility, chapters, and timeline support EPUB cover support inside the app Cover artwork embedding for exported M4A and M4B audiobooks Large audiobook exports now use less memory Batch audiobook generation progress is more accurate Temporary audiobook files are cleaned up better Removed slower MP3 export because M4A/M4B makes more sense for the audiobook flow now The audiobook feature started pretty basic, but after all the feedback around large books, storage usage, chapters, export formats, covers, and reliability. Still not perfect, and I know there’s more to improve, but v1.6.4 feels like a much more solid version than what I posted last month. I also made a short video showing what changed since the earlier version. Thanks again to everyone who tried it, criticized it, suggested things, or reported bugs. App Store: https://apps.apple.com/us/app/openvox-local-voice-ai/id6758789314?mt=12 Website: https://openvoxai.com/

01需求标签
创意与媒体缺陷修复桌面应用未解决

已收集讨论

25 条已收集

25条已收集37条 Reddit 标称评论
u/nicolas1410

That's looking cool! Can it be used with local models as well?

u/rich_awo

This is cool, is it just for EPUBs? and what do you use for diarisation on the meetings etc?

u/ritzynitz楼主回复

Thanks! No, not only EPUB. EPUB works best for audiobook workflow, but you can also use normal txt, paste content chapterwise or use Pdf ebooks. But chapter detection is not as accurate as epubs. For diarisation, OpenVox does not do that currently. It does not take a meeting recording and detect who spoke when. The Conversations feature is more for AI audio generation. Like you write/paste a dialogue, assign different voices to speakers, and then generate it locally.

u/Easy-Cobbler-1631

I really enjoy using this, it's well poslished and easy to use.

u/ritzynitz楼主回复

Great! Do let me know if you want anything improved!

u/cronberry

Love this app. Works so well. Fantastic at voice cloning. Thoroughly recommended.

u/DilshadZhou

I need someone to go through and test all of these local TTS apps that are cropping up thoroughly and tell me which to use. I don’t have time to test them all but definitely need one.

u/ritzynitz楼主回复

I mean if I call my App the best it would be not fair but from pricing perspective it's the best value for money. It's $20 for lifetime usage of 5 AI models and more models would be added in future as well.

u/ritzynitz楼主回复

Funciona super bem! Além disso, tem uma versão gratuita vitalícia que você mesmo pode testar.

u/ritzynitz楼主回复

Deve funcionar com o app/API local do OpenVox, mas eu ainda não testei pessoalmente com o Hermes Agent.

u/Old_Conversation1900

I really like the app and have been using it to create audiobooks! I have two questions/comments: 1. I cant figure out how to add cover art to an audiobook from inside the app. Can you tell me how? 2. when generating an audiobook using Qwen3 TTS Large the generate all seems to get stuck. Medium seems to work fine. My audiobook is 13 chapters of about 20,000 characters each at .9 speed. I'm on a m5 macbook pro 48 ram Tahoe 26.4.1. Thank you in advance

u/ritzynitz楼主回复

Yes it does handle thrm well and chatterbox Turbo and OmniVoice models support emotion tags as well. There is a lifetime free tier, so you can try it yourself.