Readr is a free and open source ebook reader for iPhone, iPad, and Mac that lets you ask questions about a book, listen to it, and turn highlights into Markdown notes or draft articles. It supports DRM-free EPUBs, PDFs, and Markdown.
Seed Audio AI is a browser-based text-to-speech and voice cloning tool for turning scripts into reviewable voice audio. It is aimed at creators and teams that need drafts for voiceovers, narration, podcasts, lessons, and ads.
Speko is a voice AI router that benchmarks speech models by language and routes STT, LLM, and TTS through one API. It also offers a packaged infrastructure option and an enterprise contract path.
Chariot is an AI text-to-speech product for generating English, Hindi, and Hinglish speech from text, with low-latency streaming and API access for developers. It is aimed at voice agents, app audio, and other production workflows that need natural-sounding speech.
Wisprkey is a macOS AI assistant for voice typing and text-to-speech. It works system-wide on Apple silicon Macs and is aimed at builders and teams who want to type less while working in any app.
Heard is a macOS app that turns coding agent activity into spoken updates for developers working in Claude Code or Codex. It supports different listening modes, customizable voices, and a higher-tier mobile pairing flow for hands-free follow-up.
Hitoo is a real-time AI translation platform for multilingual voice conversations. It helps people speak across languages in live calls while preserving voice identity and context.
Oakamo is a calm reading space for saving articles, reading without distractions, listening to articles as audio, and returning to a synced personal library across devices.
VersaVoice AI is a cross-lingual voice communication app for voice messages, text translation, voice cloning, and in-person translation. It is designed to help people and teams communicate across languages while keeping the speaker's voice and identity.
MonstaReel is an AI short-form video creator that turns a topic into a scripted vertical MP4 with voice, captions, and export for TikTok, Instagram Reels, and YouTube Shorts. It is designed for creators who want to produce short videos faster without handling each editing step manually.
Narration Room turns scripts into playable narrations on Apple devices using on-device voices. It supports writing, pasting, importing, or dictating text, with a local library and optional iCloud sync for script copies.
Labs AI : Text to Speech is an iPhone app that turns written text into natural-sounding speech with customizable voices and multi-language support. It is aimed at creators and professionals who need quick voiceovers or dubbed audio from text.
Gemini 3.5 Live Translate is Google’s near real-time speech translation model for developers, Google Meet, and the Google Translate app. It supports 70+ languages and is designed to produce natural-sounding translated audio during live conversations.
MAI-Voice-2 is Microsoft AI’s text-to-speech model for natural, expressive speech in assistants, support experiences, long-form narration, and accessibility use cases. It is available in Microsoft Foundry and supports 15 languages/locales, emotion control, and short-reference custom voice creation.
Voiser AI Voiceover turns text into spoken audio for voiceovers, with multilingual voice options and style controls for different narration needs. It supports a web studio workflow and shows free, paid, and enterprise paths on the site.
Our Stories is a family storytelling web app for creating, reading, and listening to custom stories in multiple languages. It is designed for multilingual households and for families who want to share bedtime stories across distance.
Wallie is an open-source AI streamer that watches your screen, hears chat, and generates live commentary in a configurable persona. It runs locally on your machine with your own keys and is aimed at faceless content, autonomous streams, and real-time reactions.
Reader Alive is an AI ebook reader for iPhone and iPad that supports EPUB, PDF, MOBI and AZW3 files. It adds translation, text-to-speech, summaries and book-aware chat for personal ebook libraries.
Selectable is a macOS OCR and text-capture utility for extracting text from anywhere on your screen, including images and videos. It supports copying, translation on newer macOS versions, text-to-speech, and output cleanup.
FlowSpeech is a context-aware text-to-speech studio that turns scripts and uploaded files into human-like audio. It offers multiple generation modes, pause and emotion control, and a free plan alongside paid tiers.