Hitoo is a real-time AI translation platform for multilingual voice calls, helping people speak across languages while preserving voice identity and context.
Oakamo is a calm reading space to save articles, read distraction-free, listen as audio, and sync your personal library across devices.
VersaVoice AI is a cross-lingual voice app for voice messages, text translation, voice cloning, and in-person translation across languages.
MonstaReel is an AI short-form video creator that turns topics into scripted vertical MP4s with voice, captions, and export for TikTok, Reels, and Shorts.
Narration Room turns scripts into playable narrations on Apple devices with on-device voices. Write, paste, import, or dictate text, plus local library and optional iCloud sync.
Labs AI : Text to Speech turns text into natural-sounding speech with customizable voices and multilingual support for quick voiceovers or dubbing.
Gemini 3.5 Live Translate is Google’s near real-time speech translation model for developers, Google Meet, and Google Translate, with 70+ languages and natural-sounding audio.
MAI-Voice-2 is Microsoft AI’s text-to-speech model for natural, expressive speech in assistants, support, narration, and accessibility. Available in Microsoft Foundry.
Voiser AI voiceover turns text into spoken audio with multilingual voices and style controls for fast, natural voiceovers in the web studio.
Our Stories is a family storytelling web app for creating, reading, and listening to custom stories in multiple languages.
Wallie is an open-source AI streamer that watches your screen, hears chat, and delivers live commentary in a configurable persona. Runs locally with your own keys.
Reader Alive is an AI ebook reader for iPhone and iPad that supports EPUB, PDF, MOBI and AZW3 files, with translation, text-to-speech, summaries and book-aware chat.
Selectable is a macOS OCR and text capture utility for grabbing text from screens, images, and videos. Copy, translate on newer macOS versions, use text-to-speech, and clean output.
FlowSpeech is a context-aware text-to-speech studio that turns scripts and uploaded files into human-like audio. Free plan and paid tiers available.
Gemini 3.1 Flash TTS is Google’s preview text-to-speech model for expressive AI speech with fine-grained style and delivery control across Gemini API, Google AI Studio, Vertex AI, and Google Vids.
Smallest.ai Lightning TTS is a low-latency text-to-speech API with multilingual speech and fast voice cloning for voice agents and production audio workflows.
Claude voice mode is a beta feature for spoken conversations with Claude on the web and in Claude Mobile for iOS and Android, with hands-free and push-to-talk.
Read the Quran online for free with audio recitation and translations, including word-by-word analysis in 18 languages on easyquran.ai.
Voxtral TTS is Mistral’s text-to-speech model for lifelike multilingual speech, voice agents, and enterprise voice workflows with low latency.
Clipchamp AI voice generator is an online text-to-speech tool for video narration and dubbing, with multilingual voices and browser-based editing.