Speech-to-text transcription and voice dictation — upload recordings or dictate anywhere you type, in 100+ languages.
Voice dictation — Composing Music & Audio (Top 50 of 223)
Speak instead of type — AI voice dictation that turns natural speech into polished text in any app.
Real-time AI voice changer, cloning, and text-to-speech — plus voice agents that handle calls end-to-end.
Text-to-speech and transcription studio: 550+ AI voices across 75+ languages, plus voice cloning, dubbing, and subtitles.
Dictate polished, formatted text with your voice — 4x faster than typing, in any app.
Turn text into natural AI speech in 50+ languages — 1,000+ voices, voice cloning and video dubbing, with best-in-class Vietnamese.
Lifelike AI text-to-speech, voice cloning, dubbing, music, and real-time voice agents in 70+ languages.
Free AI text-to-speech in 50+ languages and 300+ voices — paste a script, download the voiceover.
Iran's all-in-one Persian AI hub: chatbot, image generation and editing, text-to-speech, transcription, and voice cloning.
Conversational speech-to-text that knows when a speaker is done talking, built for real-time voice agents.
AI dictation that turns your voice into polished, app-aware text — hands-free, offline-capable, in 100+ languages.
Open-source, local-first AI dictation that turns your voice into app-aware polished text on Mac, Windows, and mobile.
Online text-to-speech synthesizer with voices across ~50 languages, SSML control, and MP3/OGG/WAV export — free up to 500 characters.
No-code talking avatars for your website — 500+ characters or one built from a photo, text-to-speech voices, and an optional AI bot layer.
Conversational speech models for real-time voice agents — human-sounding phone calls with sub-250ms time-to-first-audio.
All-in-one AI platform to generate and edit images — plus video, music, and speech — from a text prompt.
Mac and iOS voice dictation that writes what you meant, not what you mumbled, in 100+ languages
Build and deploy real-time voice agents on Cartesia's Sonic TTS and Ink STT speech models.
Assistive Chrome extension for K-12: read-aloud, dictation, 60+ language translation, and teacher voice feedback in Google Workspace.
AI transcription that turns your voice, audio, and video into accurate text — with summaries and notes in 50+ languages.
HIPAA-compliant clinical dictation that turns spoken notes into structured documentation inside any EHR — built for home health and hospice.
The Maldives' Dhivehi-first AI: draft letters, essays and poetry in Thaana, plus Dhivehi speech-to-text, TTS and OCR.
AI voice interpreter: a pocket device, app, and Sentio live-interpretation service that translate speech, signs, and meetings across 90+ languages.
AI moderation, anti-spam, analytics, and voice-to-text for Telegram communities.
iFlytek's AI voice-to-text: transcribe, translate, subtitle, and auto-summarize meetings and recordings.
Lifelike AI voices that read anything aloud — plus voice typing, voice cloning, and AI podcasts.
All-in-one AI platform for voiceovers, music, and video — 3,200+ TTS voices, voice cloning, AI songs, and text-to-video.
Clone your voice, generate AI covers, and turn text or a hummed melody into a finished song.
All-in-one AI workspace for generating images, video, voice, and text, with notes and cloud storage built in.
All-in-one AI video agent that turns text or photos into finished videos, avatars, voice, and music.
Clone any voice from a short audio sample and make it speak your text in 18 languages, with inline emotion control.
Re-voice any audio with RVC v2 — upload a TTS clip or vocal, pick from 20,000+ community voice models, export a clean WAV or MP3.
Chrome extension that transcribes WhatsApp Web voice notes to text, plus summaries, live translation, and scheduled replies.
Turn text prompts or lyrics into full royalty-free AI songs, with voice cloning, stem splitting, and mixing.
Free AI music studio: generate songs from text, clone voices for covers, split stems, and make karaoke or sheet music.
Persian-language AI voice and text assistant for chat, writing, images, translation, and everyday Iranian services.
Free, no-signup AI image generator and editor — plus text-to-video, lip-sync, and voice cloning in the browser.
Russian-language AI aggregator — chat, text, images, video and voice from GPT, Claude, Gemini and more in one browser tab, no VPN or signup.
Russian gateway to 120+ AI models — chat and OpenAI-compatible API for text, images, video, and voice, no VPN, priced in rubles.
Speak instead of typing — turns voice into polished emails, notes, posts, and meeting recaps in 90+ languages.
Japanese enterprise AI translation at TOEIC 960-level accuracy — text, Office and PDF files, voice, and video subtitles in ~20 languages.
Turn text prompts or a single image into HD AI videos — talking avatars, 800+ voices, and a timeline editor, for a one-time fee.
Talk to ChatGPT out loud — real-time, interruptible voice conversation, live translation, and camera/screen sharing.
Russian-language AI music generator — text-to-song, covers, voice cloning, vocal removal. Unofficial third-party service, not operated by Suno Inc.
Record meetings and stray thoughts, get instant transcripts, summaries, and action items — then ask AI about anything you said.
Generate celebrity and character AI voices, voice conversions, and lip-synced videos for content and memes.
Ambient AI scribe for clinicians — turns patient conversations into SOAP notes, dictation, forms, and billing codes in your EMR.
Real-time AI transcription, translation, and synthesized-voice interpretation that lets multilingual audiences follow a live event on their own phones.
Pick a celebrity voice, type your line, and get a shareable voiceover or video — 100+ AI voices for pranks, greetings, and memes.
All-in-one AI studio: type an idea and generate images, full songs, cloned-voice narration, and written copy in one library.