Heard
A voice layer for macOS that turns terminal output from Claude Code, Codex and Cursor into smart spoken summaries
Heard
A voice layer for macOS that turns terminal output from Claude Code, Codex and Cursor into smart spoken summaries
Hitoo
Real-time voice translation video call platform, using the self-developed end-to-end AI model AURIS, achieves sub-300 millisecond real-time translation between 50+ languages while retaining the identity of the speaker's voice.
Meco
A reading app that removes news subscriptions from your mailbox and uses AI audio briefings to listen to key points every day
Kyutai Pocket-TTS
An open-source end-side speech synthesis model with 100 million parameters. The CPU runs in real time and can clone any sound in 5 seconds of audio.
VocalVia
An AI workspace that turns PDFs, articles and notes into editable multi-role podcasts
Doubao App 2.0
字节跳动旗下国民级 AI 智能 assistant,集文本, 图像, 视频, 语音 and 办公 Agent 于一体
EdgeTTS Online Dubbing
Based on 微软 Edge 神经语音的免费在线文字转语音 tool,免登录, 无广告, 同步generate SRT 字幕
FreeTTS Chinese TTS
永久免费, 无需注册的中文神经语音合成 tool,覆盖普通话, 粤语 and 台湾国语,一键导出 MP3 and SRT 字幕
ElevenLabs v4
The fourth-generation AI voice model previewed by ElevenLabs focuses on human-level expressions that can whisper, sing, and have emotions and accents.
Murf 2
A full-stack AI dubbing platform for teams and enterprises that packages scripting, voice generation, video synchronization and compliance into one studio
Suno 3.0
The latest generation capabilities of the AI music generation platform, focusing on song line-level refinement, multi-lingual natural vocals and professional mastering-level split-track export
Spotify AI DJ 2.0
Spotify's AI virtual radio host created for Premium users, generates personalized music mixes and voice commentary in real time based on listening habits
GPT-Live Voice
OpenAI next-generation 全双工语音模型,让 ChatGPT 语音对话更自然, 更聪明, 更会倾听
Read PDF Aloud
Free online AI PDF voice reading tool, supports 142 languages and 470 types of AI voices, and uses a local-first architecture to process text parsing
Video to Text
AI-based video and audio transcription tool that supports 99 languages, speaker separation and multi-format export, with a pay-as-you-go model
VoiceDrop
Speak and it archives itself - a voice-driven article writing app
VibeVoice
Microsoft's open source speech AI model family supports 90-minute long audio generation
SoundView
讯飞旗下AI视频本地化创作 platform,专注视频翻译, 配音 and 跨境电商视频创作
Anylang.ai
AI video translation tool, keeping the original speaker’s voice and lip synchronization
Magicam
Real-time AI face-changing and voice-changing tool, supporting face replacement in live video calls
Yunmu Tongsheng
原声级AI视频翻译 tool,supports 20多种语言,98%还原音色 and 情绪,保留背景音乐
LPM 1.0
17 billion parameter real-time full-duplex conversation AI model for video character performance
ElevenLabs
AI speech generation and processing platform, providing text-to-speech, speech cloning and voice agent services
Speechify
AI text-to-speech application converts documents and web pages into natural speech, supporting 60+ languages and 1000+ voices
Dub AI
AI-driven video translation and dubbing platform helps quickly localize video content into more than 30 languages
Rask AI
AI-driven video localization and dubbing platform that supports automatic translation in 130+ languages
VozoAI
AI video translation, dubbing and lip synchronization tool, supporting 110+ languages
Vapi
A voice AI platform for developers to help quickly build, test and deploy voice agents
LobeChat
Open source high-performance chatbot framework that supports multi-model, speech synthesis and multi-modality
iFlyTek Spark
科大讯飞launched 的认知large model,voice interaction and 教育领域优势突出