YueLu AI intelligent voice tool, support recording to text, speaker differentiation, intelligent summary and other multi-functions, suitable for learning, work and life and other scenarios. 02,5740 AI speech generation# AI Voice
MiniMax Audio MiniMax presents an AI speech synthesis tool based on the advanced T2A-01 speech model that supports multi-language, multi-tone selection and advanced parameter control. 04,6820 AI speech generation# Speech Synthesis
Noiz AI Text-to-speech and video dubbing tools, with self-developed voice models to achieve high-quality, emotionally rich voice synthesis, suitable for multi-scene content creation. 07,3770 AI speech generation# Video Dubbing# Speech Generation
Narakeet AI text-to-speech and video dubbing tool with multi-language and multi-tone support for video narration, PPT voice presentations and subtitle generation, easy to operate and natural voice. 02,4810 AI speech generation# Video Dubbing# Speech Generation
VoiSpark AI speech generation tool that supports text-to-speech, voice cloning and voice change, helping to create high-quality voice content. 01,7780 AI speech generation# Speech Generation
KittenTTS An open source lightweight text-to-speech model that is less than 25 MB and can run in real time on ordinary CPUs, supports a variety of natural tones and can be used offline. 03,2880 AI speech generationOpen Source Project# TTS# Video Generation
MAI-Voice-1 Microsoft has introduced an efficient speech generation model that generates natural and smooth high-fidelity audio in seconds, which has been applied to scenarios such as news broadcasting, podcasting and Copilot voice interaction. 01,3930 AI speech generation# Speech Generation
Qwen3-ASR-Flash Alibaba has introduced a multi-language high-precision speech recognition model that supports complex scenes, dialect and song transcription, and can be intelligently customized for recognition in context. 04,1490 AI speech generation# Speech Recognition
UntitledPen The full-featured creation platform based on AI technology integrates intelligent writing, multi-language speech generation and audio editing, helping users efficiently complete a one-stop solution for text creation, speech customization and post-production. 01,2340 AI speech generation# AI Creation# Writing Assistance# Speech Generation
BeFreed A smart learning tool that uses AI technology to provide personalized audio learning content, supports real-time Q&A and smart recommendations, and efficiently utilizes fragmented time to facilitate knowledge acquisition. 08310 AI speech generation# Podcasting Tools# Audio Summary
AudioPod AI AI audio creation tool, voice cloning, noise reduction and translation in one click, 3 minutes to generate professional content, support 21 languages, easy to achieve globalization and dissemination. 08410 AI speech generationAI Audio Processing# Digital Split
PrismAudio Ali launched the video to generate audio framework, through the “chain of thought + reinforcement learning” technology to achieve a high degree of synchronization of audio and video, can efficiently generate environmental sound effects, applicable to film and television, games, short videos and other multi-scene creation. 07210 AI speech generation# Audio Generation
Voxtral TTS Mistral AI introduces an open source, low-latency text-to-speech model that supports cross-language timbre cloning with latency as low as 70ms and can be deployed at the edge. 08440 AI speech generationOpen Source Project# Open Source# Text-to-speech
CosyVoice Alibaba's open-source large-scale speech model supports zero-shot cloning in 3 seconds, multilingual capabilities, and command-based emotional control, enabling ultra-low-latency streaming synthesis at 150 ms. 03140 AI speech generationLarge Model# Large Language Model