AI Audio
-
Fun-AudioGen-VD Alibaba Tongyi Lab
The timbre design model launched by Alibaba Tongyi Lab supports natural language description to generate timbre, emotion and scene-based audio.
-
Gemini 3.1 Flash TTS Google
Next-generation text-to-speech model launched by Google supports 70+ languages and audio tag director-level control
-
AssemblyAI API AssemblyAI
A voice AI API platform for developers, providing a full-link infrastructure from transcription, understanding to voice agent
-
Audio Flamingo Next AF Next
An open source audio language model for long audio understanding that unifies speech, music and ambient sounds into a set of reasoning frameworks
-
sponge music 海绵音乐
ByteDance’s free AI music creation platform generates singable songs in one sentence
-
Beatoven Beatoven
AI generates copyright-free background music to customize exclusive soundtracks for videos and podcasts
-
Castmagic Castmagic
AI-driven podcast content reuse engine that automatically converts audio into show notes, social media posts, timestamps and summaries
-
Coqui Coqui
Coqui takes executable task flow as its core and emphasizes stable delivery and efficiency improvement.
-
Descript Overdub
AI voice cloning and dubbing tools, use your voice to say anything
-
Ekhos AI
AI speech synthesis and audio content creation platform
-
Abogen Abogen
Open source e-book and document conversion tool that supports synchronized subtitles and multi-voice control
-
ElevenLabs ElevenLabs
ElevenLabs is a leading AI audio platform covering speech generation, speech agents, transcription and music generation, and provides complete API capabilities.










