AI Audio
-
Fun-ASR-Realtime Alibaba Cloud
A large streaming real-time speech recognition model launched by Alibaba Qianwen supports real-time recognition of 16 dialects and 30 languages.
-
Fun-CosyVoice 3.5 FunAudioLLM
The multilingual speech generation model launched by the FunAudioLLM team of Alibaba Tongyi Laboratory supports FreeStyle natural language command control and zero-sample speech cloning.
-
3D-Speaker
ModelScope is an open source multi-modal voiceprint recognition and speaker separation toolkit.
-
Gensfx Gensfx
Free AI sound effect generator, convert text into high-quality sound effects with one click
-
ACE-Step 1.5 ACE-Step
Open source music basic model, consumer-grade graphics cards can also run
-
AI Music Generator (AI Song Maker) AI Song Maker
Enter text to generate a complete song, with AI composing, lyrics, and covers all in one package
-
Aispect Aispect
Turn live speech into visual images in real time, giving abstract speeches a shareable visual expression
-
Audiify AI Audify AI
High-quality TTS tools for free with OpenAI keys, no subscription, no surprises
-
Beatoven Beatoven
AI generates copyright-free background music to customize exclusive soundtracks for videos and podcasts
-
Coqui Coqui
Coqui takes executable task flow as its core and emphasizes stable delivery and efficiency improvement.
-
Ekhos AI
AI speech synthesis and audio content creation platform










