AI Audio
-
FunASR Tongyi Lab
Alibaba Tongyi Lab’s open source industrial-grade speech recognition toolkit, 170x real-time rate in 50+ languages, privatized deployment
-
AIVA AIVA
AI orchestral score generation platform supports 250+ styles and commercial copyright purchases
-
Audiify AI Audify AI
High-quality TTS tools for free with OpenAI keys, no subscription, no surprises
-
GPT Realtime Whisper OpenAI
OpenAI’s low-latency streaming transcription model for real-time subtitles, meeting recordings, and speech workflows
-
AudioCraft
Meta open source AI audio generation framework supports music, sound effects and compression encoding
-
Boomy Boomy
AI platform that lets anyone create original music in seconds
-
Cleanvoice AI Cleanvoice
AI audio post-processing tools for podcasting and interview scenarios, the real value is to clean up idioms, pauses and noise in batches instead of just transcribing
-
Deepgram API AIStartMap
Unified speech AI API platform, providing real-time speech recognition, synthesis, intelligent analysis and conversational AI Agent infrastructure
-
Ecrett Music ecrett
AI royalty-free soundtrack generation tool for creators, quickly generate commercially available music by scene, mood and genre
-
3D-Speaker
ModelScope is an open source multi-modal voiceprint recognition and speaker separation toolkit.
-
Eleven Music ElevenLabs
ElevenLabs’ AI music generation platform combines a creator workbench, commercial licensing marketplace, and embeddable Music API
-
Fish Audio Fish
A text-to-speech and speech cloning platform focusing on emotional expressiveness, with the open source Fish-Speech model behind it











