AI Audio
-
Fun Asr Fun-ASR
Tongyi Lab’s open source end-to-end speech recognition model supports Chinese dialects and 31 languages.
-
Fun-AudioGen-VD Alibaba Tongyi Lab
The timbre design model launched by Alibaba Tongyi Lab supports natural language description to generate timbre, emotion and scene-based audio.
-
Gemini 3.1 Flash TTS Google
Next-generation text-to-speech model launched by Google supports 70+ languages and audio tag director-level control
-
ACE-Step ACE-Step
ACE Studio and StepFun jointly open source the basic model of music generation to support efficient generation and lyrics editing.
-
Harmonai Harmonai
Stability AI open source AI music generation platform
-
Aero-1-Audio Aero-1-Audio
A lightweight audio model with only 150 million parameters, supporting 15 minutes of continuous audio processing
-
AInterview AInterview
You are the guest and AI is the host, generating a podcast interview in a few minutes
-
AssemblyAI AssemblyAI
A voice intelligence platform with API as the core, emphasizing transcription accuracy and programmable voice capabilities.
-
Audio Sds Audio Sds
Audio Sds provides AI speech synthesis and audio processing API, supporting voice cloning and multilingual TTS
-
Cleanvoice AI Cleanvoice
AI audio post-processing tools for podcasting and interview scenarios, the real value is to clean up idioms, pauses and noise in batches instead of just transcribing
-
Descript Overdub
AI voice cloning and dubbing tools, use your voice to say anything
-
Eleven Music ElevenLabs
ElevenLabs’ AI music generation platform combines a creator workbench, commercial licensing marketplace, and embeddable Music API










