HuggingFace Space hkchengrex/MMAudio.
220 AI tools matching Audio, with pricing, category and a link to each official site. Page 5 of 6.
220 tools · page 5 of 6
HuggingFace Space hkchengrex/MMAudio.
HuggingFace Space innoai/Edge-TTS-Text-to-Speech.
HuggingFace Space k2-fsa/OmniVoice.
HuggingFace Space myshell-ai/OpenVoice.
HuggingFace Space NihalGazi/Text-To-Speech-Unlimited.
HuggingFace Space hf-audio/open_asr_leaderboard.
HuggingFace Space tonyassi/voice-clone.
HuggingFace model Qwen/Qwen3-TTS-12Hz-1.7B-CustomVoice (text-to-speech).
HuggingFace Space facebook/MusicGen.
HuggingFace model MahmoudAshraf/mms-300m-1130-forced-aligner (automatic-speech-recognition).
HuggingFace model jonatasgrosman/wav2vec2-large-xlsr-53-polish (automatic-speech-recognition).
HuggingFace model openai/whisper-small (automatic-speech-recognition).
HuggingFace model jonatasgrosman/wav2vec2-large-xlsr-53-russian (automatic-speech-recognition).
HuggingFace model pyannote/segmentation (voice-activity-detection).
HuggingFace model openai/whisper-base (automatic-speech-recognition).
HuggingFace model pyannote/segmentation-3.0 (voice-activity-detection).
HuggingFace model jonatasgrosman/wav2vec2-large-xlsr-53-japanese (automatic-speech-recognition).
HuggingFace model pyannote/speaker-diarization-community-1 (automatic-speech-recognition).
HuggingFace model openai/whisper-large-v3 (automatic-speech-recognition).
HuggingFace model jonatasgrosman/wav2vec2-large-xlsr-53-portuguese (automatic-speech-recognition).
HuggingFace model pyannote/voice-activity-detection (automatic-speech-recognition).
HuggingFace model Qwen/Qwen3-ASR-1.7B (automatic-speech-recognition).
HuggingFace model Qwen/Qwen3-ASR-0.6B (automatic-speech-recognition).
HuggingFace model argmaxinc/whisperkit-coreml (automatic-speech-recognition).
HuggingFace model pyannote/speaker-diarization-3.1 (automatic-speech-recognition).
HuggingFace model coqui/XTTS-v2 (text-to-speech).
HuggingFace model openai/whisper-large-v3-turbo (automatic-speech-recognition).
HuggingFace model laion/clap-htsat-fused (audio-classification).
HuggingFace model hexgrad/Kokoro-82M (text-to-speech).
Visit tool
AI voice generator
MiniMax Music 3.0 Launches as Open-Weights AI Music Generation Model minimax.io · Hide TLDR
6 – Granola Productivity Automation & Agents Speech-To-Text Granola turns any meeting into clean
7 – Wispr Flow Speech-To-Text Dictate anything, get polished text instantly — write 3x faster withou
12 – Eleven Labs Text-To-Speech Create natural sounding voices for creators and publishers
Acton, Massachusetts teen held without bail in murders of mother, brother linked to ChatGPT
voicepod.ai
finevoice.ai
Stable Audio
MixAudio
Can’t find a tool? Tell us about it.