Category
Audio & Voice
Voice cloning, text-to-speech, music and transcription.
17 tools
ElevenLabs
Free plan
Voice cloning and text-to-speech that most people can't clock as AI.
Suno
Free plan
Full songs — vocals, lyrics and instrumentation — from a prompt.
Adobe Podcast
Free plan
Speech enhancement that makes phone audio sound studio-recorded.
AssemblyAI
Free plan
Speech-to-text API with diarisation, topics and summarisation.
Auphonic
Free plan
Automatic levelling and loudness mastering for spoken audio.
Cartesia
Free plan
Very low latency speech, built for realtime voice agents.
Deepgram
Free plan
Fast, cheap speech-to-text API built for high volume.
Fish Audio
Free plan
Open-weight speech models with a hosted API on top.
Krisp
Free plan
Strips background noise from calls in real time.
LALAL.AI
Free plan
Splits a mixed track into clean vocal and instrument stems.
Moises
Free plan
Stem separation, pitch and tempo tools aimed at musicians.
Murf
Free plan
Voiceover studio with commercial-use synthetic voices.
PlayAI
Free plan
Low-latency voice API for realtime agents and narration.
Resemble AI
Paid
Voice cloning with a deepfake-detection product alongside it.
Speechify
Free plan
Reads documents, PDFs and web pages aloud at speed.
Udio
Free plan
Music generation with fine control over sections and style.
Whisper
Free
OpenAI's open-source speech recognition, strong across accents.