Audio, Voice & Music
Voicemod alternatives
16 alternatives to Voicemod (Voicemod, Inc., Sucursal en España) from the Audio, Voice & Music category of the index.
- AIVA (AIVA Technologies) — AI music composition assistant that generates original tracks in many styles for media projects. Freemium
- AssemblyAI (AssemblyAI) — Speech-to-text API providing transcription and audio intelligence models for developers. Paid
- Cartesia (Cartesia) — Real-time voice generation platform built on state-space models for low-latency speech synthesis. Freemium
- Deepdub (Deepdub) — AI voice platform for dubbing, localization, text-to-speech, and voice conversion, supporting 100+ languages. Freemium
- Deepgram (Deepgram) — Speech recognition and voice AI API for real-time transcription and audio understanding. Paid
- Descript (Descript) — AI-powered audio and video editor that lets you edit recordings by editing their transcript. Freemium
- ElevenLabs (ElevenLabs) — AI voice platform offering text-to-speech, voice cloning, and dubbing. Freemium
- Murf AI (Murf AI) — AI text-to-speech platform offering studio-quality synthetic voices for voiceovers and narration. Freemium
- Resemble AI (Resemble AI) — Voice cloning and speech synthesis platform with real-time voice generation and deepfake detection. Paid
- Respeecher (Respeecher, Inc.) — Speech AI platform offering speech-to-speech conversion, voice cloning, a real-time text-to-speech API, a Pro Tools plugin, and a marketplace of licensed voice talents. Paid
- Sesame (Sesame AI) — Conversational voice AI from the Oculus co-founder's startup, known for its lifelike Maya voice companion demo. Free
- Speechify (Speechify) — Text-to-speech app that reads documents, articles, and books aloud in natural voices. Freemium
- Stable Audio (Stability AI) — AI music and audio generation tool that creates tracks and sound effects from text prompts. Freemium
- Suno (Suno, Inc.) — AI music generation service that creates full songs with vocals from text prompts. Freemium
- Udio (Udio) — AI music platform that creates songs with vocals and instruments from text prompts, now trained only on authorized and licensed music after a Universal Music Group agreement. Freemium
- Whisper (OpenAI) — OpenAI's open-source automatic speech recognition model; OpenAI's current transcription models are the gpt-transcribe and gpt-4o-transcribe family, while whisper-1 remains available. Open (open weights/source)
Full Voicemod profile → · All Audio, Voice & Music products →