Acoust turns text into natural AI speech with 100+ voices in 30+ languages, plus voice cloning and a built-in video editor. Free plan available.
Acoust is a browser-based AI platform for converting written text into natural-sounding speech and turning it into finished video content. It combines text-to-speech generation with a built-in video editor, letting users produce voiceovers and videos without recording equipment or voice actors.
The platform offers over 100 AI voices across 30+ languages, plus AI voice cloning that can recreate a custom voice from a 10-second audio sample. Users can fine-tune generated speech with controls for tone, style, emotion, pacing, pauses, pitch, and speed. Additional tools include AI Clips for auto-generating short-form content with auto-subtitles from longer videos, and AI Translation for converting content into multiple languages.
Acoust is aimed at creators and businesses producing social media content, YouTube videos, e-learning and corporate training material, product and real estate videos, advertising, IVR prompts, and audiobook narration. A free plan is available for testing with limited credits and no commercial use; paid plans add higher usage limits, MP3 exports, commercial licensing, and team features.
Yes, Acoust has a free Personal plan with 10K credits (about 10 minutes of AI voice generation), though it does not include audio exports or commercial usage rights.
Paid plans start at $9/month (Pro) or $29/month (Premium) when billed monthly, with about 25% savings on annual billing. An Enterprise plan with custom pricing is also available for teams.
Yes, Acoust supports AI voice cloning, creating a custom voice from a 10-second audio sample.
Acoust offers over 100 AI voices across 30+ languages.