Voice & Audio Editing
Compare the AIGC List products currently published for Voice & Audio Editing.
Yak is a cross-application voice AI editor that transforms selected text through spoken commands. Supporting over 100 languages with automatic detection, it offers shortening, lengthening, tone adjustment, translation, and rewriting. A free tier provides 100 weekly requests; Pro unlocks unlimited cloud AI for $12–15/month.
LANDR is an AI-powered music production platform centered on automated online mastering. The core engine provides three intensity and style presets calibrated by professional mastering engineers, reference track matching, and DAW integration via a dedicated Mastering Plugin. A Plugin Marketplace offers 40+ curated production tools, while a published OpenAPI specification enables programmatic mastering for developers and media organizations. The platform maintains multilingual documentation across English, French, Japanese, and German.
Descript is an AI-powered video and audio editing platform built around transcript-based editing. Import media, edit by modifying the auto-generated text transcript, and let the Underlord AI assistant handle filler word removal, Studio Sound enhancement, caption generation, and clip creation. Includes screen recording, team collaboration, multi-platform publishing, and a REST API with MCP integration at no extra cost for paying users. Free tier available with 1 hour of monthly transcription; paid plans from $12/month.
VideoAny is a multi-model AI creative platform offering video generation, image generation, and a distinctive video extender that analyzes final frames to extend clips beyond the typical 4-8 second AI generation cap. It operates on a freemium model with paid plans for higher resolution and batch processing.
Aquavoice delivers two products under one brand: a consumer AI voice keyboard on iOS and the Avalon speech-to-text API for developers. Avalon is marketed as a two-line Whisper replacement trained on developer workflows—CLI sessions, IDE captures, and real engineering dictation rather than audiobook narration. The vendor reports 97.3% accuracy on its AISpeak benchmark and claims literal transcript fidelity that preserves casing, commands, and model numbers without hallucination. The iOS keyboard brings the same speech engine to everyday dictation. Enterprise posture is supported through a Vanta trust center and a public status page, while account management runs through a web dashboard. Pricing and multilingual language specifics remain absent from the available source pack.