Automatic audio processing
Apply intelligent algorithms for loudness normalization, leveling, noise reduction, reverb reduction, filtering, and audio enhancement.
Auphonic is an automatic audio post-production web service for analyzing, improving, encoding, and publishing audio and video content. It supports browser-based productions as well as API, CLI, and workflow automation for podcast and other media processes.
Auphonic is an automatic audio post-production web service. It analyzes audio and video content and applies processing such as loudness normalization, leveling, noise and reverb reduction, filtering, enhancement, and speech recognition. Users can create productions in the web interface, select algorithms or presets, and export results in multiple formats. The service also supports automatic publishing to external services, REST API integrations, a command-line interface, and workflow automation.
Apply intelligent algorithms for loudness normalization, leveling, noise reduction, reverb reduction, filtering, and audio enhancement.
Paying users can use speech recognition features for transcripts, automatic shownotes, summaries, and chapters.
Create productions from uploaded or recorded audio and video, reuse presets, process multitrack material, and use features such as filler and silence cutting where applicable.
Generate multiple output formats and connect productions to external services for automatic publishing and delivery.
Use the REST API, Simple API, JSON API, Auphonic CLI, webhooks, watch folders, Zapier, or n8n-supported workflows to integrate processing into scripts and media pipelines.
Clean up spoken-word recordings, normalize loudness, create speech-based metadata, and prepare episodes for publishing.
Send multiple files to Auphonic through the Simple API or other automation tools when repeated manual processing would be inefficient.
Use presets, the full JSON API, or the CLI to control file formats, metadata, and outgoing services in a scripted production process.
Business teams can share team-account credits while keeping individual productions and presets separate from one another.
Yes. Productions can be created in the web service by uploading or recording audio or video and selecting algorithms or presets. Automation options are available separately for users who need scripted workflows.
Auphonic provides a Simple API for quick scripts and batch processing, a JSON REST API for detailed control, and a CLI that wraps the full API. The documentation also lists webhooks, watch folders, Zapier, and n8n integrations.
Yes. The service can encode results to multiple formats. The pricing page also states that output files such as MP3 and AAC do not change the billed processing duration.
Auphonic offers 2 hours of processed audio free each month. Paid options include recurring credits and one-time credits; recurring credits reset monthly and unused credits do not stack, while one-time credits do not expire. Billing is based on the duration of the submitted audio, with a 3-minute minimum per production.
Business users can request team-account access. Owners can invite members and admins who use shared team credits, while team members cannot access one another’s productions or presets. The owner purchases the credits.
www.waves.com
Clarity Vx is an AI-powered audio plugin for removing background noise from vocals and voice recordings. It is designed for music, podcast, production, and video workflows where clean speech or singing is needed before mixing.
xound.io
AI audio enhancement that removes noise and clarifies voices
audioenhancer.ai
Web-based AI tool for cleaning audio and video with speech, noise, loudness, and echo processing.
lalal.ai
AI audio separation for vocals, instruments, and voices
www.nvidia.com
NVIDIA Broadcast is a Windows app that uses AI to enhance microphone audio and webcam video for livestreams, voice chats, video conferences, and remote work. It provides selectable audio and video effects, with NVIDIA Broadcast chosen as the camera and microphone in a supported app.
podcast.adobe.com
Enhance Speech is Adobe Podcast’s web-based AI audio filter for reducing noise and echo in spoken recordings. It helps podcasters and video creators improve recorded dialogue without needing a perfect microphone or recording environment.