Murf
Audio & musicCreate voiceovers from scripts and refine delivery with voice, pronunciation and pacing controls. Murf Studio serves content production, while Murf’s API, dubbing and conversational-agent offerings address different workflows.
Create voiceovers from scripts and refine delivery with voice, pronunciation and pacing controls. Murf Studio serves content production, while Murf’s API, dubbing and conversational-agent offerings address different workflows.
Listen to written material through text-to-speech voices and adjustable playback. Speechify combines reading tools with premium scanning, summaries and voice features, while its Studio and API are separate products.
A free browser-based speech-recognition demo built with Transformers.js and Whisper models. It accepts audio from a file, URL or recording and performs inference on the device, with transcript text and JSON export in the documented application.
An open-source voice-cloning system that combines reference voice color with generated speech. OpenVoice supports multilingual workflows and style control through its underlying speech pipeline, with free MIT-licensed self-hosted models and code.
Browser-based speech cleanup and podcast editing with a continuing free plan. Enhance Speech can process up to one hour of audio per day, with per-file limits and premium options for video and larger workflows.
Free local AI plugins that add noise suppression, stem separation, speech transcription, music generation and audio enhancement to a compatible Audacity installation.
A locally deployable speech synthesis project with reference-voice generation, emotion and pronunciation controls; use is governed by Bilibili’s custom model agreement.
A conversational speech model for Chinese and English research, with controls for speaker characteristics and delivery; the released weights are restricted to noncommercial educational and research use.
A free open-source speech toolkit for transcription, speech activity detection, punctuation and configurable speaker pipelines, with local and service deployment options.
A free self-hosted speech generation toolkit with reference-voice synthesis, multilingual models, instruction-based controls and streaming integration paths.
A free self-hosted text-to-speech toolkit with a browser interface, reference-voice synthesis, optional fine-tuning and utilities for preparing speech datasets.
An AI-assisted audio post-production service for leveling speech, reducing noise and preparing consistent output, with 2 free processing hours per month and a jingle on free productions.