Text-to-speech generation
Generate speech from text with a large voice library. The home and voices pages describe 2,000+ AI voices across 130 languages, with plan pages showing default and Pro voice libraries.
Voicemaker is a browser-based text-to-speech platform that converts written text into synthetic speech. The site presents it as a tool for generating audio in multiple formats and for adjusting how the output sounds before download.
Across the home and pricing pages, Voicemaker positions itself around a large voice library, multilingual output, and controls for timing and delivery. The product also extends beyond basic TTS with tools for pronunciation editing, speech-to-speech, voice cloning, studio-style project work, and plan-based options for teams and higher-volume users.
Generate speech from text with a large voice library. The home and voices pages describe 2,000+ AI voices across 130 languages, with plan pages showing default and Pro voice libraries.
Adjust delivery with pause, speed, pitch, volume, emphasis, and voice-effect controls. The site also exposes pronunciation editing and metadata tag options in the web app.
Export audio in common file formats. The product pages mention MP3 and WAV downloads, and the API page also lists OGG, AAC, and OPUS output options.
Work with longer or more structured projects using the studio tools. The pricing page describes VoxStudio, Projects, background music mixing, subtitle export, file history, and cloud storage.
Use advanced voice workflows such as speech-to-speech, voice cloning, and subtitle generation. These capabilities are shown on the pricing and home pages, though availability varies by plan and voice type.
Support team and organization use with Business features such as Enterprise SSO, Team Workspace, seat management, and usage monitoring.
Turn scripts, notes, or short-form copy into spoken audio for social clips, presentations, or other publishable assets. The home page explicitly mentions YouTube Shorts, videos, and presentations as target outputs.
Use pronunciation, pause, pitch, speed, and voice-effect controls to shape delivery for narration, customer-facing audio, or branded voice work. This is useful when natural pacing matters more than raw conversion.
Build longer audio projects with VoxStudio, projects, background music, subtitle export, and cloud storage. These tools are aimed at users managing multiple audio elements in one workflow.
Use the Business plan for multi-seat workflows that need SSO, role-based access, file sharing, and usage monitoring. The pricing page frames this tier for teams and businesses scaling content production.
Convert text to speech through the API when the goal is to automate generation inside another product or pipeline. The pricing page describes a developer API with customizable speech controls and RESTful access.
You can start with the free plan by registering an account. The pricing page also shows paid Starter, Premium, and Business plans for users who need higher limits and additional features.
The source shows downloadable audio in MP3 and WAV formats on the home page, and the API/pricing pages also mention OGG, AAC, and OPUS for supported workflows.
Voicemaker includes controls for pauses, pronunciation, speed, pitch, volume, and voice effects in the web app and API. Some controls are limited to certain voice types or paid plans.
The pricing page shows support for individual, premium, and business usage, including cloud storage, file history, team workspace, SSO, and broadcasting rights on higher plans. It also states that commercial rights are included on paid plans, while broadcast rights are separate.
The refund policy says subscriptions are managed through auto-renewal and that cancellations or refunds are generally not offered, except in cases approved by support.
流量数据仅供参考。
speechify.ai
SpeechifyAI is a developer API for expressive text-to-speech and voice cloning. It provides streaming Simba models, catalog and cloned voices, SSML support, and a free starting tier for building speech into applications.
deepmind.google
Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS are audio-generation models for creating expressive voices, directing spoken performances, and producing conversational audio. They support creative teams, developers, and enterprises working across Google AI Studio, Gemini API, Gemini Enterprise, Gemini Notebook, and Google Vids.
twinelauncher.com
适用于 Android 的启动器与 AI 助手,提供类似 ChatGPT 的帮助、笔记、任务、提醒和消息撰写功能。
xiaoyi.huawei.com
小艺是华为自主研发的 AI 智慧助手,支持问答、写作、文档阅读、代码辅助和识图。
altered.ai
面向媒体制作和实时语音应用的 AI 变声与语音内容创作平台。
mmaudio.net
将视频转换为音频的 AI 语音生成工具