WhisperUI logo

WhisperUI

Rivendica

WhisperUI is a web and desktop speech-to-text tool with text-to-speech, subtitles, and free transcript utilities for creators, researchers, students, and teams.

WhisperUI preview

Overview

WhisperUI is a speech-to-text and text-to-speech product built around OpenAI models. Its core transcription flow lets you upload audio files in the browser, send them to OpenAI Whisper, and review the resulting text or subtitle output for download.

The product also includes a desktop app for local transcription on Windows and macOS, plus a set of free browser tools for converting transcript and subtitle formats, cleaning rough exports, and estimating transcript length. Pricing pages show both cloud and local workflows, with a paid subscription available for desktop access and cloud usage.

Features

Web speech-to-text uploads

Upload audio through the web app by drag and drop or file browsing, then send it through OpenAI Whisper for transcription.

Text and SRT output

Export transcriptions as plain text and transform audio files into SRT subtitles from the speech-to-text workflow.

Local desktop processing

Use WhisperUI Desktop for local transcription with unlimited jobs, no file size limit, and no file duration limit.

Cloud processing on web and desktop

Choose cloud transcription when you want browser-based processing, with limits that vary by plan.

Text-to-speech generation

Use the text-to-speech tool to generate speech from entered text with OpenAI voices and models.

Free transcript utilities

Use the free browser tools to convert SRT, VTT, and TXT, clean rough transcripts, and estimate transcript length from audio duration.

Use Cases

  • Audio transcription

    Turn uploaded audio files into text for note-taking, review, or download through the browser-based WhisperUI transcription flow.

  • Caption and subtitle preparation

    Convert transcripts into SRT subtitles or use the subtitle utilities to move between SRT, VTT, and TXT formats.

  • Local offline-style workflows

    Run local transcription on a desktop machine when you want to keep processing on your own device and avoid browser-only constraints.

  • High-volume desktop transcription

    Use the desktop app for recurring long-form audio work such as podcasts, interviews, lectures, meetings, and large audio archives.

  • Text-to-speech generation

    Generate speech from written text with the text-to-speech tool using OpenAI voices and supported audio outputs.

Pros and Cons

Pros

  • Supports both web transcription and a local desktop workflow.
  • Includes subtitle output and free transcript utility tools.
  • Desktop app offers unlimited local transcriptions and no local file size or duration ceiling.
  • Supports Windows and macOS, including Intel and Apple Silicon Macs.
  • API key is stored locally in the browser, according to the FAQ.

Cons

  • Speech-to-text in the web app depends on a working OpenAI API key.
  • Cloud transcription uses upload limits on the web workflow, including a 25 MB file limit on the home page.

FAQ

Is WhisperUI free to use?

WhisperUI is free to use with some basic features, but you need a working OpenAI API key to use the app and pay OpenAI directly for usage tied to that key.

What premium features are available?

The premium features listed on the home page are multiple-file uploads, unlimited daily file uploads, and transforming audio files into SRT files.

Is my API key stored on WhisperUI servers?

WhisperUI stores your API key locally in your browser.

What file formats does WhisperUI support?

WhisperUI supports MP3, MP4, MPEG, MPGA, M4A, WAV, OGG, and WEBM for speech-to-text, and its text-to-speech tool supports MP3, AAC, and FLAC output.

What platforms does WhisperUI Desktop support?

WhisperUI Desktop is supported on Windows 10 and 11 and on macOS for both Intel and Apple Silicon, with a minimum of 4GB of RAM.

Quick Facts

Category
Speech to text / transcription
Platform
Web and desktop
Primary users
Creators, researchers, students, and teams
Source domain
whisperui.com
Workflow
Upload audio, transcribe with OpenAI Whisper, and review text or subtitles
Desktop support
Windows 10/11 and macOS (Intel and Apple Silicon)