AIオーディオAPI

AIオーディオAPIで、音声認識や音声合成、ボイスクローニング、音声分析、リアルタイム処理をアプリに組み込めます。

AIインフラ

このコレクションを探索

製品

Daily preview
Daily logo

Daily

AIオーディオAPI

Web・モバイル・ネイティブ・デスクトップ・サーバー向けのリアルタイム通話、録画、AIエージェント用SDKと基盤

Speechmatics preview
Speechmatics logo

Speechmatics

AIオーディオAPI

Speechmatics provides speech-to-text APIs for pre-recorded and real-time audio, with support for multilingual speech, multiple speakers, and voice-agent workflows. It is designed for teams building transcription and Voice AI applications.

Loudly preview
Loudly logo

Loudly

AI音楽生成

音楽の生成、カスタマイズ、リミックス、配信を一元化するAI音楽プラットフォーム

ModelsLab preview
ModelsLab logo

ModelsLab

AIオーディオAPI

ModelsLab is a developer platform that provides APIs for image, video, audio, 3D, and LLM generation through a unified service. It supports teams building generative media features without integrating each model provider separately.

AudioStack logo
AudioStack logo

AudioStack

AIオーディオAPI

メディア・広告向けAI音声制作。ブリーフや素材から放送品質の音声を作成

SpeechifyAI preview
SpeechifyAI logo

SpeechifyAI

AIオーディオAPI

SpeechifyAI is a developer API for expressive text-to-speech and voice cloning. It provides streaming Simba models, catalog and cloned voices, SSML support, and a free starting tier for building speech into applications.

Voicemaker preview
Voicemaker logo

Voicemaker

AI音声生成ツール

多言語AI音声の生成・編集・ダウンロード・API自動化に対応するブラウザ型TTSプラットフォーム。

Apiframe preview
Apiframe logo

Apiframe

AIオーディオAPI

Apiframe is a unified REST API for generating AI images, videos, and music. It helps developers and automation teams add media generation through one API, with asynchronous jobs, webhooks, SDKs, and CDN-hosted outputs.

API in One preview
API in One logo

API in One

AIオーディオAPI

API in One is a unified AI API gateway for website owners and developers who need image, video, music, speech, chat, and AI tool capabilities through one API key and shared credit balance.

4ALL API preview
4ALL API logo

4ALL API

AIオーディオAPI

4ALL API is an API aggregation gateway for enterprises and developers that provides one access layer for text, image, video, and audio models from multiple providers. It supports OpenAI-compatible access, usage-based billing, model routing, and failover workflows.

Recall.ai Startup Program preview
Recall.ai Startup Program logo

Recall.ai Startup Program

AIオーディオAPI

A startup program for early-stage companies building products powered by meeting and conversation data. Approved applicants receive discounted recording usage, access to Recall.ai products, and support while they build and launch.

Scaleway Generative APIs preview
Scaleway Generative APIs logo

Scaleway Generative APIs

AIオーディオAPI

Scaleway Generative APIs provide OpenAI-compatible, serverless access to chat, code, vision, embedding, and audio models. They are designed for developers building AI applications without managing model-serving hardware, with endpoints hosted in European data centers and usage billed by tokens or audio minutes.

Gemini 3.8 text-to-speech preview
Gemini 3.8 text-to-speech logo

Gemini 3.8 text-to-speech

AIオーディオAPI

Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS are audio-generation models for creating expressive voices, directing spoken performances, and producing conversational audio. They support creative teams, developers, and enterprises working across Google AI Studio, Gemini API, Gemini Enterprise, Gemini Notebook, and Google Vids.

Runway Dev preview
Runway Dev logo

Runway Dev

AIオーディオAPI

Runway Dev is an API platform for adding AI-generated video, images, and audio to products and production workflows. It provides access to Runway and other providers' models, along with model routing, workflows, recipes, and team controls.

Stability AI Developer Platform preview
Stability AI Developer Platform logo

Stability AI Developer Platform

AI 3Dモデル生成

Stability AI Developer Platform provides APIs for adding generative audio, image, image editing, upscaling, control, and 3D asset creation to applications. It is intended for developers and teams building creative workflows with Stability AI models.