Cloud model exploration and execution
Browse a catalog of public and official models, including image, video, text-to-speech, and reasoning models, then run them in the cloud without managing infrastructure yourself.
Replicate is a cloud API for running machine learning models, discovering public and official models, and deploying custom models with Cog. It is aimed at developers and AI teams that need model execution, billing, and production workflow support.
Replicate is a cloud platform for running machine learning models through an API. The site positions it as a place to discover public and official models, run them in the cloud, and deploy your own custom models when you need a private workflow.
The product is aimed at developers and teams building AI-powered applications or experiments. Its docs cover model execution, custom model packaging with Cog, predictions, deployments, webhooks, billing, rate limits, and related operational topics, which suggests it is designed for both prototyping and production use.
Browse a catalog of public and official models, including image, video, text-to-speech, and reasoning models, then run them in the cloud without managing infrastructure yourself.
Use the pricing model that matches the workload: many public models bill by runtime, while some bill by input and output, which makes costs visible on each model page.
Package and deploy your own custom models with Cog, Replicate’s open-source tool for model packaging and deployment.
Work through documented APIs and SDK paths, including HTTP API, client libraries, Node.js, Python, Google Colab, and an OpenAPI schema.
Use deployment features, webhooks, secrets, and rate limits to support production workflows and automated model handling.
Train or fine-tune models and manage the lifecycle through documented model, prediction, and deployment pages.
Search the model catalog, pick a public or official model, and run it from the cloud for tasks like image generation, video generation, text-to-speech, or reasoning.
Package a model with Cog and deploy it as a private model when you need your own model code or weights rather than a public listing.
Build an application against the HTTP API or a client library, using documented workflows for Node.js, Python, or Google Colab.
Use webhooks, secrets, rate limits, and deployments to connect model execution to a production system with automation and monitoring.
Replicate is a cloud API for running machine learning models and related workflows. The docs show you can run models from Node.js, Python, Google Colab, and other integrations, and the explore page focuses on models you can run in the cloud.
The pricing page says you only pay for what you use. Public models are billed either by the time they take to run or by input and output, while private models generally run on dedicated hardware and are billed for the time instances are online, with a special case for fast-booting fine-tunes.
The docs list several ways to work with the platform, including HTTP API access, client libraries, an OpenAPI schema, MCP server support, webhooks, and agent skills. They also document model creation, predictions, deployments, secrets, and rate limits.
The explore page highlights public and official models for image generation, video generation, text-to-speech, coding, and other model categories. The docs and pricing pages also show that users can deploy custom models with Cog.
The source confirms documentation for model running, fine-tuning, deployments, webhooks, billing, and security topics, but it does not provide a complete list of every supported framework, cloud provider, or runtime from the collected pages.
Los datos de tráfico son solo como referencia.
| may | 1255416 |
|---|---|
| jun | 1218701 |
| jul | 1079703 |
Las analíticas de tráfico aún no están disponibles.
Las analíticas de tráfico aún no están disponibles.