Agent-focused repos
Source control features are designed for agents, with lightweight push and pull operations and no rate limits called out on the site.
Relace is an AI model and infrastructure platform for coding agents, focused on code retrieval, file merging, and source-control workflows. It offers hosted, on-premise, and VPC-isolated deployment options with token-based pricing for individual models.
Relace is an AI model and infrastructure platform for coding agents. It focuses on the parts of the development workflow that need fast code retrieval, reliable file edits, and source-control operations designed for automated systems.
The site positions Relace as a set of purpose-built models and supporting infrastructure rather than a general-purpose chat product. Its hosted API supports quick experimentation, while enterprise and self-hosted deployments are available for teams with stricter compliance or latency requirements.
Source control features are designed for agents, with lightweight push and pull operations and no rate limits called out on the site.
The product includes semantic code retrieval that is described as scaling to large codebases and searching a codebase in under a second.
A universal merging model applies file edits at roughly 10,000 tokens per second, which the site presents as a fast path for code changes.
The models are described as small, fast SLMs trained in-house for utility tasks such as retrieval, merging, and code generation.
Infrastructure supports sandbox workflows with lightweight push and pull, fast branching for subagents, automatic indexing for two-stage retrieval, and throughput-oriented rate limits.
Use Relace when you need an agent to find relevant code across a large repository before it makes a change. The site highlights semantic search and code retrieval as core capabilities.
Use the fast apply model when your workflow needs file edits merged back into code quickly and reliably, such as automated fixes or repeated patching.
Use the infrastructure layer when you are running subagents or sandboxed workflows that need lightweight repository operations, branching, and indexing.
Use the hosted API to try the product quickly before committing to a more controlled deployment model. The site says hosted users can start experimenting within minutes.
Use the on-premise or VPC-isolated options when code residency, compliance, or latency constraints matter more than a standard hosted setup.
Relace is built for coding workflows that need code retrieval, merging, and code generation. The site positions it for teams building coding agents and engineering pipelines that benefit from specialized models rather than general-purpose LLMs.
Yes. The pricing page lists individual models priced by token usage, including `relace-apply-3`, `relace-search`, `relace-rank`, and `relace-embed`. The site also says you can start experimenting with the hosted API and contact the team for enterprise or self-hosted setups.
The site says Relace offers on-premise and VPC-isolated deployments for teams with stricter compliance or latency requirements. It also states that self-hosted or VPC deployments keep code inside your controlled environment.
Relace says its systems are SOC 2 compliant and that data is encrypted in transit and at rest on the hosted tier. For self-hosted and VPC deployments, it says code never leaves your controlled environment.
The site says hosted users can start experimenting within minutes through the API, while enterprise and self-hosted setups get guided onboarding from the Relace team.