LLM gateway and request routing
Route OpenAI-style calls through a single gateway, or use passthrough endpoints for provider-native SDKs, while every request is logged.
Respan is an LLM engineering platform for observability, evals, prompt management, and gateway routing. Monitor production traffic and control reliability in one workflow.
Respan is an LLM engineering platform for teams that want observability, evaluation, prompt management, and gateway routing in one place. The site positions it as a system for routing, observing, evaluating, optimizing, and deploying LLM traffic from a single workflow.
The product centers on production LLM operations: traces for each call, dashboards for usage and cost, evaluators for quality checks, and gateway controls for retries, fallbacks, caching, and rate limits. It is aimed at teams shipping AI agents and other LLM-powered features that need to track behavior and reduce production issues.
Route OpenAI-style calls through a single gateway, or use passthrough endpoints for provider-native SDKs, while every request is logged.
Inspect calls in traces, view latency on spans, track logs, and filter by customer identifiers or metadata to understand production behavior.
Monitor requests, tokens, errors, latency, and cost in dashboards, with slices by model or user and alerts when thresholds are crossed.
Create evaluation workflows that combine rule checks, LLM judges, human review, datasets, and experiments across prompt and model variants.
Manage prompts with version control, collaborative editing, one-click deployment, release management, and a model playground.
Handle retries, fallback models, load balancing, caching, key limits, and spend controls from the gateway and settings.
Route model traffic through one endpoint, log each request, and use fallbacks or retries when a provider errors or rate-limits.
Watch traces, latency, token usage, and cost in dashboards so production regressions surface quickly.
Build datasets from traces or CSVs, then run experiments across prompt and model variants before merging changes.
Combine rule checks, LLM judges, and human review in a single workflow to score real production traffic and offline samples.
Use prompts, versions, rollout logic, and access to multiple models to promote workflows from the UI into production.
Respan is presented as an LLM engineering platform that combines observability, evals, prompt optimization, and an LLM gateway in one system.
The source shows a Free plan, a Team plan, and an Enterprise plan. The Team plan is listed at $199 per month billed yearly, while Enterprise is contact sales.
Yes. The pricing page lists a free tier with logs, scores, datasets, evaluators, and prompts, plus a Team plan that adds unlimited datasets, evaluators, and prompts.
The product includes an annotation of OpenTelemetry support, but the source does not provide detailed integration documentation or a full provider list.
The homepage and pricing page emphasize tracing, dashboards, evaluation workflows, prompt management, and gateway features such as fallbacks and caching.