LLM Evaluation
发现大语言模型评测工具、基准测试与平台,衡量模型质量、对比输出、识别风险并持续改进 AI 应用。
AI 基础设施
探索此分类集合产品

QAgent
AI Testing Assistant
QAgent is an AI agent testing and quality assurance platform for developers and agile teams. It connects to an agent through a webhook or endpoint, runs automated test cases, and evaluates responses for groundedness, policy adherence, prompt compliance, and related quality dimensions.

inferock-bench
AI Gateway And Routing
inferock-bench is a local diagnostic proxy for tracking LLM API usage, provider-reported costs, failures, and billing-integrity signals. It helps developers inspect calls to OpenAI, Anthropic, Gemini Developer API, and pinned OpenRouter endpoints using locally stored receipts.





