实验跟踪
跟踪机器学习实验和模型运行,并在 W&B 平台中查看以便进行对比和分析。
Weights & Biases 是面向机器学习与 AI 应用团队的开发工具。它提供一个平台,用于跟踪实验、管理模型和数据集工件,并在开发和生产工作流中观察 GenAI 系统。
该产品覆盖多种部署方式,包括托管 SaaS、专属云和客户自管选项。价格页面还展示了面向个人、团队和企业客户的不同路径,并在更高等级中提供付费支持和安全功能。
跟踪机器学习实验和模型运行,并在 W&B 平台中查看以便进行对比和分析。
对数据集和模型进行版本管理,管理工件,并在模型生命周期中跟踪血缘关系。
构建协作式仪表板和报告,与团队成员分享结果和进展。
记录 GenAI 跟踪、输入、输出和元数据,用于开发过程中的评估与生产监控。
运行评估,将 recipes 并排比较,并检查 LLM 应用的准确率、延迟和 token 使用情况。
在受支持的套餐中使用访问控制、基于团队的权限、服务账户、审计日志和 SSO。
在迭代模型和实验时跟踪运行、比较结果,并将工件整理归档。
为 LLM 应用和智能体工作流记录跟踪、评估和生产监控信号。
共享仪表板、报告和工件,方便研究人员和工程师一起审查结果。
当安全、隔离或数据驻留要求很重要时,可部署到 SaaS、专属或客户自管环境中。
使用个人、学术或试用路径,在进行更广泛推广前先评估平台。
Weights & Biases 按所选方案按月或按年预付计费 Pro 套餐。若在周期中途新增 model seats,则按比例计费;Weave 数据摄取、W&B Inference 和存储则按月后付,依据用量计费。Enterprise 套餐按年预付开票。
价格页面说明,免费套餐面向 AI 应用和模型的个人开发而设计,同时还为研究和本地使用提供个人和学术选项。Personal 套餐不允许企业用途。
Enterprise 套餐增加单租户或客户自管部署选项,以及安全控制功能,例如 SSO、自动化用户配置、审计日志和客户自管加密密钥支持,具体取决于部署选项。
支持页面列出了 Standard、Standard Plus 和 Premium 三种支持套餐。Standard Plus 增加了入门指导、专属成功团队,以及 Slack 或 Microsoft Teams 支持;Premium 则增加更快的响应时间、24/7 覆盖,以及功能优先级和路线图会议等协作功能。
部署选项页面介绍了 SaaS Cloud、Dedicated Cloud 和 Customer-Managed 部署。W&B 还在价格页面中记录了用于个人使用的本地托管服务器选项,以及可在你自己的基础设施上运行服务器的 Enterprise 试用。
流量数据仅供参考。
| 5月 | 2475536 |
|---|---|
| 6月 | 2045155 |
| 7月 | 2083772 |
暂无流量分析数据。
暂无流量分析数据。
www.lyzr.ai
OpenController is Lyzr’s control plane for discovering, evaluating, governing, and monitoring AI agents, models, tools, data, and workflows across an enterprise AI estate. It is intended for teams managing agents across clouds, frameworks, runtimes, and environments.
www.langchain.com
LangSmith is an observability and evaluation platform for AI agents and LLM applications. It helps development and production teams trace agent behavior, monitor quality and cost, investigate failures, and evaluate changes.
www.galileo.ai
Galileo is an AI observability and evaluation platform for testing, debugging, and governing LLM and agent systems across development and production. It helps teams turn evaluation results into production guardrails and monitor AI behavior at scale.
www.tensorzero.com
TensorZero is an open-source LLMOps platform for building and operating production-grade LLM applications. Its stated scope combines an LLM gateway with observability, evaluation, optimization, and experimentation tools.
arize.com
Phoenix is an open-source, local-first platform for tracing, evaluating, experimenting with, and improving AI applications and agents. It helps AI engineers inspect agent behavior, assess output quality, and test changes before deployment.
wandb.ai
W&B Weave is an observability and evaluation platform for production AI agents and applications. It helps teams trace agent behavior, evaluate changes, inspect prompts and models, and monitor production interactions.