项目介绍
Open-source, end-to-end platform for evaluating, observing, and improving LLM and AI agent applications. Tracing · Evals · Simulations · Datasets · Gateway · Guardrails. Self-hostable. Apache 2.0.
适合用来做什么
评估前沿 AI 能力
阅读源码与工程实践
用于技术选型调研
为什么值得关注
该项目在 GitHub 保持活跃,并在 AI 工程相关主题中获得持续关注。