Migrating CompileBench to Harbor: standardizing AI agent evals

Migrating CompileBench to Harbor: standardizing AI agent evals 图片 1

Standardizing AI agent evaluation with Harbor: an open-source framework for reproducible benchmarks, reinforcement learning, and collaborative evals.

添加评论
点赞收藏
点踩分享查看原文
评论
?
参与讨论