Building Evalorium · Open to AI Quality & Evaluation roles
Enterprise AI Quality Engineering · LLM & Agent Evaluation · Release Gates · Risk Monitoring · Eval-to-RL · Building Evalorium
-
Independent · Building Evalorium
- https://github.com/plwslpld-arch/evalorium
Pinned Loading
-
eval-harness-internals
eval-harness-internals Public面向开发者的中文 Eval Harness 源码教材:解析 lm-evaluation-harness、Inspect AI、OpenAI Evals、Promptfoo、DeepEval 与 Harbor,覆盖任务、运行、评分、统计与发布门禁。
Python 5
-
agent-harness-internals
agent-harness-internals Public面向开发者的中文 Agent Harness 源码教材:解析 Codex、Claude、Gemini CLI、DeepSeek Harness、pi 与 OpenCode,覆盖配置、工具循环、权限边界、状态恢复与编排。
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.
