Uh oh!
There was an error while loading. Please reload this page.
Actions: agentic-workflow-kit/eval-kit
Actions
Showing runs from all workflows
39 workflow runs
39 workflow runs
feat: harden pointwise model-judge reporting (#17)
check
#39:
Commit 5c72926
pushed
by
aryeko
feat: harden pointwise model-judge reporting
check
#38:
Pull request #17
opened
by
aryeko
fix: ignore bootstrapped eval result bundles (#15)
check
#35:
Commit 1b986b5
pushed
by
aryeko
fix: ignore bootstrapped eval result bundles
check
#34:
Pull request #15
opened
by
aryeko
docs: retire stale model-judge guidance
check
#30:
Pull request #13
opened
by
aryeko
docs: standardize model-judge calibration reporting (#12)
check
#29:
Commit 272bad8
pushed
by
aryeko
docs: standardize model-judge calibration reporting
check
#28:
Pull request #12
opened
by
aryeko
docs: align model-assisted examples with fail-closed configs (#11)
check
#27:
Commit 2616ad1
pushed
by
aryeko
docs: align model-assisted examples with fail-closed configs
check
#26:
Pull request #11
opened
by
aryeko
chore(release): v0.1.4
check
#24:
Pull request #10
opened
by
aryeko
chore(release): v0.1.3
check
#22:
Pull request #9
synchronize
by
aryeko
fix(cli): enforce disabled eval methods
check
#19:
Pull request #8
opened
by
aryeko
docs: clarify eval policy quickstart
check
#17:
Pull request #7
opened
by
aryeko
docs: clarify eval-kit help and adapter contract (#6)
check
#16:
Commit 2d3bf2d
pushed
by
aryeko
docs: clarify eval-kit help and adapter contract
check
#15:
Pull request #6
opened
by
aryeko