Skip to content
goard
Go to anything
Open any GitHub user, repo or page
You're not signed in, so some data is limited. Sign in with GitHub for full access.

Production LLM evaluation: golden datasets, model-as-judge scoring, CI/CD regression gates

Python 7 KB Created Apr 22, 2026
Stars 0 Forks 0 Watchers 0 Open issues 0 Open PRs 0
Contributors
Stargazers
Forks
Watchers