Skip to content

2026 modernization: reliable AI judge guardrails + security baseline - #8

Merged
duhfreakinduh merged 2 commits into
mainfrom
upgrade/2026-ai-maintenance
Aug 31, 2026
Merged

duhfreakinduh merged 2 commits into
mainfrom
upgrade/2026-ai-maintenance

Conversation

@duhfreakinduh

Copy link
Copy Markdown
Owner

Modernization pass around the existing reliability-first AI Judge architecture.

Changes:

  • add AGENTS.md requiring deterministic fallbacks, bounded loading states, validated model output, manual override, and private evidence handling
  • add SECURITY.md for uploaded evidence, participant data, model/provider tokens, and score/winner validation

Audit note: the repo already has Judge v3/v4 hotfix logic using @huggingface/transformers 4.2.0, a deterministic instant referee, a small MobileBERT classifier, hard timeouts, and safe fallback behavior, so this pass does not replace that working architecture.

@duhfreakinduh
duhfreakinduh merged commit 6d8cf4c into main Aug 31, 2026
2 checks passed
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant