This is a tiny, runnable starter to build a 20–120 paragraph benchmark to test simple AI-text detectors.
- Fill
data/data.csvwith paragraphs (human + AI). - Install deps:
pip install -r requirements.txt - Run:
python src/eval.py - See:
results/metrics.csvandresults/summary.txt - Write your 1‑pager using
report/one_pager_template.md
Columns (keep these headers exactly): id,text,label,domain,paraphrase_level,language,source,notes
label:humanoraidomain:stem,policy,narrative(use these three to start)paraphrase_level:none,light,heavylanguage:enores->en(if you machine-translate to Spanish and back)source: e.g.,self,ku_website,wikipedia,chatgptnotes: short free text (e.g., prompt name, link slug)- Wrap the
textfield in double quotes. If your text contains quotes, double them.
A few sample rows are already in data/data.csv. Replace them with your own.