Highlights
Popular repositories Loading
-
-
trmlu-audit
trmlu-audit PublicAPI-only cross-lingual contamination detection tooling for TR-MMLU (Turkish MMLU benchmark for LLM evaluation)
Python
-
tr-contamination-audit
tr-contamination-audit PublicMulti-benchmark contamination audit of Turkish LLM benchmarks (M1/M2/M3 probes, black-box API)
Python
-
agent-skills
agent-skills PublicForked from moonlight-lupin/agent-skills
A collection of AI agent skills for Hermes Agent — research, creative, productivity, devops, and more. Each skill is self-contained and tested.
Python
-
imla-tr-orthography-benchmark
imla-tr-orthography-benchmark PublicImla - rule-verifiable Turkish orthography benchmark for LLM evaluation (TDK-grounded caret probe; free-tier audit finds 61.1% caret drop). Pre-registered, zero-cost, paper + executable record.
TeX
-
tr-mmlu-ecosystem-audit
tr-mmlu-ecosystem-audit PublicPre-registered census + provenance audit of the Turkish MMLU benchmark ecosystem (91-repo name-collision census, gating/license findings, 29-node provenance DAG) - paper + full evidence chain.
TeX
If the problem persists, check the GitHub status page or contact support.