Can small local models query the databases?
Open-weight models — given nothing but the project's published skill files — answer 48 research questions across six task categories, from simple counts to multi-step analytical queries. Their SQL runs verbatim against the live databases; every answer is scored against pre-registered references. Includes the full question-by-question record with every model's SQL, and the skill-hardening feedback loop the first round produced.
Read the report