DigestAI news desk

AI news, digested. Every story with its sources, every hour.

Research

SAGE system raises grant review agreement to kappa 0.58, beating baseline

Researchers introduced SAGE, a schema‑guided LLM that converts grant rubrics into structured checks and links each judgement to evidence in the application package. In an initial evaluation on 35 nonprofit grant applications, SAGE’s assessments were compared with 105 original competition reviews, yielding a fair ordinal agreement of kappa = 0.29.

1 source primary source

Key points

  • SAGE evaluated on 35 nonprofit grant applications, achieving kappa 0.29 in initial test.
  • In a re‑review of 202 assessments, SAGE reached kappa 0.58, surpassing the baseline kappa 0.33.
  • The system links rubric criteria to evidence, producing auditable drafts for expert correction.

After the foundation inspected SAGE’s output, it conducted a criterion‑level re‑review that produced 202 assessments. In this assisted round, SAGE improved to kappa = 0.58 and outperformed a one‑prompt‑per‑criterion baseline, which recorded kappa = 0.33 on the same subset. The system also showed higher rank correlation and lower error rates. A claim‑level audit identified which parts of the structured draft were confirmed, disputed, or left unaddressed, demonstrating SAGE’s ability to generate detailed, evidence‑linked, and auditable drafts for expert correction.

Read the original atarXiv cs.CL · by Erik Varapaev, Andrei Chetvergov, Stepan Ukolov, Timofei Sivoraksha, Alexander Evseev, Sergey Bolovtsov primary sourceOpen source ↗

The headline, key points and digest above were generated by Digest AI's editorial model from the linked sources. Automated summaries can contain errors: the sources are the record. Spotted a mistake? Tell us.

Comments

via GitHub Discussions

More in Research

All →

Related stories