Commitment Checklist: Auditing Author Commitments in Peer Review
2026-03-04T19:47:20Z•7259b41a8081493525a2e473c8ff69a7bb8b5a5b48aba55da39c9faa93522260
Commitment ChecklistEU AI ActLLM evaluationMOSAICPaperReproVLMsauthor commitmentsautomated reproducibilitychatbotscustomizationeducation technologyenvironmental impactethicsgreen AIlarge language modelsmeasurement sciencemental modelsmoral benchmarkspeer review auditingregulationreproducibilityrisk classificationsurvey synthesissynthetic data
What happened
Collection of newly announced arXiv papers (Mar 3, 2026) covering AI evaluation, ethics, reproducibility, education, governance, and environmental impacts. Key contributions include: a large-scale LLM-driven audit of author commitments in peer review ("Commitment Checklist") finding ~25% of commitments unfulfilled; MOSAIC, a broad benchmark assessing moral, social and individual dimensions of LLMs; PaperRepro, a two-stage multi-agent system that improves automated computational reproducibility (21.9% relative improvement on REPRO-Bench); a measurement-science framing for AI capabilities as "d⇢
Why it matters
A reviewed impact interpretation has not been published for this record.
Evidence and limitations
- Source ID
- arxiv_cs_cy
- Record identifier
- 7259b41a8081493525a2e473c8ff69a7bb8b5a5b48aba55da39c9faa93522260
- Enrichment time
- 2026-03-04T19:47:20Z
- AI-assisted enrichment
- Yes
This record may overlap with other records. Its enrichment can be incomplete or wrong, and machine assistance was used. Validate consequential decisions against the linked source and your own environment.