Commitment Checklist: Auditing Author Commitments in Peer Review

2026-03-04T19:47:20Z7259b41a8081493525a2e473c8ff69a7bb8b5a5b48aba55da39c9faa93522260
Commitment ChecklistEU AI ActLLM evaluationMOSAICPaperReproVLMsauthor commitmentsautomated reproducibilitychatbotscustomizationeducation technologyenvironmental impactethicsgreen AIlarge language modelsmeasurement sciencemental modelsmoral benchmarkspeer review auditingregulationreproducibilityrisk classificationsurvey synthesissynthetic data

What happened

Collection of newly announced arXiv papers (Mar 3, 2026) covering AI evaluation, ethics, reproducibility, education, governance, and environmental impacts. Key contributions include: a large-scale LLM-driven audit of author commitments in peer review ("Commitment Checklist") finding ~25% of commitments unfulfilled; MOSAIC, a broad benchmark assessing moral, social and individual dimensions of LLMs; PaperRepro, a two-stage multi-agent system that improves automated computational reproducibility (21.9% relative improvement on REPRO-Bench); a measurement-science framing for AI capabilities as "d⇢

Why it matters

A reviewed impact interpretation has not been published for this record.

Evidence and limitations

Source ID
arxiv_cs_cy
Record identifier
7259b41a8081493525a2e473c8ff69a7bb8b5a5b48aba55da39c9faa93522260
Enrichment time
2026-03-04T19:47:20Z
AI-assisted enrichment
Yes

This record may overlap with other records. Its enrichment can be incomplete or wrong, and machine assistance was used. Validate consequential decisions against the linked source and your own environment.

Record · Commitment Checklist: Auditing Author Commitments in Peer Review · Baitaphish