LLMbench: A Comparative Close Reading Workbench for Large Language Models

arXiv 2604.15508•32c2cf24d37799949127b2ca6421c0d820f53928622ada810ef573a95d2a5f80
AI ethicsAI governanceLLM evaluationPolymarketUMAaccessibility as defenseanthropomorphic deceptiondark patternsdelayed ground truthdrift detectionhuman-AI interactionlarge language modelslog-probability analysismodel interpretabilityon-chain dispute resolutionprediction marketsprivacy & health AIprobability visualizationproxy monitoring

Paper metadata

arXiv ID
2604.15508
Version
Not specified by this published record
Category
Computer Science — Computers and Society (cs.CY)

The PDF link points to arxiv.org. Baitaphish does not expose a private stored PDF.

Evidence and limitations

Source ID
arxiv_cs_cy
Record identifier
32c2cf24d37799949127b2ca6421c0d820f53928622ada810ef573a95d2a5f80
Enrichment time
2026-04-20T07:23:52Z
AI-assisted enrichment
Yes

This record may overlap with other records. Its enrichment can be incomplete or wrong, and machine assistance was used. Validate consequential decisions against the linked source and your own environment.

LLMbench: A Comparative Close Reading Workbench for Large Language Models · Baitaphish