OriginBlame: Record- and Token-Level Data Provenance for AI Training Datasets

arXiv 2607.13037•cfd0b64daa81e607e783231c1afed555b9901bb8ad44e146b6f90ceff6fc008d
AI-insuranceAI-riskLLM-auditingOracle-Databaseagent-memoryagentic-AIbenchmarkingchain-of-thoughtdata-provenancegraph-neural-networkshuman-preferencesinterventional-auditsmachine-unlearningmodel-evaluationmolecular-predictionmulti-agent-systemsneuro-symbolic-AIprivacyprobabilistic-reasoningrecord-level-provenancerobotics-deploymentsafe-RLself-improving-agentssoftware-package

Paper metadata

arXiv ID
2607.13037
Version
Not specified by this published record
Category
Computer Science — Artificial Intelligence (cs.AI)

The PDF link points to arxiv.org. Baitaphish does not expose a private stored PDF.

Evidence and limitations

Source ID
arxiv_cs_ai
Record identifier
cfd0b64daa81e607e783231c1afed555b9901bb8ad44e146b6f90ceff6fc008d
Enrichment time
2026-07-16T08:52:18Z
AI-assisted enrichment
Yes

This record may overlap with other records. Its enrichment can be incomplete or wrong, and machine assistance was used. Validate consequential decisions against the linked source and your own environment.