Research Reliability
A publication count is not an accuracy score. Evidence reference checks and semantic support evaluation answer different questions. See how Research works.
Snapshot updated .
Continuous Eval V1 has no reconciled public results yet. The legacy inventory below does not measure the new GPT-6 pipeline.
Legacy RI1
Methodology: legacy-ri1. Current published inventory; no processing-window denominator available.
Comparable model-profile telemetry is unavailable for this inventory.
| Outcome / stage | Source versions |
|---|---|
| Source versions discovered | Unavailable |
| Source eligible | Unavailable |
| Acquired | Unavailable |
| Normalized | Unavailable |
| Classification admitted | Unavailable |
| Synthesis attempted | Unavailable |
| Semantically evaluated | Unavailable |
| Publication eligible | Unavailable |
| Published | 10 |
| Quarantined | Unavailable |
| Abstained | Unavailable |
| Capacity deferred | Unavailable |
| Budget deferred | Unavailable |
| Infrastructure failed | Unavailable |
| Dimension | Observed / assessed |
|---|---|
| Evidence sufficiency | Unavailable |
| Claim coverage | Unavailable |
| Semantic support | Unavailable |
| Scope correctness | Unavailable |
| Wrong lens | Unavailable |
| Quantitative validation | Unavailable |
| Limitations preserved | Unavailable |
| Citation and provenance integrity | Unavailable |
| Normalization dependency | Unavailable |
| Material omission | Unavailable |
| Section fit | Unavailable |
| Publication usefulness | Unavailable |
Observed failure categories
No comparable failure-category data is available in this snapshot.
Usage and latency
Comparable model usage and cost are unavailable.
Unavailable means the retained data cannot establish that measurement. Cohorts are not combined into a headline pass rate. Counts describe automated processing outcomes and do not establish human scientific approval.