Preprints.ai
Evidence wall

What we measure, and what we don't

Every credibility claim on preprints.ai gets its own page — with a denominator, a sample size, a source-code link, and a list of things it does not measure. No marketing claims, no cherry-picked screenshots.

When we don't yet have data, the page should say so plainly. External benchmarks are attributed and separated from production performance. This wall documents screening modules and their failure modes; it is not evidence that the overall system performs peer review.

01

Layer 1 audit modules

02

Model screening panel

03

External grounding

04

Quality measurement

05

Case studies