Recent models, without an invented ranking.
A searchable working record of release claims, outside observations, direct health evidence, policy behavior, access, cost, and unresolved questions.
← Health-AI WatchFind the model. Inspect what is actually known.
Reverse-chronological cards for major models a health researcher could plausibly use. Search, filter, or compare up to three without collapsing unlike evidence into one score.
It does not infer a health rank, treat evidence maturity as capability, or equate a base model with a consumer product. API price is not total benchmark cost, and an open-weight release still has hardware and deployment costs. The ledger identifies candidates and evidence gaps; it does not establish clinical readiness.
A release record chooses the candidate. It does not answer the question.
The main Watch asks whether selected systems can conduct a rigorous, reproducible health study—and records result fidelity, method, calibration, audit recovery, human burden, and cost.