Tag: benchmark
-
Benchmark reality for bio foundation models: how data leakage, missing baselines and metric choice inflate “beats baseline” claims
Evidence-first notes on bioscience and deep tech, at the edge of the lab and the market. Information only — not investment advice. The 30-second version What. A leaderboard number for a bio foundation model (bio-FM) is the product of model ability and how forgiving the evaluation is. Data leakage is close to the default, not…
-
Have cell / virtual-cell foundation models beaten the linear baseline? An honest read (Part 3)
Evidence-first notes on bioscience and deep tech, at the edge of the lab and the market. Information only — not investment advice. The 30-second version What. This is the most skeptical axis of the bio-foundation-model series. As of 2026-07, cell / virtual-cell foundation models (scGPT, Geneformer, scFoundation, Arc STATE and others) do not yet consistently…
-
The landscape of biology foundation models: do they actually beat the old specialized baselines?
Evidence-first notes on bioscience and deep tech, at the edge of the lab and the market. Information only — not investment advice. The 30-second version What. Protein, genome and cell “foundation models” (bio-FMs) are usually sold with headlines like “simulating millions of years of evolution.” The firm’s one question is narrower: on a specific task,…
-
AI protein design’s validation reality — high in-silico scores, low wet-lab hit rates, and why an honest success-rate benchmark is still missing
Evidence-first notes on bioscience and deep tech, at the edge of the lab and the market. Information only — not investment advice, not medical advice. This is the bottleneck layer of the AI-protein-design series, where its central question resolves. Self-reported in-silico success rates and improvement factors (from original authors and vendors) are separated throughout from…