Tag: model evaluation
-
Benchmark reality for bio foundation models: how data leakage, missing baselines and metric choice inflate “beats baseline” claims
Evidence-first notes on bioscience and deep tech, at the edge of the lab and the market. Information only — not investment advice. The 30-second version What. A leaderboard number for a bio foundation model (bio-FM) is the product of model ability and how forgiving the evaluation is. Data leakage is close to the default, not…
-
AI protein design’s validation reality — high in-silico scores, low wet-lab hit rates, and why an honest success-rate benchmark is still missing
Evidence-first notes on bioscience and deep tech, at the edge of the lab and the market. Information only — not investment advice, not medical advice. This is the bottleneck layer of the AI-protein-design series, where its central question resolves. Self-reported in-silico success rates and improvement factors (from original authors and vendors) are separated throughout from…