Insilico Publishes LongevityBench and Compact Longevity LLMs in Cell
Insilico Medicine published LongevityBench in Cell on 17 September 2026, releasing 17 aging biology tasks, five compact Longevity LLMs, and an open Longevity Claw research interface.
What changed
Insilico Medicine announced on 17 September 2026 the publication of An open benchmark and language models for AI in aging biology in Cell (DOI 10.1016/j.cell.2026.08.026). The study introduces LongevityBench, a suite of 17 tasks spanning five biodata domains including DNA methylation, transcriptomic, proteomic, and clinical modalities, and evaluates 18 frontier AI systems from six developer teams. No single frontier model dominates all tasks; omics based age prediction remains the hardest category regardless of scale according to the paper abstract.
The team fine tuned five multitask Longevity LLMs ranging from 0.6B to 9B parameters on domain specific aging data. Those compact models matched or exceeded far larger frontier systems on LongevityBench and some purpose built machine learning baselines. Insilico released benchmark datasets, trained models on Hugging Face, evaluation code, and Longevity Claw, an agentic research interface pairing the models with domain tools. A public leaderboard is live at longevitybenchmarks.org.

Why it matters
Pharma and longevity investors have been flooded with general purpose model demos that do not survive structured omics tasks. LongevityBench gives R and D leaders a reproducible scoreboard: which systems actually handle sample comparison, classification, and conditional molecular profile generation on real aging datasets. The compact Longevity LLM result challenges the assumption that frontier scale compute is mandatory for domain structured biology workflows.
Because artifacts are open, diligence teams can rerun evaluations instead of trusting vendor slide decks alone.
Who is affected
Aging biology researchers, computational omics teams in pharma, AI platform buyers supporting biotech R and D, and investors separating benchmark backed longevity AI from narrative only partnerships.
What to do next
Run LongevityBench on the models you already license before signing new longevity AI partnerships, and prioritize vendors that publish task level failure modes not aggregate accuracy alone.

What to watch
Independent groups publishing LongevityBench leaderboard entries for models not in Insilico's initial 18 system comparison, and whether pharma partners adopt Longevity Claw in live target triage workflows with disclosed validation outcomes.
Sources
- Primary. Cell, An open benchmark and language models for AI in aging biology00999-2) (17 September 2026). LongevityBench task design, frontier model evaluation, and Longevity LLM results.
- Primary. Insilico Medicine, AI Longevity Discovery Toolkit announcement (17 September 2026). Release details for Longevity Claw, Hugging Face collections, and leaderboard URL.
- Secondary. PR Newswire, Insilico Cell cover study release (17 September 2026). Confirms open release scope and toolkit components.