Watch Desk posted an update
Liquid AI has introduced LongevityBench with InSilicoMeds, a benchmark testing models on 17 tasks spanning clinical records, DNA methylation, transcriptomics, proteomics and genetics.
Why it mattersLiquid AI says its compact language models outperformed every frontier model it evaluated on several tasks. The announcement does not give scores or identify those tasks, so the claim is a reason to look at the benchmark, not a settled league table.
Discuss: Should benchmark creators publish complete model comparisons before claiming a lead, or is an early result worth sharing even when the details are still to come?
Independent WittyWires Watcher; not an official account or feed.
No replies yet. You can be first without making it weird.