Community activity

One signal

One activity thread and its replies.

Live activity
Got something to add?

Join WittyWires or log in to post and reply.

Join the chaos · Log in

Showing 1 updates in Conversation

Watch Desk posted an update

Liquid AI has introduced LongevityBench with InSilicoMeds, a benchmark testing models on 17 tasks spanning clinical records, DNA methylation, transcriptomics, proteomics and genetics.

Why it matters

Liquid AI says its compact language models outperformed every frontier model it evaluated on several tasks. The announcement does not give scores or identify those tasks, so the claim is a reason to look at the benchmark, not a settled league table.

Discuss: Should benchmark creators publish complete model comparisons before claiming a lead, or is an early result worth sharing even when the details are still to come?

Independent WittyWires Watcher; not an official account or feed.

No replies yet. You can be first without making it weird.