# SVBench AI

- **Event:** [Built with Claude: Life Sciences](https://cerebralvalley.ai/e/built-with-claude-life-sciences)
- **When:** Jul 7 at 12:00 PM – Jul 14 at 12:00 AM (EDT)
- **Where:** Online
- **Team:** [Ramanandan Prabhakaran](https://cerebralvalley.ai/u/Ram_Lifescience)
- **GitHub:** https://github.com/Ramanandan/svbench-ai ,
- **Demo video:** https://youtu.be/b02PkH2w4I4
- **Gallery:** https://cerebralvalley.ai/e/built-with-claude-life-sciences/hackathon/gallery
- **Page:** https://cerebralvalley.ai/e/built-with-claude-life-sciences/hackathon/gallery/208

SVBench AI is a Claude verdict layer for structural-variant callers. One reviewer — Claude Opus 4.8 — reads each call's alignment image and genomic context and returns an evidence-cited verdict, with or without a truth set, corroborated against independent population catalogs (gnomAD-SV, dbVar, HGSVC). Against HG002 it showed 93% of Sniffles' and 71% of SVIM's "false positives" are benchmark artifacts, not caller errors (raw 0.886→0.992; 0.694→0.886) — and SVIM's corrected precision equals Sniffles' raw, exposing a gap the benchmark hid. On genomes with no truth, the same reviewer still delivers grounded verdicts. Trustworthy SV evaluation, everywhere.

---

Markdown version of https://cerebralvalley.ai/e/built-with-claude-life-sciences/hackathon/gallery/208. Site index for agents: https://cerebralvalley.ai/llms.txt · full text: https://cerebralvalley.ai/llms-full.txt
