Summa summarum, scientia !
Built at Built with Opus 4.6: a Claude Code hackathon · Feb 10, 2026
With AI being increasingly deployed for scientific tasks, we are seeing a surge in scientific hypotheses formulated by AI. However, in clinical research, evaluating hypotheses is expensive with the limiting factors being constraints of the physical world. In software development we have tools for automatic reviews of pull requests that can make it very easy to triage code generated by AI (or humans) and makes it much easier to integrate AI into the coding workflow. Our idea here was that it would be great to have a similar kind of automatic review of scientific claims states in publications, investor decks, or even free text. We used Opus 4.6 logical reasoning capabilities and developed a “Critical AI scientist” that searches and surfaces what is already known, and identifies and structures the right context to provide a nuanced view of a hypothesis across many axes. DEMO: https://critical-ai-scientist.jonas-walheim.workers.dev user: scicrit pwd: Xo8gN4DciL1Oont56IDH2mp3