Turning published AI-safety claims into runnable, independently reproducible test packages that expose fragile assumptions, hidden failure modes, and unsupported conclusions.
Turning published AI-safety claims into runnable, independently reproducible test packages that expose fragile assumptions, hidden failure modes, and unsupported conclusions.
Project Details
Updated 07/05/26 · Provided via application · VerifiedPublished AI-safety claims can be difficult to reproduce because code, test data, assumptions, and evaluation procedures are often incomplete/fragmented. I plan to select several safety claims and convert each into a runnable package. Each package will include code, synthetic test cases, validation checks, failure examples, documentation, and reproducibility results. I will conduct the work through DBbun as the sole principal researcher and publicly release the resulting tools where feasible.
Theory of Impact
Updated 07/05/26 · By grantmaking.aiAI decisions may depend on evaluations that have not been independently reproduced. Weak/misleading safety evidence can create false confidence. Executable reproduction allows researchers to test assumptions and detect fragile conclusions. Better verification improves the evidence available to labs/evaluators/policymakers. This does not solve alignment directly, but reduces one pathway by which unreliable safety claims could influence high-stakes deployment decisions.
People
Updated 07/05/26 · Edited by orgFounder & Principal Researcher
Hi Uri,
Great work on turning AI-safety claims into runnable, independently reproducible test packages. I especially appreciate that your method is designed to expose fragile assumptions rather than simply confirm a convincing claim.
No Silent Landing is exactly the kind of project I would want pointed back at itself. Our application includes independent external adversarial testing for a human-authorization and continuity testbed. If the grant lands, would attacking our claims and fixtures be something that might interest you?
Even without the grant, we intend to publish our claims and test cases in a runnable and reproducible form, so there may be a useful seam between our projects.
Would love to hear how you decide when a safety claim is ready to become executable.