An open-source benchmark that measures whether tool-using AI agents faithfully report their actions, failures, and policy violations, using deterministic execution traces as ground truth. Read more
Actively Fundraising
AI safety projects actively seeking funding: what they’re working on and how much they need.



