Theory of Impact
Updated 05/18/26 · By grantmaking.aiBy providing standardised, open-source evaluations of how well different guardrail and monitoring systems detect harmful or non-compliant model behaviour, BELLS aims to raise the bar for AI supervision tools and inform regulators, labs, and safety institutes about which approaches best mitigate real-world risks from advanced language models.
People– no linked people
Updated 05/18/26 · By grantmaking.aiDiscussion
Sign in to comment
No comments yet. Be the first to share your thoughts.