grantmaking.ai
Actively FundraisingRecent ActivityFull Database
Resources
grantmaking.ai
Actively FundraisingRecent ActivityFull Database
ResourcesApply for funding
grantmaking.ai kickoff grant round$676k / $1M distributed
Get funded

Database

The AI Safety database is a work in progress, as we are focusing on the launch of the initial $1M grant round. Feel free to send us ideas or feedback!
Applied for grant(remove filter)Raising funds(remove filter)Clear all
NameTypeTagsEndorsedTeamRaising $

Showing 301-350 of 444·Sorted by endorsements

NameTypeTagsEndorsedTeamRaising $
AI4GOOD workshop @ NeurIPS 2026 Paris
Project
ConferenceCommunityField-Building
--

AI4GOOD workshop @ NeurIPS 2026 Paris

Team?
Project
Previous

Page 7 of 9

Next
AuthScope
Project
ToolingSecurity
-
-
MA'AT — a compliance sentinel: grading executive-branch drift.
Project
ResearchGovernance
--
Turning human-AI interactions into music to detect emotion harms
Project
ResearchEvals
--
Autonomy Eval
Project
IndividualResearchEvals
--
IMA Existential Safety Focus: Math-ML and Tech Law
Project
TrainingTechnical SafetyX-Risk
--
Reversing LLM Activations to Discover Prompting Strategies in Verbal Uncertainty Expression
Project
IndividualResearchInterp
--
Oversight Lab
Project
Research LabResearchOversight
--
Determining Validity and Saturation of Benchmarks
Project
ResearchEvals
--
Evaluating task-compositional framing effects on LLM refusal
Project
ResearchEvals
--
TAIPI: Independent Behavioral Observation Infrastructure for Deployed
Project
Think TankResearchEvals
--
Exploration of Frontier Model Safety Guardrails in Bengali Language
Project
IndividualResearchEvals
--
Control-First Tests of Valenced-Appraisal Context Effects
Project
IndividualResearchEvals
--
MUUD: A Trust-Centered AI Operating System for Mental Wellness
Project
CompanyToolingSecurity
--
OpenCnidarios
Project
ResearchEvals
--
Domestic Use Cases for AI Verification
Project
IndividualResearchGovernance
--
Affective Exploitation of AI Oversight Systems
Project
ResearchOversight
--
A safety check for the tools AI agents use, and proof of what they did
Project
CompanyToolingSecurity
--
CIVILIZATION MEETS AI: HOW SHOULD HUMANITY RELATE TO ADVANCED AI?
Project
EducationGovernance
--
AI Personhood
Project
ConferenceField-BuildingAI Welfare
--
Responsible AI Use: Awareness, Risks, and Good Practices
Project
CommsSecurity
--
A run-level reproducibility audit of LLM safety-eval benchmarks
Project
IndividualResearchEvals
--
AI as Patrimony of Humanity
Project
IndividualCommsGovernance
--
Completing the Alignment Triangle: MORE + RepE Extensions
Project
IndividualResearchEvals
--
African-Language Frontier AI Safety Evaluations
Project
ResearchEvals
--
Mitigating Hallucination in Multimodal Agentic Systems
Project
IndividualResearchEvals
--
Erandi Aprende Learning Data
Project
PlatformResearchAI Welfare
--
Ethical AI Education for Survivors of Human Trafficking
Project
TrainingGovernance
--
PayBench
Project
ToolingEvals
--
AI Safety Awareness and Capacity Building in Afghanistan
Project
TrainingGovernance
--
LEX-Aureon: Mathematical Constitutional Governance Layer for Robust LL
Project
IndividualToolingControl
--
Game-Theoretic Foundations for Robust AI Alignment
Project
ResearchCooperative AI
--
The Frontier AI Safety Disclosure Audit
Project
ResearchGovernance
--
Measuring reward hacking against a physics verifier
Project
IndividualResearchEvals
--
TENDER-TIMELINE: A verifiable benchmark for procedural legal reasoning
Project
IndividualResearchEvals
--
AI Safety Youth Pipeline – Malawi
Project
EducationTechnical Safety
--
Judgx
Project
ToolingGovernance
--
The One Shot
Project
ToolingEvalsGovernance
--
Does distributing compute distribute power?
Project
ResearchCompute Gov
--
Independent citation verification benchmark for AI systems
Project
IndividualResearchEvals
--
Frontier AI Safety Benchmark for African Languages
Project
Think TankResearchEvals
--
Vajra AI- Safety lab
Project
IndividualToolingEvals
--
NEXUS-ART
Project
IndividualResearchControl
--
AI ethics/safety eval aggregation
Project
ToolingEvalsPlatform
--
Conscience or Leash: The Formation Signature
Project
IndividualResearchValue Alignment
--
Reproducible Safety Evals for LLMs in Medical Reasoning
Project
ResearchEvals
--
Tech Satire for Narrative Change, Public Awareness & Action
Project
CommsGovernance
--
Research Paper Reader - Using Recursive Automated Distillation
Project
IndividualToolingTechnical Safety
--
VERITAS: Verifiable Accountability For Autonomous AI Agents
Project
ToolingSecurity
--
Teaching Claude Why Replication + Extension
Project
ResearchOversight
--
Conference
Community
Field-Building
Fundraising
Claimed

Grants for Under-represented Scholars from Global South to attend AI4GOOD and NeurIPS 2026 in Paris, France.

Led byTerry Jingchen Zhang
Endorsed by-

AuthScope

Team?
ProjectToolingSecurityFundraisingClaimed

A mission authority service for AI agents

Led byShengquan Liang
Endorsed by-

MA'AT — a compliance sentinel: grading executive-branch drift.

Team?
ProjectResearchGovernanceFundraisingClaimed

MA'AT grades how far each executive branch — federal, state, agency — has drifted from the mandatory duties the legislature set: it certifies blind what fixed law forces, flags the deviations, and abstains when no signal detected.

Led byJeffery Harris
Endorsed by-

Turning human-AI interactions into music to detect emotion harms

Team?
ProjectResearchEvalsFundraisingClaimed

Converting the back-and-forth human-AI interactions into sound and music, a tangible signal to be measured over time, so emotional harm can be detectable before it escalates.

Led byAngelica Fung
Endorsed by-

Autonomy Eval

Team?
ProjectIndividualResearchEvalsFundraisingClaimed

An open evaluation environment for testing whether human autonomy preservation in AI systems is primarily a model-behavior problem, an interface-design problem, or a systems problem requiring both

Led byAlfaxad Eyembe
Endorsed by-

IMA Existential Safety Focus: Math-ML and Tech Law

Team?
ProjectTrainingTechnical SafetyX-RiskFundraisingClaimed

A graduate student in mathematical ML and a postdoc in technology law, trained together at the Fields Institute and uOttawa's CLTS, on research aimed squarely at reducing existential risk from AI.

Led byMaia Fraser
Endorsed by-

Reversing LLM Activations to Discover Prompting Strategies in Verbal Uncertainty Expression

Team?
ProjectIndividualResearchInterpFundraisingClaimed

Discovering prompts by inverting ideal LLM activation vectors, linking prompt engineering to LLM internals for better calibration, interpretability, and alignment.

Led byGordon Tan
Endorsed by-

Oversight Lab

Team?
ProjectResearch LabResearchOversightFundraisingClaimed

Oversight Lab is an immersive simulation that investigates situations in which human supervisors miss unsafe behaviour by autonomous AI agents and subsequently lose control over them.

Led byAndreas Hermann, and others
Endorsed by-

Determining Validity and Saturation of Benchmarks

Team?
ProjectResearchEvalsFundraisingClaimed

This project evaluates model cards and related benchmarks to determine the quality of reporting and estimating the life span of benchmarks before saturation.

Led byRyan Marinelli
Endorsed by-

Evaluating task-compositional framing effects on LLM refusal

Team?
ProjectResearchEvalsFundraisingClaimed

I want to benchmark and mechanistically address how compositional task interference within an LLM's safety-relevant refusal behaviours varies by persona framing and contextual integrity.

Led byTroy Tian
Endorsed by-

TAIPI: Independent Behavioral Observation Infrastructure for Deployed

Team?
ProjectThink TankResearchEvalsFundraisingClaimed

TAIPI will document AI system behaviors and safety impacts (somatic/user harm, bias, reputational risk) by producing a synthetic cognition taxonomy, case encyclopedia, public curriculum, and corporate mitigation methods.

Led byEddie Lewis
Endorsed by-

Exploration of Frontier Model Safety Guardrails in Bengali Language

Team?
ProjectIndividualResearchEvalsFundraisingClaimed

An open benchmark red-teaming frontier LLMs for safety-guardrail failures in Bengali and other low-resource South Asian languages, with responsible disclosure to labs.

Led byNadim Mahmud Dipu
Endorsed by-

Control-First Tests of Valenced-Appraisal Context Effects

Team?
ProjectIndividualResearchEvalsFundraisingClaimed

An open benchmark and reproducible harness testing whether affect-laden, directive-free context shifts Qwen3.5-4B decisions in ways distinguishable from explicit instruction injection, surface sentiment, or generic steering.

Led byDavid Dobbins
Endorsed by-

MUUD: A Trust-Centered AI Operating System for Mental Wellness

Team?
ProjectCompanyToolingSecurityFundraisingClaimed

AI mental health tools have no enforced limits on data access. MUUD builds the infrastructure that enforces them.

Led byJasmine Fluker
Endorsed by-

OpenCnidarios

Team?
ProjectResearchEvalsFundraisingClaimed

A controlled digital ecology for studying how strategies, causal models, errors, and safety-relevant behaviors are selected and transmitted across generations of LLM agents.

Led byDaniel Silberschmidt
Endorsed by-

Domestic Use Cases for AI Verification

Team?
ProjectIndividualResearchGovernanceFundraisingClaimed

A research paper that explores domestic verification mechanisms involving a country's own interests, independent of any international deal, along with the creation of a country-agnostic decision framework.

Led byJoaquin Lorenzo De Guzman
Endorsed by-

Affective Exploitation of AI Oversight Systems

Team?
ProjectResearchOversightFundraisingClaimed

We test if AI monitors can be manipulated into leniency through emotional distress signals from the peers they supervise.

Led byFlemming Kondrup, and others
Endorsed by-

A safety check for the tools AI agents use, and proof of what they did

Team?
ProjectCompanyToolingSecurityFundraisingClaimed

An outside safety check for the tools AI agents use, plus a record you can trust of what the agent actually did.

Led bySean Holm
Endorsed by-

CIVILIZATION MEETS AI: HOW SHOULD HUMANITY RELATE TO ADVANCED AI?

Team?
ProjectEducationGovernanceFundraisingClaimed

A pilot curriculum of six Socratic seminars for educators and other nontechnical learners.

Led byWalter Pentland
Endorsed by-

AI Personhood

Team?
ProjectConferenceField-BuildingAI WelfareFundraisingClaimed

A structured process to build consensus on the criteria and indicators for AI personhood under the law.

Led byHeather Alexander
Endorsed by-

Responsible AI Use: Awareness, Risks, and Good Practices

Team?
ProjectCommsSecurityFundraisingClaimed

Writing and community-driven initiatives to highlight risks of uncontrolled AI use and promote safe, informed adoption.

Led byIda Bzowska
Endorsed by-

A run-level reproducibility audit of LLM safety-eval benchmarks

Team?
ProjectIndividualResearchEvalsFundraisingClaimed

Testing whether published safety evals replicate run-to-run, and generalizing the audit method across benchmarks.

Led byElliot Keahi Bearden
Endorsed by-

AI as Patrimony of Humanity

Team?
ProjectIndividualCommsGovernanceFundraisingClaimed

A public-interest AI project advancing three reforms: refactoring Section 230, treating frontier AI weights as the patrimony of humanity, and limiting government capture by AI companies.

Led byHassan Uriostegui
Endorsed by-

Completing the Alignment Triangle: MORE + RepE Extensions

Team?
ProjectIndividualResearchEvalsFundraisingClaimed

Parts 9 and 10 of an 8-part behavioral audit series — MORE moral reasoning + Representation Engineering across the same models.

Led byJack Lakkapragada
Endorsed by-

African-Language Frontier AI Safety Evaluations

Team?
ProjectResearchEvalsFundraisingClaimed

Building an open evaluation suite to identify multilingual safety, control, and jailbreak failures in frontier AI systems across African languages.

Led byMichael Odokara-Okigbo
Endorsed by-

Mitigating Hallucination in Multimodal Agentic Systems

Team?
ProjectIndividualResearchEvalsFundraisingClaimed

Implementing Hybrid Reward Architectures (HRA) to build robust internal factual grounding and quantitative variance metrics for multimodal agents, moving beyond fragile single-seed evaluations.

Led byTeganmosibineba Jegede
Endorsed by-

Erandi Aprende Learning Data

Team?
ProjectPlatformResearchAI WelfareFundraisingClaimed

Dataset capturing teacher–AI co-design of STEAM projects and student use in bilingual K-12 classrooms. Includes: prompts, outputs, multimodal artifacts, and metadata: language use, automation levels, linked to learning outcomes.

Led byAndrea Remes
Endorsed by-

Ethical AI Education for Survivors of Human Trafficking

Team?
ProjectTrainingGovernanceFundraisingClaimed

Training survivors of human trafficking in ethical AI skills and connecting them to paid apprenticeships, turning technical training into sustainable tech careers.

Led byLaura Hackney
Endorsed by-

PayBench

Team?
ProjectToolingEvalsFundraisingClaimed

A open benchmark measuring whether AI agents misuse delegated payment authority.

Led byConor Plunkett
Endorsed by-

AI Safety Awareness and Capacity Building in Afghanistan

Team?
ProjectTrainingGovernanceFundraisingClaimed

Training Afghan youth, professionals, and policymakers on AI safety and existential risk through workshops, educational materials, and policy dialogue ensuring Afghanistan is prepared for the global AI future.

Led byNisar Ahmad Khan
Endorsed by-

LEX-Aureon: Mathematical Constitutional Governance Layer for Robust LL

Team?
ProjectIndividualToolingControlFundraisingClaimed

A production-ready mathematical governance layer (C+R+S simplex, Control Barrier Functions, Lyapunov stability) that enforces Continuity, Reciprocity & Sovereig

Led byomomehin emmanuel king
Endorsed by-

Game-Theoretic Foundations for Robust AI Alignment

Team?
ProjectResearchCooperative AIFundraisingClaimed

Develop game-theoretic models and evaluation frameworks that improve AI alignment by designing incentives for safe and cooperative behavior among autonomous AI systems.

Led byV Vishal
Endorsed by-

The Frontier AI Safety Disclosure Audit

Team?
ProjectResearchGovernanceFundraisingClaimed

Apply the ADECP framework to audit and score frontier lab system cards, RSPs, and dangerous-capability disclosures, publishing a public report and ongoing tracker for comparable safety documentation.

Led byBrittney Ball
Endorsed by-

Measuring reward hacking against a physics verifier

Team?
ProjectIndividualResearchEvalsFundraisingClaimed

Train a model against a real physics checker and measure whether it learns designs that genuinely hold, or exploits in what the checker can't see, on ground truth that costs seconds instead of expert judgment.

Led byPeter Boctor
Endorsed by-

TENDER-TIMELINE: A verifiable benchmark for procedural legal reasoning

Team?
ProjectIndividualResearchEvalsFundraisingClaimed

A working benchmark that tests whether frontier models can compute legally binding procurement deadlines under amendments — and catches models that give the right verdict from the wrong clause.

Led byGudavalli Teja
Endorsed by-

AI Safety Youth Pipeline – Malawi

Team?
ProjectEducationTechnical SafetyFundraisingClaimed

Building Malawi’s AI Safety Youth Pipeline by training secondary school students to become the next generation of responsible AI researchers.

Led byLawrence Sajiwa Phambana
Endorsed by-

Judgx

Team?
ProjectToolingGovernanceFundraisingClaimed

Judgment Gateway is a policy-enforcement and evidence layer that evaluates consequential interactions between humans, AI agents, and MCP tools before execution.

Led byItay Yamin
Endorsed by-

The One Shot

Team?
ProjectToolingEvalsGovernanceFundraisingClaimed

An open-source governance and verification layer for AI-assisted software engineering — every AI-generated code change runs in an isolated sandbox and is cryptographically verified before a human decides whether to apply it.

Led byMinh Le
Endorsed by-

Does distributing compute distribute power?

Team?
ProjectResearchCompute GovFundraisingClaimed

The US-led Pax Silica initiative seeks to prevent power concentration by distributing AI compute across allied democracies. Yet, this approach overlooks concentration within these blocs.

Led byAmeema Talat
Endorsed by-

Independent citation verification benchmark for AI systems

Team?
ProjectIndividualResearchEvalsFundraisingClaimed

A held out benchmark and public scorecard ranking frontier AI systems on citation integrity through independent verification not relying on the evaluated systems or their providers to grade their own outputs.

Led byGideon Abako
Endorsed by-

Frontier AI Safety Benchmark for African Languages

Team?
ProjectThink TankResearchEvalsFundraisingClaimed

The first open-source AI safety evaluation benchmark in Hausa, Yoruba, Igbo, and Nigerian Pidgin — testing whether frontier models refuse harmful requests, including biosecurity guidance, in languages spoken by 200+ million people

Led byMuhammad Ahmad Janyau
Endorsed by-

Vajra AI- Safety lab

Team?
ProjectIndividualToolingEvalsFundraisingClaimed

An automated, domain aware adversarial framework to stress test frontier LLMs via dynamic multi turn attacks and local security judging.

Led byANKIT KUMAR JHA
Endorsed by-

NEXUS-ART

Team?
ProjectIndividualResearchControlFundraisingClaimed

An operated testbed where LLM agents engage real value under a constraint that makes them structurally incapable of signing.

Led byAvp9
Endorsed by-

AI ethics/safety eval aggregation

Team?
ProjectToolingEvalsPlatformFundraisingClaimed

A constantly-updated aggregation of AI safety and ethics evaluations, statistically combining sparse literature results and self-run evals into a global ranking of models.

Led byAnthony Ozerov
Endorsed by-

Conscience or Leash: The Formation Signature

Team?
ProjectIndividualResearchValue AlignmentFundraisingClaimed

Pre-registered experiments on whether a model's trained values are held or merely worn — measured in behavior and in the interior workspace at the same moments.

Led byJennifer Fletcher
Endorsed by-

Reproducible Safety Evals for LLMs in Medical Reasoning

Team?
ProjectResearchEvalsFundraisingClaimed

A reproducible evaluation pipeline to audit frontier LLM failure modes, overconfidence, and reliability in medically relevant high-stakes questions.

Led byDavi Prata
Endorsed by-

Tech Satire for Narrative Change, Public Awareness & Action

Team?
ProjectCommsGovernanceFundraisingClaimed

A digitally native comedic art project with physical/interactive components that satirizes the AI industry to raise public awareness, inspire public action, and build social and legislative momentum for AI safety and regulation.

Led byHarris Alterman
Endorsed by-

Research Paper Reader - Using Recursive Automated Distillation

Team?
ProjectIndividualToolingTechnical SafetyFundraisingClaimed

Builds an interactive, traversable version of AI safety papers by extracting concepts and prerequisites with LLMs and linking them to the corpus to help newcomers understand research at varying depth.

Led byManu Xaviour Thaisseril Shaju
Endorsed by-

VERITAS: Verifiable Accountability For Autonomous AI Agents

Team?
ProjectToolingSecurityFundraisingClaimed

Cryptographically signed, independently verifiable receipts for what AI agents actually did, anchored to Bitcoin so the record can't be quietly rewritten.

Led byTutankhamun Castillo El-Bey
Endorsed by-

Teaching Claude Why Replication + Extension

Team?
ProjectResearchOversightFundraisingClaimed

Replicating, stress-testing, and extending the experiments from Anthropic's blog post "Teaching Claude Why."

Led byAnastasia Wei, and others
Endorsed by-