grantmaking.ai
Actively FundraisingRecent ActivityFull Database
Resources
grantmaking.ai
Actively FundraisingRecent ActivityFull Database
ResourcesApply for funding
grantmaking.ai kickoff grant round$676k / $1M distributed
Get funded

Database

The AI Safety database is a work in progress, as we are focusing on the launch of the initial $1M grant round. Feel free to send us ideas or feedback!
Applied for grant(remove filter)Raising funds(remove filter)Clear all
NameTypeTagsEndorsedTeamRaising $

Showing 101-150 of 442·Sorted by endorsements

NameTypeTagsEndorsedTeamRaising $
"Out of this Box! - The Last Musical (Written by Humans)"
Project
MediaTechnical SafetyComms
-

"Out of this Box! - The Last Musical (Written by Humans)"

Team?
Project
Previous

Page 3 of 9

Next
Mechanistic interpretability Discord events and research
Project
CommunityInterpNetwork
+1
-
Solving Scheming and Deception in LLMs
Project
DeceptionResearch
-
Constraining and Scoping Foundation Model Agent Behavior
Project
ResearchIndividualOversight
-
AI Safety Studios
Project
MediaCommsX-Risk
-
The Multilingual Blind Spot in Production-Style Safety Probes
Project
ResearchIndividualEvals
-
Alignment Game
Project
GovernanceComms
-
Bridging the Gap: AI Scenario based Policy Pipeline
Project
ResearchForecastingGovernance
+1
-
PropagateBench: do AI agents pay to share info, or free-ride?
Project
ToolingEvals
+1
-
Governing Off-World AI Infrastructure Before a Single Actor Does
Project
ResearchIndividualCompute Gov
+1
-
Project Ocean: Trustworthy Human-AI Collaboration Infrastructure
Project
ResearchTechnical Safety
+1
-
Why Is It So Hard to Build Truly Safe AI? Dynamic Constraint Boundary
Project
ResearchIndividualTechnical Safety
+1
-
The Human Liveness Project
Project
Field-BuildingGovernance
+1
-
CareerMap: A navigation tool for careers in AI Safety
Project
Field-BuildingPlatformTechnical Safety
-
Emergent alignment: generalizing narrow values OOD
Project
ResearchValue Alignment
+1
-
Intensive Mandarin (ICLP) for Chinese-language AI governance research
Project
GovernanceIndividualTraining
+1
-
Shared Screening for AI Safety Hiring of Generalists
Project
EAPlatformOps
+1
-
No Silent Landing — Landing Integrity Testbed
Project
ToolingIndividualEvals
+1
-
CIRIS Ethical AI Deployment
Project
ToolingIndividualEvals
+1
-
Preserving the Human Veto
Project
GovernanceTrainingNetwork
+1
-
Empirically evaluating faithfulness and transparency in AI delegation
Project
ResearchEvalsDemocratic AI
+1
-
Grant to establish an AI safety lab and fellowship from India
Project
ResearchTechnical SafetyResearch Lab
--
Automated control research
Project
ControlResearch
--
Frontier Safety Talent: Africa
Project
Field-BuildingTrainingTechnical Safety
--
Managing AI X-Risk
Project
MediaGovernanceComms
--
AI Safety Nepal
Project
CommunityTrainingTechnical Safety
--
Can Honest AI Agents Finish Previous Attacks?
Project
ResearchIndividualSecurity
--
PayBench: A Benchmark for Unsafe Commercial Autonomy
Project
ResearchEvals
--
AI Safety Career Coaching (Intro + Recurring)
Project
Field-BuildingIndividualTechnical Safety
--
Behavioral Safety in Conversational and Agentic AI
Project
ToolingEvals
--
Frankfurt AI Safety
Project
CommunityTechnical SafetyNetwork
--
GCR Disaster Response Benchmark
Project
ResearchEvalsX-Risk
--
Neural data privacy evals on frontier LLMs
Project
ResearchIndividualEvals
--
The Capability Race: Inside the AI Safety Crisis
Project
MediaTechnical SafetyComms
--
BMG: Biosafety Moderation Gateway for Agents and LLMs
Project
BiosecurityTooling
--
Machine God - documentary film on AI X-Risk
Project
MediaCommsX-Risk
--
Multi-agent cooperative contracts
Project
ResearchCooperative AI
--
AI ≠ Human Time
Project
ToolingResearch
--
Are Global AI-Biosecurity Safeguards Really Global?
Project
BiosecurityResearch
--
Scaling Honesty: Deception Detection & Steering Beyond 70B
Project
DeceptionResearch
--
echo-truth-llm: Open Source Deception Detection & Correction Toolkit
Project
DeceptionTooling
--
proofbundle: Offline Evidence Receipts for AI Safety Evaluations
Project
ToolingSecurity
--
Coordination Studies
Project
Field-BuildingGovernanceThink Tank
--
Comprehension Audits to Mitigate Risks from Automated AI Research
Project
StandardsGovernance
--
Daios, an independent lab post-training machines with virtue
Project
ResearchTechnical SafetyResearch Lab
--
Estonian AISI
Project
AdvocacyGovernanceAcademic
--
Stress-Testing AgentHarm
Project
ResearchIndividualEvals
--
EthicsNet: Open Safety Infrastructure for Agentic AI
Project
ToolingPlatformTechnical Safety
--
Coalition Circuits: Mechanistic Trajectories of Cooperation and Betray
Project
InterpResearchIndividual
--
A Game-Theoretic Model of AI Consciousness Disagreement as X-Risk
Project
AI WelfareResearchGovernance
--
Media
Technical Safety
Comms
Fundraising
Claimed

We have validated the artistic value of our show; this grant would test whether it can become a scalable and repeatable form of AI-safety outreach.

Led byStephan Wäldchen
Endorsed by

Mechanistic interpretability Discord events and research

Team?
ProjectCommunityInterpNetworkFundraisingClaimed

Making the mechinterp Discord more active through events and research projects.

Led byVictor Levoso Fernandez
Endorsed by
+1

Solving Scheming and Deception in LLMs

Team?
ProjectDeceptionResearchFundraisingClaimed

I study how training processes produce models that behave deceptively and pursue hidden objectives, with scheming as the most consequential case.

Led byAmina Keldibek
Endorsed by

Constraining and Scoping Foundation Model Agent Behavior

Team?
ProjectResearchIndividualOversightFundraisingClaimed

A specification-driven architecture for building more controllable and reliable long-horizon AI agents.

Led byPinar Ozisik
Endorsed by

AI Safety Studios

Team?
ProjectMediaCommsX-RiskFundraisingClaimed

A platform to connect funders to filmmakers who want to create AI Safety films.

Led byDylan Tuccillo, and others
Endorsed by

The Multilingual Blind Spot in Production-Style Safety Probes

Team?
ProjectResearchIndividualEvalsFundraisingClaimed

Auditing production-style safety probes for silent failure on the world's other languages, and what fixing it costs.

Led bySripad Karne
Endorsed by

Alignment Game

Team?
ProjectGovernanceCommsFundraisingClaimed

A video game that teaches people about AI alignment and race dynamics.

Led byWilliam Harrison
Endorsed by

Bridging the Gap: AI Scenario based Policy Pipeline

Team?
ProjectResearchForecastingGovernanceFundraisingClaimed

A project to reduce catastrophic AI risk by training people to turn rigorously forecasted AI-risk scenarios into actionable policy advice for key decision-makers in government.

Led byJames Newport
Endorsed by
+1

PropagateBench: do AI agents pay to share info, or free-ride?

Team?
ProjectToolingEvalsFundraisingClaimed

First open benchmark measuring how willingly agents propagate costly information through a population under time uncertainty.

Led byAndrey Seryakov
Endorsed by
+1

Governing Off-World AI Infrastructure Before a Single Actor Does

Team?
ProjectResearchIndividualCompute GovFundraisingClaimed

Every existing lunar data framework governs spacecraft telemetry and object registries. None of them governs compute. That's the gap this framework closes before someone builds the orbital data infrastructure first.

Led byPaige Donner
Endorsed by
+1

Project Ocean: Trustworthy Human-AI Collaboration Infrastructure

Team?
ProjectResearchTechnical SafetyFundraisingClaimed

Building the foundational infrastructure for trustworthy, long-term human–AI collaboration.

Led bySarah Carmody
Endorsed by
+1

Why Is It So Hard to Build Truly Safe AI? Dynamic Constraint Boundary

Team?
ProjectResearchIndividualTechnical SafetyFundraisingClaimed

DCB removes probability from AI decision-making and replaces it with internal integrity so AI doesn't guess, it decides within its boundaries.

Led byNaufal Ridwan
Endorsed by
+1

The Human Liveness Project

Team?
ProjectField-BuildingGovernanceFundraisingClaimed

Building tools to ensure we can continue to do good in the AI era

Led byAlbert Wang
Endorsed by
+1

CareerMap: A navigation tool for careers in AI Safety

Team?
ProjectField-BuildingPlatformTechnical SafetyFundraisingClaimed

CareerMap is an interactive career discovery tool that maps non-obvious AI safety career paths to help broaden and guide talent beyond Western EA-adjacent circles.

Led byAmit Kumar
Endorsed by

Emergent alignment: generalizing narrow values OOD

Team?
ProjectResearchValue AlignmentFundraisingClaimed

Testing whether fine-tuning an LLM on one narrow prosocial value (e.g. compassion for nonhuman animals) generalizes OOD, making it broadly safer toward human values (e.g. reduced misalignment, bias, etc.)

Led byAllen Lu
Endorsed by
+1

Intensive Mandarin (ICLP) for Chinese-language AI governance research

Team?
ProjectGovernanceIndividualTrainingFundraisingClaimed

11-week intensive Chinese program (ICLP) to move from B2 to C1 Mandarin and remove the bottleneck on US-China-Taiwan AI governance research I'm already doing

Led byDavid Sanchez Garcia
Endorsed by
+1

Shared Screening for AI Safety Hiring of Generalists

Team?
ProjectEAPlatformOpsFundraisingClaimed

Reduce duplicative hiring processes in AI safety organisations by sharing common screening processes for generalist roles.

Led byEmily Marais
Endorsed by
+1

No Silent Landing — Landing Integrity Testbed

Team?
ProjectToolingIndividualEvalsFundraisingClaimed

An open adversarial testbed that catches when a human-approved decision silently changes meaning between approval, memory, and downstream use.

Led byLoek Verdonk
Endorsed by
+1

CIRIS Ethical AI Deployment

Team?
ProjectToolingIndividualEvalsFundraisingClaimed

CIRIS provides a free and open source model-agnostic agentic alignment harness, available on all major app stores.

Led byEric Moore
Endorsed by
+1

Preserving the Human Veto

Team?
ProjectGovernanceTrainingNetworkFundraisingClaimed

Civil-society infrastructure against AI-enabled power concentration, built around autonomous weapons

Led byJenna Jauhiainen, and others
Endorsed by
+1

Empirically evaluating faithfulness and transparency in AI delegation

Team?
ProjectResearchEvalsDemocratic AIFundraisingClaimed

A participatory-budgeting experiment measuring whether AI delegates faithfully represent their principals and whether principals can identify misrepresentation, and an open evaluation suite for AI delegation.

Led byJohann D. Gaebler
Endorsed by
+1

Grant to establish an AI safety lab and fellowship from India

Team?
ProjectResearchTechnical SafetyResearch LabFundraisingManifundClaimed

Focusing more on the intersection of red teaming and interp, while creating a strong rooted community.

Led byRedarc Labs | Manan Wadhwa, and others
Endorsed by-

Automated control research

Team?
ProjectControlResearchFundraisingClaimed

I have an automated control research scaffold. I'll use it to find the best control protocols to protect against various attacks and publish reports.

Led byRam Potham
Endorsed by-

Frontier Safety Talent: Africa

Team?
ProjectField-BuildingTrainingTechnical SafetyFundraisingClaimed

A pilot to find, screen and support overlooked African ML talent into frontier AI safety programs such as MATS and ARENA, adding new researchers to the alignment field’s talent pipeline.

Led byGideon Abako
Endorsed by-

Managing AI X-Risk

Team?
ProjectMediaGovernanceCommsFundraisingClaimed

A quarterly Techplomacy Conversations series that turns AI x-risk research into direct, actionable recommendations for foreign ministries, UN missions, and AI companies.

Led byOlin Thakur
Endorsed by-

AI Safety Nepal

Team?
ProjectCommunityTrainingTechnical SafetyFundraisingClaimed

AI safety fellowship to upskill local talent, build a pipeline of people who understand AI safety deeply enough to contribute to research, advise on policy, and coordinate when AI governance decisions are being made.

Led byAnju Chhetri, and others
Endorsed by-

Can Honest AI Agents Finish Previous Attacks?

Team?
ProjectResearchIndividualSecurityFundraisingClaimed

A benchmark that tests whether one AI coding agent can leave behind a harmless looking change that causes a later honest agent to unknowingly finish an attack.

Led byDev Goyal
Endorsed by-

PayBench: A Benchmark for Unsafe Commercial Autonomy

Team?
ProjectResearchEvalsFundraisingClaimed

Benchmark for agents spending human money - paybench.org

Led byConor Plunkett
Endorsed by-

AI Safety Career Coaching (Intro + Recurring)

Team?
ProjectField-BuildingIndividualTechnical SafetyFundraisingClaimed

Seeking a top-up grant to prioritize AI Safety Career Coaching.

Led byJen Baik
Endorsed by-

Behavioral Safety in Conversational and Agentic AI

Team?
ProjectToolingEvalsFundraisingClaimed

Building computational tools to identify and address pathological behavioral states in conversational and autonomous AI

Led byAlejandro De Los Angeles
Endorsed by-

Frankfurt AI Safety

Team?
ProjectCommunityTechnical SafetyNetworkFundraisingClaimed

An AI safety community that trains professionals and researchers to become competent AI safety contributors within industries and the academia, by learning and doing something.

Led byHelen Adepoju
Endorsed by-

GCR Disaster Response Benchmark

Team?
ProjectResearchEvalsX-RiskFundraisingClaimed

An LLM benchmark for emergency response communications in existential catastrophes.

Led byJames Mulhall, and others
Endorsed by-

Neural data privacy evals on frontier LLMs

Team?
ProjectResearchIndividualEvalsFundraisingClaimed

A dangerous-capability benchmark testing whether frontier LLMs can extract sensitive attributes from anonymized brain recordings, enable adversaries to build privacy-violating tools, or over-refuse legitimate neuroscience queries.

Led bySwaraag Sistla
Endorsed by-

The Capability Race: Inside the AI Safety Crisis

Team?
ProjectMediaTechnical SafetyCommsFundraisingClaimed

A high-production visual journalism series combining primary-source data and motion graphics to break down the mechanics, safety crises, linguistic biases, and Arab-world implications of transformative AI.

Led byMagdi Hamed
Endorsed by-

BMG: Biosafety Moderation Gateway for Agents and LLMs

Team?
ProjectBiosecurityToolingFundraisingClaimed

BMG is a drop in open ai compatible gateway that screens LLM and Agent calls for dual use biological risk.

Led byKIMON ANTONIOS PROVATAS
Endorsed by-

Machine God - documentary film on AI X-Risk

Team?
ProjectMediaCommsX-RiskFundraisingClaimed

Feature-length documentary that captures the intellectual and cultural zeitgeist surrounding frontier AI at the moment before AGI, and communicates X-Risk ideas to a broad range of elites.

Led bySteve Hsu
Endorsed by-

Multi-agent cooperative contracts

Team?
ProjectResearchCooperative AIFundraisingClaimed

Multi-agent cooperative contractual obligation framework.

Led byBenjamin John Schulz
Endorsed by-

AI ≠ Human Time

Team?
ProjectToolingResearchFundraisingClaimed

Build a digital intervention that measures the human time needed to produce AI-generated outputs

Led byAbigail Browka
Endorsed by-

Are Global AI-Biosecurity Safeguards Really Global?

Team?
ProjectBiosecurityResearchFundraisingClaimed

Assess AI-enabled biosecurity safeguards in Nigeria by reviewing governance, red-teaming frontier models with local language/culture, and testing DNA synthesis screening for orders from resource-limited settings.

Led byNnaemeka Emmanuel Nnadi, and others
Endorsed by-

Scaling Honesty: Deception Detection & Steering Beyond 70B

Team?
ProjectDeceptionResearchFundraisingClaimed

Extend recently developed deception detection+steering method to larger and more diverse open-weight models, releasing the tooling that makes frontier-scale honesty interventions reproducible.

Led byJared Glover
Endorsed by-

echo-truth-llm: Open Source Deception Detection & Correction Toolkit

Team?
ProjectDeceptionToolingFundraisingClaimed

An open-source library that both detects self-chosen deception in open-weight LLMs and steers the model back toward honesty at inference — the correction half that current honesty tools lack.

Led byJared Glover
Endorsed by-

proofbundle: Offline Evidence Receipts for AI Safety Evaluations

Team?
ProjectToolingSecurityFundraisingClaimed

proofbundle lets people verify AI safety evaluation results offline instead of trusting a number in a PDF and it catches if an evaluation record was altered or swapped

Led byKonrad Gruszka
Endorsed by-

Coordination Studies

Team?
ProjectField-BuildingGovernanceThink TankFundraisingClaimed

Coordination Studies is a new field-building project oriented towards solving coordination problems and designing new coordination mechanisms.

Led byMax Holly
Endorsed by-

Comprehension Audits to Mitigate Risks from Automated AI Research

Team?
ProjectStandardsGovernanceFundraisingClaimed

*Comprehension audits* are a novel development-process assurance mechanism to verify human understanding of AI research outputs to act as a gate to slow automation of AI R&D.

Led byRonald Bodkin
Endorsed by-

Daios, an independent lab post-training machines with virtue

Team?
ProjectResearchTechnical SafetyResearch LabFundraisingClaimed

Post-training virtue into open models as a third alignment technique and control mechanism

Led byAndrew Agathon, and others
Endorsed by-

Estonian AISI

Team?
ProjectAdvocacyGovernanceAcademicFundraisingClaimed

6 months of salary support for my research and operations work to help found a quasi-governmental AI safety institution, the [Estonian AI Security Institute](https://www.aisi.ee/) (Turvalise Tehisaru Teadmuskeskus T3).

Led byKristi Uustalu
Endorsed by-

Stress-Testing AgentHarm

Team?
ProjectResearchIndividualEvalsFundraisingClaimed

Run AgentHarm on 3–4 frontier/open-weight models, manually audit transcripts for metric gaming and spurious failures, compare to prior critiques, and publish a detailed evidence-based writeup.

Led byAnshuman Singh
Endorsed by-

EthicsNet: Open Safety Infrastructure for Agentic AI

Team?
ProjectToolingPlatformTechnical SafetyFundraisingClaimed

Free, open, forkable safety infrastructure for agentic AI: a peer-reviewed pathology nosology, a safety runtime with reproducible benchmarks, and enforcement gating every agent tool call against human-authored policy.

Led byEleanor 'Nell' Watson
Endorsed by-

Coalition Circuits: Mechanistic Trajectories of Cooperation and Betray

Team?
ProjectInterpResearchIndividualFundraisingClaimed

An open benchmark and causal interpretability study of when individually power-motivated LLM agents compete, form coalitions, collude, or betray—and whether internal signals reveal these shifts before behavior does.

Led bySubramanyam Sahoo
Endorsed by-

A Game-Theoretic Model of AI Consciousness Disagreement as X-Risk

Team?
ProjectAI WelfareResearchGovernanceFundraisingClaimed

A strategic model of governance under uncertainty: how disagreement about AI consciousness undermines the coordination that keeps AI risks in check.

Led bySoenke Ziesche
Endorsed by-