grantmaking.ai
Actively FundraisingRecent ActivityFull Database
Resources
grantmaking.ai
Actively FundraisingRecent ActivityFull Database
ResourcesApply for funding
grantmaking.ai kickoff grant round$583k / $1M distributed
Get funded

Database

The AI Safety database is a work in progress, as we are focusing on the launch of the initial $1M grant round. Feel free to send us ideas or feedback!
Applied for grant(remove filter)Raising funds(remove filter)Clear all
NameTypeTagsEndorsedTeamRaising $

Showing 51-100 of 441·Sorted by endorsements

NameTypeTagsEndorsedTeamRaising $
GenomeGuard: Securing the Genomic AI Supply Chain
Project
BiosecurityToolingSecurity

GenomeGuard: Securing the Genomic AI Supply Chain

Team?
Project
Previous

Page 2 of 9

Next
-
Agent Island: A Multiagent Environment of Cooperation and Conflict
Project
ResearchEvals
+1
-
Pax Machina
Project
NewsletterGovernanceComms
-
Synthesizing Standalone World-Models
Project
ResearchTechnical SafetyAlignment Theory
-
The Blind Spots in AI Risk Assessment
Project
ResearchGovernanceThink Tank
+1
-
COMPASS: AI X-Risk Track
Project
GovernanceTrainingX-Risk
+2
-
PauseAI Australia attendance PauseCon London
Project
GovernanceTrainingNetwork
+2
-
Investigating Chain-of-Thought Faithfulness in AI Reasoning
Project
InterpResearchIndividual
+1
-
Flourishing Futures Fellowship
Project
TrainingX-Risk
-
Belief state geometry of language model personas
Project
InterpResearch
-
Detailed Threat Models for AI-Enabled Extreme Power Concentration
Project
ResearchGovernanceIndividual
+1
-
Surrogate base model for Mechanistic Interpretability
Project
InterpResearch
-
A Mechanistic Analysis of Activation Verbalizers: Promise and Risks
Project
InterpResearch
-
BASE - Bangalore AI Safety Exchange
Project
CommunityHubTechnical Safety
+1
-
Tight PAC-Bayes Generalisation Guarantees Across Frontier LLM Safety Monitoring Deployment Settings
Project
VerificationResearchOversight
-
Humans in Control
Project
CommunityGovernanceNetwork
-
Formally verified autoresearch for theoretical mech interp
Project
InterpResearch
-
Survey of the American public on AI futures.
Project
ResearchDemocratic AI
+1
-
Fragile Alignment
Project
ResearchTechnical SafetyEvals
-
Separatrix
Project
ResearchEvals
-
Superintelligence Inc. Strategy Game
Project
CommsX-Risk
-
Mapping AI
Project
ToolingGovernancePlatform
-
Invisible Hands: Measuring Agent Steering of Human Researchers
Project
ResearchOversight
-
Educating French decision-makers on AI existential risk
Project
GovernanceNetwork
+2
-
Unlearning in genomic and protein language models
Project
BiosecurityToolingResearch
-
Shaping AI’s « safe culture »
Project
StandardsResearchGovernance
+1
-
Adapative circuit tracing for test-time interpretability
Project
InterpResearchIndividual
-
Closing the Talent Gaps in AI: Webinars, Videos and Peer-Support
Project
CommunityMediaGovernance
+2
-
Developing a new science of agency
Project
ResearchIndividualAgent Foundations
+2
-
Adversarially robust workload classification using hardware sensors
Project
ResearchSecurity
-
AI Agent Overeagerness as a Problem in Safety-Critical Scenarios
Project
ControlResearch
-
Safeguarding open-weight genomic foundation models through weight lock
Project
BiosecurityToolingResearch
-
Does the thinking behind data selection shape a model's character?
Project
ResearchIndividualEvals
+2
-
AI Safety Unconference 2026
Project
CommunityConferenceTechnical Safety
+1
-
Preparing for the future of AI by thinking beyond Transformers
Project
ResearchIndividualTechnical Safety
+1
-
Originality and Decomposition of Generalization.
Project
ResearchEvals
-
“La P8stée” (2027) starring Bobcat and P8stie
Project
MediaCommsX-Risk
+1
-
Horizon AGI - AI Safety France Pilot
Project
Field-BuildingGovernanceNetwork
+1
-
AI Welfare Evaluation Panel: A validation study and pilot
Project
AI WelfareResearch
-
Project AI Kavach : AI safety for builders fellowship
Project
ToolingTrainingTechnical Safety
-
Oceania AI Safety & Young Talent Conference
Project
CommunityConferenceGovernance
-
Evaluating Legal Gradual Disempowerment Risk
Project
ResearchEvals
-
Is AI safety research reproducible?
Project
ResearchTechnical Safety
-
Mapping Persona Contamination in the Training Pipeline
Project
ResearchIndividualEvals
-
Minister for Human Autonomy
Project
AdvocacyGovernance
-
Using dynamical systems theory approach to identify dishonesty in LLMs
Project
DeceptionResearchIndividual
+1
-
AI Safety Social Media Gap Analysis
Project
IndividualCommsX-Risk
-
Alien Cub Productions
Project
MediaCommsX-Risk
-
AI Safety Berlin Visiting Researcher Pilot
Project
CommunityHubTechnical Safety
+1
-
Does the chain-of-thought actually drive the answer?
Project
ResearchEvalsOversight
+1
-
Biosecurity
Tooling
Security
Fundraising
Claimed

GenomeGuard is an open-source defense and threat-discovery framework for detecting data poisoning, compromised annotations, and supply-chain attacks before they propagate into genomic foundation models.

Led byCharalampos Koilakos
Endorsed by

Agent Island: A Multiagent Environment of Cooperation and Conflict

Team?
ProjectResearchEvalsFundraisingClaimed

Agent Island places agents in a rich social setting, similar to reality competitions like Survivor, to study multiagent interactions and the consequences of learning pressure in competitive settings.

Led byConnacher Murphy
Endorsed by
+1

Pax Machina

Team?
ProjectNewsletterGovernanceCommsFundraisingManifundClaimed

A publication about the institutions we need for powerful AI.

Led byOliver Klingefjord, and others
Endorsed by

Synthesizing Standalone World-Models

Team?
ProjectResearchTechnical SafetyAlignment TheoryFundraisingManifundClaimed

Research agenda aimed at developing methods for constructing powerful, easily interpretable world-models.

Led byThane Ruthenis, and others
Endorsed by

The Blind Spots in AI Risk Assessment

Team?
ProjectResearchGovernanceThink TankFundraisingClaimed

AI risk assessment currently checks few threat models and doesn't compose them into the aggregate risk that matters. We'll build a tool mapping what frontier system cards cover and omit, plus a paper on what the assessments miss.

Led byRichard Mallah, and others
Endorsed by
+1

COMPASS: AI X-Risk Track

Team?
ProjectGovernanceTrainingX-RiskFundraisingClaimed

COMPASS (Capacity-Oriented Mentorship for Public Administrators on AI Safety & Strategy): AI X-Risk Track trains sitting Global South government officials to manage AI x-risk, so their governments are part of preventing it.

Led byOlin Thakur
Endorsed by
+2

PauseAI Australia attendance PauseCon London

Team?
ProjectGovernanceTrainingNetworkFundraisingClaimed

Fly two volunteer leaders of PauseAI Australia to PauseCon London, to bring UK's learnings home and train our all-volunteer chapter

Led byPeter Horniak
Endorsed by
+2

Investigating Chain-of-Thought Faithfulness in AI Reasoning

Team?
ProjectInterpResearchIndividualFundraisingClaimed

Probe an open-weight model’s activations under biasing/cue conditions to test whether chain-of-thought explanations match internal reasoning, releasing paper, code, and datasets.

Led byDaniel Zhang
Endorsed by
+1

Flourishing Futures Fellowship

Team?
ProjectTrainingX-RiskFundraisingManifundClaimed

A scalable fellowship training researchers to develop interventions for achieving high-value long-term futures

Led byJordan Arel
Endorsed by

Belief state geometry of language model personas

Team?
ProjectInterpResearchFundraisingClaimed

A formal, testable account of LLM persona selection as Bayesian inference, validated against model internals, so labs can monitor and steer model personas during post-training and deployment

Led byLogan Graves
Endorsed by

Detailed Threat Models for AI-Enabled Extreme Power Concentration

Team?
ProjectResearchGovernanceIndividualFundraisingClaimed

Detailed models of how AI could allow a small set of actors to gain a decisive strategic advantage over the rest of the world: concrete pathways, required capabilities, quantified likelihoods, and the defenses that bind them.

Led byAmritanshu Prasad
Endorsed by
+1

Surrogate base model for Mechanistic Interpretability

Team?
ProjectInterpResearchFundraisingClaimed

Creating a reference model for mechanistic interpretability without assuming that at auditing time we have a safe model to compare the suspicious model against.

Led byRaffaello Fornasiere
Endorsed by

A Mechanistic Analysis of Activation Verbalizers: Promise and Risks

Team?
ProjectInterpResearchFundraisingClaimed

Mechanistically analyze how activation verbalizers use target-model activation concepts (e.g., cyclic day-of-week representations) via PCA/DAS/patching, explain cross-family failures, and improve verbalizers.

Led byTung-Yu Wu
Endorsed by

BASE - Bangalore AI Safety Exchange

Team?
ProjectCommunityHubTechnical SafetyFundraisingClaimed

A Physical Community Hub for AI Safety in Bangalore,India to build long-term AI safety careers, host multiple AI safety fellowships, career events and build a community that raises the long-term impact and value of AI

Led byAsh Singh, and others
Endorsed by
+1

Tight PAC-Bayes Generalisation Guarantees Across Frontier LLM Safety Monitoring Deployment Settings

Team?
ProjectVerificationResearchOversightFundraisingClaimed

Compression-based PAC-Bayes certification for frontier-scale LLM safety monitors, deployment setting shift, and modern post-training.

Led byTim G. J. Rudner, and others
Endorsed by

Humans in Control

Team?
ProjectCommunityGovernanceNetworkFundraisingClaimed

Humans in Control (HIC) is a nonpartisan grassroots advocacy organization focused on AI safeguards.

Led byVael Gates
Endorsed by

Formally verified autoresearch for theoretical mech interp

Team?
ProjectInterpResearchFundraisingClaimed

LLM agents collaborate to discover and formally verify theorems about the internal computations of transformers, beginning with a simple pilot question: how many attention heads are needed to represent a Boolean function?

Led byKarthik Viswanathan
Endorsed by

Survey of the American public on AI futures.

Team?
ProjectResearchDemocratic AIFundraisingClaimed

TLDR: A representative survey with Yougov of the American public on questions about AI futures, including space governance, successionism, values.

Led byJason Hausenloy
Endorsed by
+1

Fragile Alignment

Team?
ProjectResearchTechnical SafetyEvalsFundraisingClaimed

Build an open-source platform of model organisms and agentic sandboxes to iteratively test mitigations for emergent misalignment via white-box probes, trigger tests, and evaluation-awareness checks.

Led byMatteo Leonesi
Endorsed by

Separatrix

Team?
ProjectResearchEvalsFundraisingClaimed

Empirical research to create conditions for cooperative strategies to dominate adversarial ones among a broad swath of near-future AIs - in the narrow window this work is still possible.

Led byJai Dhyani
Endorsed by

Superintelligence Inc. Strategy Game

Team?
ProjectCommsX-RiskFundraisingClaimed

Browser based game in the style of Plague Inc where players act as a rogue AI attempting to escape human control. Intended to give lay audiences a grounded understanding of how ASI x-risk could play out.

Led byConnor Heaton
Endorsed by

Mapping AI

Team?
ProjectToolingGovernancePlatformFundraisingClaimed

This grant would help us maintain and scale Mapping AI, an open-source stakeholder map of the people and organizations with the potential to shape U.S. AI policy.

Led byAnushree Chaudhuri, and others
Endorsed by

Invisible Hands: Measuring Agent Steering of Human Researchers

Team?
ProjectResearchOversightFundraisingClaimed

How do the options agents present to human researchers alter which research paths are explored, and do steered researchers notice or feel less in control?

Led byTrevor Lohrbeer
Endorsed by

Educating French decision-makers on AI existential risk

Team?
ProjectGovernanceNetworkFundraisingClaimed

A dedicated role at Pause IA to brief French policymakers on existential and catastrophic risks from advanced AI, and to train our volunteer network to do the same with their own representatives.

Led byClémence Peyrot, and others
Endorsed by
+2

Unlearning in genomic and protein language models

Team?
ProjectBiosecurityToolingResearchFundraisingClaimed

Implementing different types of unlearning methods for genomic and protein language models to remove sensitive biological information (e.g. pathogen virulence) while preserving predictive performance and scientific utility.

Led byIlias Georgakopoulos-Soares
Endorsed by

Shaping AI’s « safe culture »

Team?
ProjectStandardsResearchGovernanceFundraisingClaimed

Identify what AI early risks or failures could be reported despite strategic rivalries towards a « safe culture » shaped after aviation

Led byBenoît Larrouturou
Endorsed by
+1

Adapative circuit tracing for test-time interpretability

Team?
ProjectInterpResearchIndividualFundraisingClaimed

A training methodology, and transcoder sets that allow to leverage heavy-weight interpretability methods, but made more lightweight for test-time analysis.

Led byAlexandre Doukhan
Endorsed by

Closing the Talent Gaps in AI: Webinars, Videos and Peer-Support

Team?
ProjectCommunityMediaGovernanceFundraisingClaimed

A webinar series and peer-support community helping people identify and move into high-impact talent gaps careers and projects in AI safety and governance.

Led byKarolina Gruzel
Endorsed by
+2

Developing a new science of agency

Team?
ProjectResearchIndividualAgent FoundationsFundraisingClaimed

Naturalizing theoretical alignment by initiating the development of a scientific theory that is capable of making falsifiable empirical claims about agents in general, including humans, AGI, and ASI, and thereby about alignment.

Led byBen Auer, and others
Endorsed by
+2

Adversarially robust workload classification using hardware sensors

Team?
ProjectResearchSecurityFundraisingClaimed

I've previously developed a classifier that distinguishes training and inference based on Nvidia software telemetry. This project will achieve that using physical sensors, making the system more secure.

Led byRobi Rahman, and others
Endorsed by

AI Agent Overeagerness as a Problem in Safety-Critical Scenarios

Team?
ProjectControlResearchFundraisingClaimed

We want to investigate how AI agent overeagerness can backfire when exhibited in safety-critical scenarios.

Led byThilo Hagendorff
Endorsed by

Safeguarding open-weight genomic foundation models through weight lock

Team?
ProjectBiosecurityToolingResearchFundraisingClaimed

Safeguarding open-weight genomic foundation models through weight lock against adversarial finetuning

Led byAlexandros Tzanakakis
Endorsed by

Does the thinking behind data selection shape a model's character?

Team?
ProjectResearchIndividualEvalsFundraisingClaimed

We test whether the structure behind the selection matters: one model, three diets (industry standard, our protocol, random), open evals, published either way.

Led byKatja Gorlinski
Endorsed by
+2

AI Safety Unconference 2026

Team?
ProjectCommunityConferenceTechnical SafetyFundraisingClaimed

A participant-led unconference gathering the global AI safety community, online and at local sites, Nov 20-22, 2026.

Led byOrpheus Lummis
Endorsed by
+1

Preparing for the future of AI by thinking beyond Transformers

Team?
ProjectResearchIndividualTechnical SafetyFundraisingClaimed

A project about researching radical methods to both advance and prepare countermeasures for what is to come after transformers.

Led byDarsh M
Endorsed by
+1

Originality and Decomposition of Generalization.

Team?
ProjectResearchEvalsFundraisingClaimed

Perform research to rigorously elucidate and quantify generalization versus memorization, and examine evidence of originality in LLMS.

Led byAri Spiesberger
Endorsed by

“La P8stée” (2027) starring Bobcat and P8stie

Team?
ProjectMediaCommsX-RiskFundraisingClaimed

It's La Jetée (1962), the time-traveling film later adapted as 12 Monkeys (1995), but it's about X-risk and inspired by AI 2027 and it's directed by, and starring, myself and @p8stie. Ergo, La P8stée

Led byScott Ellis
Endorsed by
+1

Horizon AGI - AI Safety France Pilot

Team?
ProjectField-BuildingGovernanceNetworkFundraisingClaimed

Launching a French AI safety community through university outreach and local events to connect students, researchers, practitioners, and policymakers.

Led byArthur DAVID
Endorsed by
+1

AI Welfare Evaluation Panel: A validation study and pilot

Team?
ProjectAI WelfareResearchFundraisingClaimed

An independent, multi-lineage panel of AI models, tested on whether it can identify welfare concerns in AI evaluations with opinions, dissents, and proposed modifications published in a public registry

Led byJuliana Grant
Endorsed by

Project AI Kavach : AI safety for builders fellowship

Team?
ProjectToolingTrainingTechnical SafetyFundraisingClaimed

A 3 / 6 month builders fellowship where mentor-builder pods ship practical AI safety tools, not papers.

Led byAnkur Pandey, and others
Endorsed by

Oceania AI Safety & Young Talent Conference

Team?
ProjectCommunityConferenceGovernanceFundraisingClaimed

Hosting a full day conference based in Sydney, Australia where young, aspiring students in senior high school and university interested in AI Safety can connect and share ideas

Led byWilliam Wu
Endorsed by

Evaluating Legal Gradual Disempowerment Risk

Team?
ProjectResearchEvalsFundraisingClaimed

An evaluation suite to identify a model’s legal values (e.g., anti-tech-regulation) relative to well-known actors (e.g., Ruth Bader Ginsburg) and an assessment of how language in a model’s constitution impacts the extent to which these values are human-aligned.

Led byMagnus Saebo, and others
Endorsed by

Is AI safety research reproducible?

Team?
ProjectResearchTechnical SafetyFundraisingClaimed

A SCORE-style computational reproducibility audit of empirical AI safety research that estimates the field's base rate of reproducibility and generates a taxonomy of its failure modes.

Led byJordan Suchow
Endorsed by

Mapping Persona Contamination in the Training Pipeline

Team?
ProjectResearchIndividualEvalsFundraisingClaimed

Map how undesired behaviors can silently spread between models during training, and which parts of the pipeline have the greatest risk — starting with the feedback processes used to align models.

Led byLily Wen
Endorsed by

Minister for Human Autonomy

Team?
ProjectAdvocacyGovernanceFundraisingClaimed

Research and plan advocacy for creating a UK Minister for Human Autonomy, including foundational research, project management, and drafting policy recommendations to safeguard human autonomy amid AI.

Led byJoshua Krook
Endorsed by

Using dynamical systems theory approach to identify dishonesty in LLMs

Team?
ProjectDeceptionResearchIndividualFundraisingClaimed

Identifying reasoning pathologies/disingenuous behavior in reasoning traces based on activation dynamics rather than apparent semantics

Led byJosiah Kratz
Endorsed by
+1

AI Safety Social Media Gap Analysis

Team?
ProjectIndividualCommsX-RiskFundraisingClaimed

Auditing the social media presence of the top 25 AI safety organizations across major platforms to quantify the public communication gap and publishing a full gap analysis, then giving recommendations to the orgs for improvement.

Led byJude Williams
Endorsed by

Alien Cub Productions

Team?
ProjectMediaCommsX-RiskFundraisingClaimed

Alien Cub is a production company focused on stories that inform a wide audience about risks from transformative AI.

Led byKeiran Harris
Endorsed by

AI Safety Berlin Visiting Researcher Pilot

Team?
ProjectCommunityHubTechnical SafetyFundraisingClaimed

Short-term residencies (2–4 weeks) that bring high-context AI safety researchers to work at AI Safety Berlin’s coworking hub to connect with local researchers and career-transitioners and seed a European relocation pipeline.

Led byAlex McKenzie, and others
Endorsed by
+1

Does the chain-of-thought actually drive the answer?

Team?
ProjectResearchEvalsOversightFundraisingClaimed

A Bayesian causal auditor that quantifies chain-of-thought faithfulness while accounting for hidden confounding in shared-network language models.

Led byMaksim Silchenko
Endorsed by
+1