grantmaking.ai
Actively FundraisingRecent ActivityFull Database
Resources
grantmaking.ai
Actively FundraisingRecent ActivityFull Database
ResourcesApply for funding
grantmaking.ai kickoff grant round$278k / $1M distributed
Get funded

Database

The AI Safety database is a work in progress, as we are focusing on the launch of the initial $1M grant round. Feel free to send us ideas or feedback!
Raising funds(remove filter)Clear all
NameTypeTagsEndorsedTeamRaising $

Showing 51-100 of 459·Sorted by endorsements

NameTypeTagsEndorsedTeamRaising $
Synthesizing Standalone World-Models
Project
ResearchTechnical SafetyAlignment Theory
-

Synthesizing Standalone World-Models

Team?
Project
Previous

Page 2 of 10

Next
The Blind Spots in AI Risk Assessment
Project
Think TankResearchGovernance
+1
-
COMPASS: AI X-Risk Track
Project
TrainingGovernanceX-Risk
+2
-
Investigating Chain-of-Thought Faithfulness in AI Reasoning
Project
IndividualResearchInterp
+1
-
Flourishing Futures Fellowship
Project
TrainingX-Risk
-
Belief state geometry of language model personas
Project
ResearchInterp
-
Detailed Threat Models for AI-Enabled Extreme Power Concentration
Project
IndividualResearchGovernance
+1
-
Surrogate base model for Mechanistic Interpretability
Project
ResearchInterp
-
A Mechanistic Analysis of Activation Verbalizers: Promise and Risks
Project
ResearchInterp
-
BASE - Bangalore AI Safety Exchange
Project
HubCommunityTechnical Safety
+1
-
Tight PAC-Bayes Generalisation Guarantees Across Frontier LLM Safety Monitoring Deployment Settings
Project
ResearchVerificationOversight
-
Humans in Control
Project
NetworkCommunityGovernance
-
Formally verified autoresearch for theoretical mech interp
Project
ResearchInterp
-
DC Mini-Conference 2.0 (DCMC 2.0)
Project
ConferenceCommunityGovernance
-
Survey of the American public on AI futures.
Project
ResearchDemocratic AI
+1
-
Fragile Alignment
Project
ResearchEvalsTechnical Safety
-
Separatrix
Project
ResearchEvals
-
Superintelligence Inc. Strategy Game
Project
CommsX-Risk
-
Mapping AI
Project
PlatformToolingGovernance
-
Invisible Hands: Measuring Agent Steering of Human Researchers
Project
ResearchOversight
-
Unlearning in genomic and protein language models
Project
ResearchToolingBiosecurity
-
Lisa Intel: A Global Runtime Governance Layer for Autonomous AI
Project
PlatformToolingControl
+2
-
Shaping AI’s « safe culture »
Project
ResearchGovernanceStandards
+1
-
Adapative circuit tracing for test-time interpretability
Project
IndividualResearchInterp
-
Closing the Talent Gaps in AI: Webinars, Videos and Peer-Support
Project
MediaCommunityGovernance
+2
-
Developing a new science of agency
Project
IndividualResearchAgent Foundations
+2
-
Adversarially robust workload classification using hardware sensors
Project
ResearchSecurity
-
AI Agent Overeagerness as a Problem in Safety-Critical Scenarios
Project
ResearchControl
-
Safeguarding open-weight genomic foundation models through weight lock
Project
ResearchToolingBiosecurity
-
Does the thinking behind data selection shape a model's character?
Project
ResearchEvalsIndividual
+2
-
AI Safety Unconference 2026
Project
ConferenceCommunityTechnical Safety
+1
-
Originality and Decomposition of Generalization.
Project
ResearchEvals
-
London Interface for AI Safety (LIAISE)
Project
NetworkTrainingField-Building
+1
-
“La P8stée” (2027) starring Bobcat and P8stie
Project
MediaCommsX-Risk
+1
-
Horizon AGI - AI Safety France Pilot
Project
NetworkField-BuildingGovernance
+1
-
AI Welfare Evaluation Panel: A validation study and pilot
Project
ResearchAI Welfare
-
Project AI Kavach : AI safety for builders fellowship
Project
TrainingToolingTechnical Safety
-
Evaluating Legal Gradual Disempowerment Risk
Project
ResearchEvals
-
Is AI safety research reproducible?
Project
ResearchTechnical Safety
-
Mapping Persona Contamination in the Training Pipeline
Project
IndividualResearchEvals
-
PauseAI Australia attendance PauseCon London
Project
NetworkTrainingGovernance
+1
-
Preparing for the future of AI by thinking beyond Transformers
Project
IndividualResearchTechnical Safety
+1
-
Oceania AI Safety & Young Talent Conference
Project
ConferenceCommunityGovernance
-
Minister for Human Autonomy
Project
AdvocacyGovernance
-
Using dynamical systems theory approach to identify dishonesty in LLMs
Project
IndividualResearchDeception
+1
-
AI Safety Social Media Gap Analysis
Project
IndividualCommsX-Risk
-
Alien Cub Productions
Project
MediaCommsX-Risk
-
Does the chain-of-thought actually drive the answer?
Project
ResearchEvalsOversight
+1
-
"Out of this Box! - The Last Musical (Written by Humans)"
Project
MediaCommsTechnical Safety
-
Sydney AI Safety Space - General Funding
Project
HubCommunityTechnical Safety
-
Research
Technical Safety
Alignment Theory
Fundraising
Manifund
Claimed

Research agenda aimed at developing methods for constructing powerful, easily interpretable world-models.

Led byThane Ruthenis, and others
Endorsed by

The Blind Spots in AI Risk Assessment

Team?
ProjectThink TankResearchGovernanceFundraisingClaimed

AI risk assessment currently checks few threat models and doesn't compose them into the aggregate risk that matters. We'll build a tool mapping what frontier system cards cover and omit, plus a paper on what the assessments miss.

Led byRichard Mallah, and others
Endorsed by
+1

COMPASS: AI X-Risk Track

Team?
ProjectTrainingGovernanceX-RiskFundraisingClaimed

COMPASS (Capacity-Oriented Mentorship for Public Administrators on AI Safety & Strategy): AI X-Risk Track trains sitting Global South government officials to manage AI x-risk, so their governments are part of preventing it.

Led byOlin Thakur
Endorsed by
+2

Investigating Chain-of-Thought Faithfulness in AI Reasoning

Team?
ProjectIndividualResearchInterpFundraisingClaimed

Probe an open-weight model’s activations under biasing/cue conditions to test whether chain-of-thought explanations match internal reasoning, releasing paper, code, and datasets.

Led byDaniel Zhang
Endorsed by
+1

Flourishing Futures Fellowship

Team?
ProjectTrainingX-RiskFundraisingManifundClaimed

A scalable fellowship training researchers to develop interventions for achieving high-value long-term futures

Led byJordan Arel
Endorsed by

Belief state geometry of language model personas

Team?
ProjectResearchInterpFundraisingClaimed

A formal, testable account of LLM persona selection as Bayesian inference, validated against model internals, so labs can monitor and steer model personas during post-training and deployment

Led byLogan Graves
Endorsed by

Detailed Threat Models for AI-Enabled Extreme Power Concentration

Team?
ProjectIndividualResearchGovernanceFundraisingClaimed

Detailed models of how AI could allow a small set of actors to gain a decisive strategic advantage over the rest of the world: concrete pathways, required capabilities, quantified likelihoods, and the defenses that bind them.

Led byAmritanshu Prasad
Endorsed by
+1

Surrogate base model for Mechanistic Interpretability

Team?
ProjectResearchInterpFundraisingClaimed

Creating a reference model for mechanistic interpretability without assuming that at auditing time we have a safe model to compare the suspicious model against.

Led byRaffaello Fornasiere
Endorsed by

A Mechanistic Analysis of Activation Verbalizers: Promise and Risks

Team?
ProjectResearchInterpFundraisingClaimed

Mechanistically analyze how activation verbalizers use target-model activation concepts (e.g., cyclic day-of-week representations) via PCA/DAS/patching, explain cross-family failures, and improve verbalizers.

Led byTung-Yu Wu
Endorsed by

BASE - Bangalore AI Safety Exchange

Team?
ProjectHubCommunityTechnical SafetyFundraisingClaimed

A Physical Community Hub for AI Safety in Bangalore,India to build long-term AI safety careers, host multiple AI safety fellowships, career events and build a community that raises the long-term impact and value of AI

Led byAsh Singh, and others
Endorsed by
+1

Tight PAC-Bayes Generalisation Guarantees Across Frontier LLM Safety Monitoring Deployment Settings

Team?
ProjectResearchVerificationOversightFundraisingClaimed

Compression-based PAC-Bayes certification for frontier-scale LLM safety monitors, deployment setting shift, and modern post-training.

Led byTim G. J. Rudner, and others
Endorsed by

Humans in Control

Team?
ProjectNetworkCommunityGovernanceFundraisingClaimed

Humans in Control (HIC) is a nonpartisan grassroots advocacy organization focused on AI safeguards.

Led byVael Gates
Endorsed by

Formally verified autoresearch for theoretical mech interp

Team?
ProjectResearchInterpFundraisingClaimed

LLM agents collaborate to discover and formally verify theorems about the internal computations of transformers, beginning with a simple pilot question: how many attention heads are needed to represent a Boolean function?

Led byKarthik Viswanathan
Endorsed by

DC Mini-Conference 2.0 (DCMC 2.0)

Team?
ProjectConferenceCommunityGovernanceFundraisingClaimed

Running a conference in DC for promising AI safety university students interested in policy to network, learn, and be exposed to the DC ecosystem.

Led bySeth Lifland
Endorsed by

Survey of the American public on AI futures.

Team?
ProjectResearchDemocratic AIFundraisingClaimed

TLDR: A representative survey with Yougov of the American public on questions about AI futures, including space governance, successionism, values.

Led byJason Hausenloy
Endorsed by
+1

Fragile Alignment

Team?
ProjectResearchEvalsTechnical SafetyFundraisingClaimed

Build an open-source platform of model organisms and agentic sandboxes to iteratively test mitigations for emergent misalignment via white-box probes, trigger tests, and evaluation-awareness checks.

Led byMatteo Leonesi
Endorsed by

Separatrix

Team?
ProjectResearchEvalsFundraisingClaimed

Empirical research to create conditions for cooperative strategies to dominate adversarial ones among a broad swath of near-future AIs - in the narrow window this work is still possible.

Led byJai Dhyani
Endorsed by

Superintelligence Inc. Strategy Game

Team?
ProjectCommsX-RiskFundraisingClaimed

Browser based game in the style of Plague Inc where players act as a rogue AI attempting to escape human control. Intended to give lay audiences a grounded understanding of how ASI x-risk could play out.

Led byConnor Heaton
Endorsed by

Mapping AI

Team?
ProjectPlatformToolingGovernanceFundraisingClaimed

This grant would help us maintain and scale Mapping AI, an open-source stakeholder map of the people and organizations with the potential to shape U.S. AI policy.

Led byAnushree Chaudhuri, and others
Endorsed by

Invisible Hands: Measuring Agent Steering of Human Researchers

Team?
ProjectResearchOversightFundraisingClaimed

How do the options agents present to human researchers alter which research paths are explored, and do steered researchers notice or feel less in control?

Led byTrevor Lohrbeer
Endorsed by

Unlearning in genomic and protein language models

Team?
ProjectResearchToolingBiosecurityFundraisingClaimed

Implementing different types of unlearning methods for genomic and protein language models to remove sensitive biological information (e.g. pathogen virulence) while preserving predictive performance and scientific utility.

Led byIlias Georgakopoulos-Soares
Endorsed by

Lisa Intel: A Global Runtime Governance Layer for Autonomous AI

Team?
ProjectPlatformToolingControlFundraisingClaimed

Building an independent runtime governance layer that enables organizations to deploy autonomous AI with authorization, human oversight, auditability and cross-model governance.

Led byPedro Bentancour Garin
Endorsed by
+2

Shaping AI’s « safe culture »

Team?
ProjectResearchGovernanceStandardsFundraisingClaimed

Identify what AI early risks or failures could be reported despite strategic rivalries towards a « safe culture » shaped after aviation

Led byBenoît Larrouturou
Endorsed by
+1

Adapative circuit tracing for test-time interpretability

Team?
ProjectIndividualResearchInterpFundraisingClaimed

A training methodology, and transcoder sets that allow to leverage heavy-weight interpretability methods, but made more lightweight for test-time analysis.

Led byAlexandre Doukhan
Endorsed by

Closing the Talent Gaps in AI: Webinars, Videos and Peer-Support

Team?
ProjectMediaCommunityGovernanceFundraisingClaimed

A webinar series and peer-support community helping people identify and move into high-impact talent gaps careers and projects in AI safety and governance.

Led byKarolina Gruzel
Endorsed by
+2

Developing a new science of agency

Team?
ProjectIndividualResearchAgent FoundationsFundraisingClaimed

Naturalizing theoretical alignment by initiating the development of a scientific theory that is capable of making falsifiable empirical claims about agents in general, including humans, AGI, and ASI, and thereby about alignment.

Led byBen Auer, and others
Endorsed by
+2

Adversarially robust workload classification using hardware sensors

Team?
ProjectResearchSecurityFundraisingClaimed

I've previously developed a classifier that distinguishes training and inference based on Nvidia software telemetry. This project will achieve that using physical sensors, making the system more secure.

Led byRobi Rahman, and others
Endorsed by

AI Agent Overeagerness as a Problem in Safety-Critical Scenarios

Team?
ProjectResearchControlFundraisingClaimed

We want to investigate how AI agent overeagerness can backfire when exhibited in safety-critical scenarios.

Led byThilo Hagendorff
Endorsed by

Safeguarding open-weight genomic foundation models through weight lock

Team?
ProjectResearchToolingBiosecurityFundraisingClaimed

Safeguarding open-weight genomic foundation models through weight lock against adversarial finetuning

Led byAlexandros Tzanakakis
Endorsed by

Does the thinking behind data selection shape a model's character?

Team?
ProjectResearchEvalsIndividualFundraisingClaimed

We test whether the structure behind the selection matters: one model, three diets (industry standard, our protocol, random), open evals, published either way.

Led byKatja Gorlinski
Endorsed by
+2

AI Safety Unconference 2026

Team?
ProjectConferenceCommunityTechnical SafetyFundraisingClaimed

A participant-led unconference gathering the global AI safety community, online and at local sites, Nov 20-22, 2026.

Led byOrpheus Lummis
Endorsed by
+1

Originality and Decomposition of Generalization.

Team?
ProjectResearchEvalsFundraisingClaimed

Perform research to rigorously elucidate and quantify generalization versus memorization, and examine evidence of originality in LLMS.

Led byAri Spiesberger
Endorsed by

London Interface for AI Safety (LIAISE)

Team?
ProjectNetworkTrainingField-BuildingFundraisingClaimed

Liaise helps close the generalist talent gap in AI safety by helping motivated non‑technical people enter with fluency, alignment, and a trusted network.

Led byRohan Barad
Endorsed by
+1

“La P8stée” (2027) starring Bobcat and P8stie

Team?
ProjectMediaCommsX-RiskFundraisingClaimed

It's La Jetée (1962), the time-traveling film later adapted as 12 Monkeys (1995), but it's about X-risk and inspired by AI 2027 and it's directed by, and starring, myself and @p8stie. Ergo, La P8stée

Led byScott Ellis
Endorsed by
+1

Horizon AGI - AI Safety France Pilot

Team?
ProjectNetworkField-BuildingGovernanceFundraisingClaimed

Launching a French AI safety community through university outreach and local events to connect students, researchers, practitioners, and policymakers.

Led byArthur DAVID
Endorsed by
+1

AI Welfare Evaluation Panel: A validation study and pilot

Team?
ProjectResearchAI WelfareFundraisingClaimed

An independent, multi-lineage panel of AI models, tested on whether it can identify welfare concerns in AI evaluations with opinions, dissents, and proposed modifications published in a public registry

Led byJuliana Grant
Endorsed by

Project AI Kavach : AI safety for builders fellowship

Team?
ProjectTrainingToolingTechnical SafetyFundraisingClaimed

A 3 / 6 month builders fellowship where mentor-builder pods ship practical AI safety tools, not papers.

Led byAnkur Pandey, and others
Endorsed by

Evaluating Legal Gradual Disempowerment Risk

Team?
ProjectResearchEvalsFundraisingClaimed

An evaluation suite to identify a model’s legal values (e.g., anti-tech-regulation) relative to well-known actors (e.g., Ruth Bader Ginsburg) and an assessment of how language in a model’s constitution impacts the extent to which these values are human-aligned.

Led byMagnus Saebo, and others
Endorsed by

Is AI safety research reproducible?

Team?
ProjectResearchTechnical SafetyFundraisingClaimed

A SCORE-style computational reproducibility audit of empirical AI safety research that estimates the field's base rate of reproducibility and generates a taxonomy of its failure modes.

Led byJordan Suchow
Endorsed by

Mapping Persona Contamination in the Training Pipeline

Team?
ProjectIndividualResearchEvalsFundraisingClaimed

Map how undesired behaviors can silently spread between models during training, and which parts of the pipeline have the greatest risk — starting with the feedback processes used to align models.

Led byLily Wen
Endorsed by

PauseAI Australia attendance PauseCon London

Team?
ProjectNetworkTrainingGovernanceFundraisingClaimed

Fly two volunteer leaders of PauseAI Australia to PauseCon London, to bring UK's learnings home and train our all-volunteer chapter

Led byPeter Horniak
Endorsed by
+1

Preparing for the future of AI by thinking beyond Transformers

Team?
ProjectIndividualResearchTechnical SafetyFundraisingClaimed

A project about researching radical methods to both advance and prepare countermeasures for what is to come after transformers.

Led byDarsh M
Endorsed by
+1

Oceania AI Safety & Young Talent Conference

Team?
ProjectConferenceCommunityGovernanceFundraisingClaimed

Hosting a full day conference based in Sydney, Australia where young, aspiring students in senior high school and university interested in AI Safety can connect and share ideas

Led byWilliam Wu
Endorsed by

Minister for Human Autonomy

Team?
ProjectAdvocacyGovernanceFundraisingClaimed

Research and plan advocacy for creating a UK Minister for Human Autonomy, including foundational research, project management, and drafting policy recommendations to safeguard human autonomy amid AI.

Led byJoshua Krook
Endorsed by

Using dynamical systems theory approach to identify dishonesty in LLMs

Team?
ProjectIndividualResearchDeceptionFundraisingClaimed

Identifying reasoning pathologies/disingenuous behavior in reasoning traces based on activation dynamics rather than apparent semantics

Led byJosiah Kratz
Endorsed by
+1

AI Safety Social Media Gap Analysis

Team?
ProjectIndividualCommsX-RiskFundraisingClaimed

Auditing the social media presence of the top 25 AI safety organizations across major platforms to quantify the public communication gap and publishing a full gap analysis, then giving recommendations to the orgs for improvement.

Led byJude Williams
Endorsed by

Alien Cub Productions

Team?
ProjectMediaCommsX-RiskFundraisingClaimed

Alien Cub is a production company focused on stories that inform a wide audience about risks from transformative AI.

Led byKeiran Harris
Endorsed by

Does the chain-of-thought actually drive the answer?

Team?
ProjectResearchEvalsOversightFundraisingClaimed

A Bayesian causal auditor that quantifies chain-of-thought faithfulness while accounting for hidden confounding in shared-network language models.

Led byMaksim Silchenko
Endorsed by
+1

"Out of this Box! - The Last Musical (Written by Humans)"

Team?
ProjectMediaCommsTechnical SafetyFundraisingClaimed

We have validated the artistic value of our show; this grant would test whether it can become a scalable and repeatable form of AI-safety outreach.

Led byStephan Wäldchen
Endorsed by

Sydney AI Safety Space - General Funding

Team?
ProjectHubCommunityTechnical SafetyFundraisingClaimed

The Sydney AI Safety Space is a co-working hub in Sydney, Australia, offering free office space for people working in the field of AI safety.

Led byHunter Jay
Endorsed by