grantmaking.ai
Actively FundraisingRecent ActivityFull Database
Resources
grantmaking.ai
Actively FundraisingRecent ActivityFull Database
ResourcesApply for funding
grantmaking.ai kickoff grant round$278k / $1M distributed
Get funded

Database

The AI Safety database is a work in progress, as we are focusing on the launch of the initial $1M grant round. Feel free to send us ideas or feedback!
Entity: Grantmaker(remove filter)Entity: Project(remove filter)Clear all
NameTypeTagsEndorsedTeamRaising $

Showing 101-150 of 1.1k·Sorted by endorsements

NameTypeTagsEndorsedTeamRaising $
Invisible Hands: Measuring Agent Steering of Human Researchers
Project
ResearchOversight
-

Invisible Hands: Measuring Agent Steering of Human Researchers

Team?
Project
Previous

Page 3 of 23

Next
Unlearning in genomic and protein language models
Project
ResearchToolingBiosecurity
-
Lisa Intel: A Global Runtime Governance Layer for Autonomous AI
Project
PlatformToolingControl
+2
-
Shaping AI’s « safe culture »
Project
ResearchGovernanceStandards
+1
-
Adapative circuit tracing for test-time interpretability
Project
IndividualResearchInterp
-
Closing the Talent Gaps in AI: Webinars, Videos and Peer-Support
Project
MediaCommunityGovernance
+2
-
Developing a new science of agency
Project
IndividualResearchAgent Foundations
+2
-
Adversarially robust workload classification using hardware sensors
Project
ResearchSecurity
-
AI Agent Overeagerness as a Problem in Safety-Critical Scenarios
Project
ResearchControl
-
Safeguarding open-weight genomic foundation models through weight lock
Project
ResearchToolingBiosecurity
-
Does the thinking behind data selection shape a model's character?
Project
ResearchEvalsIndividual
+2
-
Safe AI Germany (SAIGE)
Project
NetworkCommunityTechnical Safety
--
Mox: 2026 Fundraiser
Project
HubCommunityEA
--
Ambitious AI Alignment Seminar
Project
TrainingTechnical Safety
---
Research Staff for AI Safety Research Projects
Project
Research LabResearchTechnical Safety
--
Lyptus Research: Bridge Funding
Project
Research LabResearchTechnical Safety
---
Exploring feature interactions in transformer LLMs through sparse autoencoders
Project
IndividualResearchInterp
--
Investigating and informing the public about the trajectory of AI
Project
Think TankResearchForecasting
--
AI Safety Unconference 2026
Project
ConferenceCommunityTechnical Safety
+1
-
Help Apart Expand Global AI Safety Research
Project
Research LabTrainingTechnical Safety
--
Joseph Bloom - Independent AI Safety Research
Project
IndividualResearchInterp
--
Avoiding Incentives for Performative Prediction in AI
Project
IndividualResearchTechnical Safety
--
AI Governance Exchange (focus on China, AI safety), Seed Funding
Project
Think TankCommunityGovernance
--
Originality and Decomposition of Generalization.
Project
ResearchEvals
-
London Interface for AI Safety (LIAISE)
Project
NetworkTrainingField-Building
+1
-
“La P8stée” (2027) starring Bobcat and P8stie
Project
MediaCommsX-Risk
+1
-
Forecasting - AI Governance Policies
Project
ResearchForecasting
--
Travel grant to attend HAAISS Prague 2026
Project
TrainingTechnical Safety
--
Compute and other expenses for LLM alignment research
Project
ResearchTechnical Safety
--
Train great open-source sparse autoencoders
Project
ResearchInterp
--
Horizon AGI - AI Safety France Pilot
Project
NetworkField-BuildingGovernance
+1
-
Guaranteed Safe AI Seminars 2026
Project
ConferenceCommunityVerification
--
Run five international hackathons on AI safety research
Project
ConferenceField-BuildingTechnical Safety
--
AI safety fieldbuilding in Warsaw, Poland (funding for 1 semester)
Project
Field-BuildingTechnical SafetyCommunity
--
SafePlanBench: evaluating a Guaranteed Safe AI Approach for LLM-based Agents
Project
ResearchTechnical SafetyEvals
--
AI Welfare Evaluation Panel: A validation study and pilot
Project
ResearchAI Welfare
-
Project AI Kavach : AI safety for builders fellowship
Project
TrainingToolingTechnical Safety
-
Evaluating Legal Gradual Disempowerment Risk
Project
ResearchEvals
-
Is AI safety research reproducible?
Project
ResearchTechnical Safety
-
Mapping Persona Contamination in the Training Pipeline
Project
IndividualResearchEvals
-
PauseAI Australia attendance PauseCon London
Project
NetworkTrainingGovernance
+1
-
Concept control in transformers with Sparse Concept Anchoring
Project
IndividualResearchControl
+1
--
Precautionary AGI Governance
Project
IndividualResearchGovernance
--
FragGuard: Cross-Session Malicious Activity Detection for Model APIs
Project
IndividualResearchControl
--
Benchmarking and comparing different evaluation awareness metrics
Project
ResearchEvalsDeception
--
Groundless Alignment Residency 2025
Project
CommunityTechnical SafetyTooling
---
3 Months Career Transition into AI Safety Research & Software Engineering
Project
IndividualResearchEvals
---
Testing and spreading messages to reduce AI x-risk
Project
CommsGovernanceX-Risk
--
Cadenza Labs: AI Safety research group working on own interpretability agenda
Project
Research LabResearchInterp
--
The Deal of the Century: Targeted Persuasion Campaign for a US-China AI Treaty
Project
AdvocacyGovernance
--
Research
Oversight
Fundraising
Claimed

How do the options agents present to human researchers alter which research paths are explored, and do steered researchers notice or feel less in control?

Led byTrevor Lohrbeer
Endorsed by

Unlearning in genomic and protein language models

Team?
ProjectResearchToolingBiosecurityFundraisingClaimed

Implementing different types of unlearning methods for genomic and protein language models to remove sensitive biological information (e.g. pathogen virulence) while preserving predictive performance and scientific utility.

Led byIlias Georgakopoulos-Soares
Endorsed by

Lisa Intel: A Global Runtime Governance Layer for Autonomous AI

Team?
ProjectPlatformToolingControlFundraisingClaimed

Building an independent runtime governance layer that enables organizations to deploy autonomous AI with authorization, human oversight, auditability and cross-model governance.

Led byPedro Bentancour Garin
Endorsed by
+2

Shaping AI’s « safe culture »

Team?
ProjectResearchGovernanceStandardsFundraisingClaimed

Identify what AI early risks or failures could be reported despite strategic rivalries towards a « safe culture » shaped after aviation

Led byBenoît Larrouturou
Endorsed by
+1

Adapative circuit tracing for test-time interpretability

Team?
ProjectIndividualResearchInterpFundraisingClaimed

A training methodology, and transcoder sets that allow to leverage heavy-weight interpretability methods, but made more lightweight for test-time analysis.

Led byAlexandre Doukhan
Endorsed by

Closing the Talent Gaps in AI: Webinars, Videos and Peer-Support

Team?
ProjectMediaCommunityGovernanceFundraisingClaimed

A webinar series and peer-support community helping people identify and move into high-impact talent gaps careers and projects in AI safety and governance.

Led byKarolina Gruzel
Endorsed by
+2

Developing a new science of agency

Team?
ProjectIndividualResearchAgent FoundationsFundraisingClaimed

Naturalizing theoretical alignment by initiating the development of a scientific theory that is capable of making falsifiable empirical claims about agents in general, including humans, AGI, and ASI, and thereby about alignment.

Led byBen Auer, and others
Endorsed by
+2

Adversarially robust workload classification using hardware sensors

Team?
ProjectResearchSecurityFundraisingClaimed

I've previously developed a classifier that distinguishes training and inference based on Nvidia software telemetry. This project will achieve that using physical sensors, making the system more secure.

Led byRobi Rahman, and others
Endorsed by

AI Agent Overeagerness as a Problem in Safety-Critical Scenarios

Team?
ProjectResearchControlFundraisingClaimed

We want to investigate how AI agent overeagerness can backfire when exhibited in safety-critical scenarios.

Led byThilo Hagendorff
Endorsed by

Safeguarding open-weight genomic foundation models through weight lock

Team?
ProjectResearchToolingBiosecurityFundraisingClaimed

Safeguarding open-weight genomic foundation models through weight lock against adversarial finetuning

Led byAlexandros Tzanakakis
Endorsed by

Does the thinking behind data selection shape a model's character?

Team?
ProjectResearchEvalsIndividualFundraisingClaimed

We test whether the structure behind the selection matters: one model, three diets (industry standard, our protocol, random), open evals, published either way.

Led byKatja Gorlinski
Endorsed by
+2

Safe AI Germany (SAIGE)

Team?
ProjectNetworkCommunityTechnical SafetyField-BuildingGovernanceManifund

Germany’s talents are critical to the global effort of reducing catastrophic risks brought by artificial intelligence.

Led byJessica P. Wang, and others
Endorsed by

Mox: 2026 Fundraiser

Team?
ProjectHubCommunityEAIncubatorTechnical SafetyManifund

An incubator & community space in SF; for doers of good and masters of craft

Led byMox
Endorsed by

Ambitious AI Alignment Seminar

Team?
ProjectTrainingTechnical SafetyManifund

One Month to Study, Explain, and Try to Solve Superintelligence Alignment

Led byMateusz Bagiński
-

Research Staff for AI Safety Research Projects

Team?
ProjectResearch LabResearchTechnical SafetyManifund

CAIS seeks funding to hire research staff and cover compute/datasets to develop scalable AI safety evals, robustness/jailbreak defenses, internal control benchmarks, and biohazard knowledge benchmarks/unlearning.

Led byDan Hendrycks
Endorsed by

Lyptus Research: Bridge Funding

Team?
ProjectResearch LabResearchTechnical SafetyManifund

An early-stage AI safety research group based in Sydney, Australia

Led bySean Peters
Endorsed by-

Exploring feature interactions in transformer LLMs through sparse autoencoders

Team?
ProjectIndividualResearchInterpManifund

Develop methods using sparse autoencoders to map and relate transformer features across components, track feature evolution/manifolds, and build scalable circuit-search algorithms via feature interventions.

Led byKunvar Thaman
Endorsed by

Investigating and informing the public about the trajectory of AI

Team?
ProjectThink TankResearchForecastingManifund

Epoch AI seeks $10M/2yrs general support to expand public research on AI trajectories via data tracking (models/hardware/clusters), independent benchmarking, and economic modeling for policy and industry.

Led byEpoch AI
Endorsed by

AI Safety Unconference 2026

Team?
ProjectConferenceCommunityTechnical SafetyFundraisingClaimed

A participant-led unconference gathering the global AI safety community, online and at local sites, Nov 20-22, 2026.

Led byOrpheus Lummis
Endorsed by
+1

Help Apart Expand Global AI Safety Research

Team?
ProjectResearch LabTrainingTechnical SafetyManifund

Incubate AI safety research and develop the next generation of global AI safety talent via research sprints and research fellowships

Led byApart Research
Endorsed by

Joseph Bloom - Independent AI Safety Research

Team?
ProjectIndividualResearchInterpManifund

Trajectory Models and Agent Simulators

Led byJoseph Bloom
Endorsed by

Avoiding Incentives for Performative Prediction in AI

Team?
ProjectIndividualResearchTechnical SafetyManifund

Funds four months to develop theory and experiments showing joint evaluation of multiple predictors can remove incentives for manipulative performative prediction, including extensions to prediction/decision markets.

Led byRubi Hudson
Endorsed by

AI Governance Exchange (focus on China, AI safety), Seed Funding

Team?
ProjectThink TankCommunityGovernanceManifund

Building bridges between Western and Chinese AI governance efforts to address global AI safety challenges.

Led byYuanyuan Sun
Endorsed by

Originality and Decomposition of Generalization.

Team?
ProjectResearchEvalsFundraisingClaimed

Perform research to rigorously elucidate and quantify generalization versus memorization, and examine evidence of originality in LLMS.

Led byAri Spiesberger
Endorsed by

London Interface for AI Safety (LIAISE)

Team?
ProjectNetworkTrainingField-BuildingFundraisingClaimed

Liaise helps close the generalist talent gap in AI safety by helping motivated non‑technical people enter with fluency, alignment, and a trusted network.

Led byRohan Barad
Endorsed by
+1

“La P8stée” (2027) starring Bobcat and P8stie

Team?
ProjectMediaCommsX-RiskFundraisingClaimed

It's La Jetée (1962), the time-traveling film later adapted as 12 Monkeys (1995), but it's about X-risk and inspired by AI 2027 and it's directed by, and starring, myself and @p8stie. Ergo, La P8stée

Led byScott Ellis
Endorsed by
+1

Forecasting - AI Governance Policies

Team?
ProjectResearchForecastingManifund

Samotsvety Forecasting will publish and improve conditional forecasts on an AI training pause and an AI treaty, and draft treaty language and priorities to support international AI governance.

Led byTolga Bilge
Endorsed by

Travel grant to attend HAAISS Prague 2026

Team?
ProjectTrainingTechnical SafetyManifund

Requesting $1.2k to cover remaining travel, visa, and living costs to attend the Human-Aligned AI Summer School (HAAISS) in Prague to transition into technical AI safety research.

Led byAnju Chhetri
Endorsed by

Compute and other expenses for LLM alignment research

Team?
ProjectResearchTechnical SafetyManifund

4 different projects (finding RLHF alignment failures, debate, improving CoT faithfulness, and model organisms)

Led byEthan Josean Perez
Endorsed by

Train great open-source sparse autoencoders

Team?
ProjectResearchInterpManifund

Find the best settings for SAE training we can, then scale across models

Led byTom McGrath

Horizon AGI - AI Safety France Pilot

Team?
ProjectNetworkField-BuildingGovernanceFundraisingClaimed

Launching a French AI safety community through university outreach and local events to connect students, researchers, practitioners, and policymakers.

Led byArthur DAVID
Endorsed by
+1

Guaranteed Safe AI Seminars 2026

Team?
ProjectConferenceCommunityVerificationManifund

Seminars on quantitative/guaranteed AI safety (formal methods, verification, mech-interp), with recordings, debates, and the guaranteedsafe.ai community hub.

Led byOrpheus Lummis
Endorsed by

Run five international hackathons on AI safety research

Team?
ProjectConferenceField-BuildingTechnical SafetyManifund

Six-month support for a Program Manager to organize and execute international AI safety hackathons with Apart Research

Led byEsben Kran Christensen
Endorsed by

AI safety fieldbuilding in Warsaw, Poland (funding for 1 semester)

Team?
ProjectField-BuildingTechnical SafetyCommunityIndividualManifund

Reach the university that trained close to 20% of OpenAI early employees

Led byPiotr Zaborszczyk
Endorsed by

SafePlanBench: evaluating a Guaranteed Safe AI Approach for LLM-based Agents

Team?
ProjectResearchTechnical SafetyEvalsVerificationManifund

Seeking funding to develop and evaluate a new benchmark for systematically assesing safety of LLM-based agents

Led byAgustín Martinez Suñé
Endorsed by

AI Welfare Evaluation Panel: A validation study and pilot

Team?
ProjectResearchAI WelfareFundraisingClaimed

An independent, multi-lineage panel of AI models, tested on whether it can identify welfare concerns in AI evaluations with opinions, dissents, and proposed modifications published in a public registry

Led byJuliana Grant
Endorsed by

Project AI Kavach : AI safety for builders fellowship

Team?
ProjectTrainingToolingTechnical SafetyFundraisingClaimed

A 3 / 6 month builders fellowship where mentor-builder pods ship practical AI safety tools, not papers.

Led byAnkur Pandey, and others
Endorsed by

Evaluating Legal Gradual Disempowerment Risk

Team?
ProjectResearchEvalsFundraisingClaimed

An evaluation suite to identify a model’s legal values (e.g., anti-tech-regulation) relative to well-known actors (e.g., Ruth Bader Ginsburg) and an assessment of how language in a model’s constitution impacts the extent to which these values are human-aligned.

Led byMagnus Saebo, and others
Endorsed by

Is AI safety research reproducible?

Team?
ProjectResearchTechnical SafetyFundraisingClaimed

A SCORE-style computational reproducibility audit of empirical AI safety research that estimates the field's base rate of reproducibility and generates a taxonomy of its failure modes.

Led byJordan Suchow
Endorsed by

Mapping Persona Contamination in the Training Pipeline

Team?
ProjectIndividualResearchEvalsFundraisingClaimed

Map how undesired behaviors can silently spread between models during training, and which parts of the pipeline have the greatest risk — starting with the feedback processes used to align models.

Led byLily Wen
Endorsed by

PauseAI Australia attendance PauseCon London

Team?
ProjectNetworkTrainingGovernanceFundraisingClaimed

Fly two volunteer leaders of PauseAI Australia to PauseCon London, to bring UK's learnings home and train our all-volunteer chapter

Led byPeter Horniak
Endorsed by
+1

Concept control in transformers with Sparse Concept Anchoring

Team?
ProjectIndividualResearchControlInterpManifund

Develop a training-time method for transformers that puts concepts where you can find them, so removal has predictable efficacy and bounded side-effects.

Led bySandy Fraser
Endorsed by
+1

Precautionary AGI Governance

Team?
ProjectIndividualResearchGovernanceManifund

Limited Legal Personhood as a Reversible Safety Instrument

Led byKarsten Brensing
Endorsed by

FragGuard: Cross-Session Malicious Activity Detection for Model APIs

Team?
ProjectIndividualResearchControlManifund

Build an AI control research tool that auto-generates red-team datasets to bypass monitors (e.g., ControlArena) and use it to develop and evaluate new monitor-training control protocols.

Led byLinh Le
Endorsed by

Benchmarking and comparing different evaluation awareness metrics

Team?
ProjectResearchEvalsDeceptionManifund

LLMs often know when they are being evaluated. We’ll do a study comparing various methods to measure and monitor this capability.

Led byJord Nguyen
Endorsed by

Groundless Alignment Residency 2025

Team?
ProjectCommunityTechnical SafetyToolingManifund

Practicing Embodied Protocols that work with Live Interfaces

Led byAditya Arpitha Prasad
Endorsed by-

3 Months Career Transition into AI Safety Research & Software Engineering

Team?
ProjectIndividualResearchEvalsManifund

I'd like to explore a research agenda at the intersection of time horizon model evaluation and control protocols.

Led bySean Peters
Endorsed by-

Testing and spreading messages to reduce AI x-risk

Team?
ProjectCommsGovernanceX-RiskManifund

Educating the general public about AI and risks in most efficient ways and leveraging this to achieve good policy outcomes

Led byAI Safety and Governance Fund
Endorsed by

Cadenza Labs: AI Safety research group working on own interpretability agenda

Team?
ProjectResearch LabResearchInterpManifund

We're a team of SERI-MATS alumni working on interpretability, seeking funding to continue our research after our LTFF grant ended.

Led byCadenza Labs
Endorsed by

The Deal of the Century: Targeted Persuasion Campaign for a US-China AI Treaty

Team?
ProjectAdvocacyGovernanceManifund

Persuading a critical mass of key potential influencers of Trump's AI policy to champion a bold, timely and proper US-China-led global AI treaty

Led byRufo Guerreschi
Endorsed by