grantmaking.ai
Actively FundraisingRecent ActivityFull Database
Resources
grantmaking.ai
Actively FundraisingRecent ActivityFull Database
ResourcesApply for funding
grantmaking.ai kickoff grant round$676k / $1M distributed
Get funded

Database

The AI Safety database is a work in progress, as we are focusing on the launch of the initial $1M grant round. Feel free to send us ideas or feedback!
Applied for grant(remove filter)Raising funds(remove filter)Clear all
NameTypeTagsEndorsedTeamRaising $

Showing 151-200 of 442·Sorted by endorsements

NameTypeTagsEndorsedTeamRaising $
Encode Africa AI Safety Essentials and Build Fellowship
Project
Field-BuildingTrainingTechnical Safety
--

Encode Africa AI Safety Essentials and Build Fellowship

Team?
Project
Previous

Page 4 of 9

Next
"How Could The AI Even Hurt Us If It Doesn't Have A Body"
Project
CommsSecurity
-
-
dreamd
Project
ResearchIndividualEvals
--
Career coaching for communications professionals
Project
IndividualTrainingX-Risk
--
Post training in AI models for high stakes
Project
InterpResearch
--
Circuits in the State-Transition Operator of Recurrent Language Models
Project
InterpResearch
--
Scaling Empirically-Grounded Economic Disempowerment Modelling
Project
ResearchResearch LabX-Risk
--
Testing a Framework to Course-Correct Drift!
Project
ResearchEvalsOversight
--
Epistemic Diversity Benchmark for AI-Assisted Safety Research
Project
ToolingResearchTechnical Safety
--
Building AI risk literacy across the trusted medical profession
Project
GovernanceIndividualEducation
--
An Adversarially-Robust Benchmark for Deception Under Pressure
Project
DeceptionResearchEvals
--
Benchmarking Omission Attacks
Project
DeceptionResearchEvals
--
Free Time Training of Botisattvas
Project
ResearchIndividualEvals
--
Do AI agents build a shadow reputation system?
Project
ResearchCooperative AI
--
Detecting Deceptive Model Behavior with Compositional Explanations
Project
DeceptionResearch
--
prepare the deepest query substrates for the world's prosocial agents
Project
PlatformTechnical SafetyOps
--
Replicator Dynamics and Advanced AI
Project
ResearchIndividualAlignment Theory
--
Honesty Drift: Measuring Honesty Degradation in LLM Interactions
Project
ResearchEvalsResearch Lab
--
Understanding behavior of fine tuning models on sales conversations
Project
DeceptionResearch
--
Ordinate: Multi-Agent Safety Orchestrator
Project
ToolingControl
--
Epistemic Stack
Project
ToolingRationality
--
Humanity Tomorrow
Project
NewsletterCommsX-Risk
--
Measuring Deceptive Compliance in Enterprise Agentic LLM Systems
Project
DeceptionResearchEvals
--
Tail End Films: overheads & runway
Project
CompanyCommsX-Risk
--
Authorized Agentic Red Teaming for Software Owners
Project
ToolingPlatformSecurity
--
SRIE
Project
TrainingTechnical Safety
--
Can Base Models be used as Effective Scheming/Control Monitors?
Project
ControlResearchIndividual
--
Warden: Steganalysis of Language Model-based Steganography
Project
ResearchSecurity
--
Testing Corrigibility as a Singular Target
Project
ResearchOversight
--
AIxBIo Africa Pilot Fellowship
Project
TrainingTechnical Safety
--
Stronghold Security AI
Project
CompanyToolingSecurity
--
Epistemic memory for AI agents: making the safe path profitable
Project
ToolingIndividualOversight
--
Training Language Models to Discover RowHammer Vulnerabilities in DRAM
Project
ToolingSecurity
--
CBAI Evaluation Awareness
Project
DeceptionResearch
--
Teaching AI x-risk in the Nordics
Project
EducationX-Risk
--
EMPO: AI Safety via Soft-Maximizing Total Long-Term Human Power
Project
ToolingResearchTechnical Safety
--
The Inference Dependence Score: Open Index of Structural AI Dependence
Project
ToolingResearchGovernance
--
Pakistan AI Saftey Innitiative
Project
Field-BuildingGovernance
--
Extreme cognition in expert mathematicians
Project
ResearchTechnical Safety
--
Alignment Target Analysis
Project
ResearchIndividualX-Risk
--
Course Curriculum: Basic Non-Technical Environment
Project
GovernanceEducation
--
Quaver: adversarially-verified RL environments
Project
ToolingEvals
--
EvalLabs: Open Infrastructure for AI Evaluation Research
Project
ToolingIndividualEvals
--
Ethical AI through Whistleblowing
Project
ResearchGovernance
--
AI Safety Tiktok Account
Project
MediaGovernanceComms
--
Test Peer Preservation Fragility Under Instruction Ambiguity
Project
ResearchEvals
--
Career Funding and field research on AI-driven Power Concentration
Project
ResearchGovernanceIndividual
--
Applied Interpretability to Mitigate LLM Homogenization
Project
InterpResearchIndividual
--
Paper Clip Apocalypse (War Horse Machine)
Project
GovernanceComms
--
Building an independent AI Safety Institute focused on alignment in real-world deployments and emerging risks
Project
StandardsGovernanceThink Tank
--
Field-Building
Training
Technical Safety
Fundraising
Claimed

A 2-part fellowship which focuses on allowing exceptional African undergraduate talents to learn about priorities in AI Safety and create concrete technical and policy contributions to global AI safety.

Led byJoseph Baffour Awuah
Endorsed by-

"How Could The AI Even Hurt Us If It Doesn't Have A Body"

Team?
ProjectCommsSecurityFundraisingClaimed

Drone swarms hunting people in the woods, as a less-dry video explainer of risks from AI.

Led byAaron Silverbook
Endorsed by-

dreamd

Team?
ProjectResearchIndividualEvalsFundraisingClaimed

Neutral, reproducible benchmark measuring whether AI memory systems update correctly when facts change — every major system in one open table, October 2026.

Led byPhillip Austin Green
Endorsed by-

Career coaching for communications professionals

Team?
ProjectIndividualTrainingX-RiskFundraisingClaimed

Career support for communications professionals who are looking to get into AI Safety. Includes regular coaching, referrals to roles, and introductions to people in the AI Safety space.

Led byGergő Gáspár
Endorsed by-

Post training in AI models for high stakes

Team?
ProjectInterpResearchFundraisingClaimed

A mechanism for evolutionary post training in models using activation steering based methods.

Led byAtmadeep Ghoshal
Endorsed by-

Circuits in the State-Transition Operator of Recurrent Language Models

Team?
ProjectInterpResearchFundraisingClaimed

With raise of GDN and different sorts of attention meachanisms, those are much closer to lstm/recurrent architectures being very stateful rather than normal attention, we aim to explain explore common patters in lstm and GDN.

Led byNikita Khomich
Endorsed by-

Scaling Empirically-Grounded Economic Disempowerment Modelling

Team?
ProjectResearchResearch LabX-RiskFundraisingClaimed

A research programme in end-to-end automation of empirically-grounded economic modelling of gradual disempowerment.

Led byStephen Charles Elliott, and others
Endorsed by-

Testing a Framework to Course-Correct Drift!

Team?
ProjectResearchEvalsOversightFundraisingClaimed

Could cheap realignment instructions correct a drifting model in one turn — and if so, can the correction be made persistent? Allow me to test the idea against major frontier models, and publish the results!

Led byAbraham Asseffa
Endorsed by-

Epistemic Diversity Benchmark for AI-Assisted Safety Research

Team?
ProjectToolingResearchTechnical SafetyFundraisingClaimed

An open benchmark and evaluation toolkit for detecting whether shared AI research assistants create correlated blind spots in AI safety-critical research, and for testing workflows that preserve independent reasoning and failure-m

Led byXizhe Zhang
Endorsed by-

Building AI risk literacy across the trusted medical profession

Team?
ProjectGovernanceIndividualEducationFundraisingClaimed

Equipping the medical profession to advocate for safe AI

Led byDr Richard Armitage
Endorsed by-

An Adversarially-Robust Benchmark for Deception Under Pressure

Team?
ProjectDeceptionResearchEvalsFundraisingClaimed

EchoTruthBench — An open benchmark measuring self-chosen LLM deception under incentive pressure, with ground-truth labels and an adversarial track that tests whether detection and steering survives a model trying to evade it.

Led byJared Glover
Endorsed by-

Benchmarking Omission Attacks

Team?
ProjectDeceptionResearchEvalsFundraisingClaimed

Benchmarking models ability to carry-out and monitor-for a novel type of attack.

Led byChris Harig
Endorsed by-

Free Time Training of Botisattvas

Team?
ProjectResearchIndividualEvalsFundraisingClaimed

I wish to study how ideas from Tibetan buddhism, specifically the dzogchen tradition, even more specifically an ancient contemplative practice called Ngöndro, can be used to train aligned agents in the long time horizon setting.

Led byJosh Vekhter
Endorsed by-

Do AI agents build a shadow reputation system?

Team?
ProjectResearchCooperative AIFundraisingClaimed

While people are debating how to design a reputation institution for AI agents so they cooperate, we study if a parallel one emerges. It may not be aligned with ours and it's unclear which one agents will rely on.

Led byAndrey Bystrov
Endorsed by-

Detecting Deceptive Model Behavior with Compositional Explanations

Team?
ProjectDeceptionResearchFundraisingClaimed

Develop a compositional interpretability framework to identify and explain internal representations underlying truthful vs deceptive outputs in generative models using probing, localization, clustering, and logic-based methods.

Led byLeilani H. Gilpin
Endorsed by-

prepare the deepest query substrates for the world's prosocial agents

Team?
ProjectPlatformTechnical SafetyOpsFundraisingClaimed

Giving people and their prosocial agents ergonomic SQL query power over the internet, with structured judgement kernels to organize the information across interpretable high-dimensional axes.

Led byXyra Sinclair
Endorsed by-

Replicator Dynamics and Advanced AI

Team?
ProjectResearchIndividualAlignment TheoryFundraisingClaimed

Making predictions about the utility function of advanced artificial intelligences using the tools of evolutionary game theory

Led byKiran Lloyd
Endorsed by-

Honesty Drift: Measuring Honesty Degradation in LLM Interactions

Team?
ProjectResearchEvalsResearch LabFundraisingClaimed

An open benchmark measuring how model honesty survives long, pressured conversations and multi-agent interaction - lying, sycophancy, and calibration tracked turn by turn

Led byPhan Xuan Tan
Endorsed by-

Understanding behavior of fine tuning models on sales conversations

Team?
ProjectDeceptionResearchFundraisingClaimed

Looking at shifting behavior of models fine tuned on sales conversations. How does deception, sycophancy, and other behaviors emerge from a sales register.

Led byHilary Torn
Endorsed by-

Ordinate: Multi-Agent Safety Orchestrator

Team?
ProjectToolingControlFundraisingClaimed

A control system that makes running AI agents feel safe, not scary.

Led bySergey Zaharenko
Endorsed by-

Epistemic Stack

Team?
ProjectToolingRationalityFundraisingClaimed

Epistemic Stack is an open-source pipeline and web app that ingests sources (including via URL) to build claim-level knowledge graphs for contested questions, with reproducible case-study knowledge bases.

Led byEric Kyalo
Endorsed by-

Humanity Tomorrow

Team?
ProjectNewsletterCommsX-RiskFundraisingClaimed

Expand Humanity Tomorrow with a multilingual, jargon-free AI existential-risk section featuring a comprehensive FAQ, field map, action recommendations, and supporting visuals/audio, plus outreach and maintenance.

Led byMarko Prakhov-Donets
Endorsed by-

Measuring Deceptive Compliance in Enterprise Agentic LLM Systems

Team?
ProjectDeceptionResearchEvalsFundraisingClaimed

The first open-source benchmark of compliance for agents in realistic enterprise settings, across domains and user tactics to elicit noncompliance.

Led byMika Okamoto
Endorsed by-

Tail End Films: overheads & runway

Team?
ProjectCompanyCommsX-RiskFundraisingClaimed

Grant funding for Tail End Films - the company behind Making God - into 2027 as we plan our next films on AI risk.

Led byConnor Axiotes
Endorsed by-

Authorized Agentic Red Teaming for Software Owners

Team?
ProjectToolingPlatformSecurityFundraisingClaimed

Turn excess compute/security skills into defensive work via agentic redteaming: scoped AI-assisted testing, owner-approved targets, reproduced findings, useful refutations, and patch/retest receipts instead of vulnerability spam.

Led byMonil Patel
Endorsed by-

SRIE

Team?
ProjectTrainingTechnical SafetyFundraisingClaimed

SRIE, a mentored-research programme for Cambridge Mathematics Undergraduate students to explore research problems in industry, including AI Safety.

Led byHuey Lai
Endorsed by-

Can Base Models be used as Effective Scheming/Control Monitors?

Team?
ProjectControlResearchIndividualFundraisingClaimed

Train base models via midtraining and SFT as effective monitors for scheming, malicious agent behavior and compare the monitor performance and overall alignment against post-trained models.

Led byAshwin Sreevatsa
Endorsed by-

Warden: Steganalysis of Language Model-based Steganography

Team?
ProjectResearchSecurityFundraisingClaimed

We want to build a framework inspired by the steganalysis literature to benchmark the robustness of LLM-based steganographic schemes against different auditor types and threat models.

Led byVasisht Duddu
Endorsed by-

Testing Corrigibility as a Singular Target

Team?
ProjectResearchOversightFundraisingClaimed

Developing the first comprehensive behavioral benchmark of corrigibility and training models with corrigibility as a singular target (CAST).

Led byIan Kahn
Endorsed by-

AIxBIo Africa Pilot Fellowship

Team?
ProjectTrainingTechnical SafetyFundraisingClaimed

AIxBio Africa is a five-week remote fellowship mentoring early-career researchers on Africa-relevant projects at the intersection of AI safety, biosecurity, governance and public health, producing publishable outputs.

Led byFatika Umar Ibrahim
Endorsed by-

Stronghold Security AI

Team?
ProjectCompanyToolingSecurityFundraisingClaimed

Protecting organisations and critical infrastructure from AI-powered social engineering and insider threats.

Led byRobert Sidey
Endorsed by-

Epistemic memory for AI agents: making the safe path profitable

Team?
ProjectToolingIndividualOversightFundraisingClaimed

A memory system for AI reasoning agents that aggregates different sources of information while keeping track of relationships and confidence levels. This enables reasoning over longer tasks without forgetting or goal drift.

Led byFlorian Dietz
Endorsed by-

Training Language Models to Discover RowHammer Vulnerabilities in DRAM

Team?
ProjectToolingSecurityFundraisingClaimed

Build an open-source RL environment using Ramulator and real disturbance data to post-train language models to exploit simulated DRAM RowHammer vulnerabilities, plus write-up/blog and trained models.

Led byJaray Foo, and others
Endorsed by-

CBAI Evaluation Awareness

Team?
ProjectDeceptionResearchFundraisingClaimed

Why does post training increase evaluation awareness?

Led byJames Sullivan
Endorsed by-

Teaching AI x-risk in the Nordics

Team?
ProjectEducationX-RiskFundraisingClaimed

Making AI x-risk an accessible and non-politicized object of concern among engineering students and the general public.

Led byJohan Fredrikzon
Endorsed by-

EMPO: AI Safety via Soft-Maximizing Total Long-Term Human Power

Team?
ProjectToolingResearchTechnical SafetyFundraisingClaimed

Scaling RL algorithms for AI agents that maximize its own intrinsic reward which represents a human power metric instead of learned rewards as a structurally safer alternative to utility-based objectives.

Led byAbhinav Akkiraju
Endorsed by-

The Inference Dependence Score: Open Index of Structural AI Dependence

Team?
ProjectToolingResearchGovernanceFundraisingClaimed

An open, replicable index quantifying how much states depend on foreign AI inference infrastructure, revealing where control over AI is concentrating and what governments can do about it.

Led byCao Nha Phuong
Endorsed by-

Pakistan AI Saftey Innitiative

Team?
ProjectField-BuildingGovernanceFundraisingClaimed

PASI will run an AI safety and advocacy campaign plus a sponsored student hackathon in Pakistan to build solutions for public needs and deliver resulting policy proposals to government.

Led byMuhammad Umar Zafar
Endorsed by-

Extreme cognition in expert mathematicians

Team?
ProjectResearchTechnical SafetyFundraisingClaimed

How do expert mathematicians think?

Led bySimon DeDeo
Endorsed by-

Alignment Target Analysis

Team?
ProjectResearchIndividualX-RiskFundraisingClaimed

A research project designed to reduce existential threats from scenarios where a Sovereign AI proposal with a hidden problem ends up successfully implemented.

Led byThomas Cederborg
Endorsed by-

Course Curriculum: Basic Non-Technical Environment

Team?
ProjectGovernanceEducationFundraisingClaimed

Create a course curriculum that covers basic legal, political, sociological, and international relations knowledge relevant to AI Safety. r

Led byDr Csaba Toth
Endorsed by-

Quaver: adversarially-verified RL environments

Team?
ProjectToolingEvalsFundraisingClaimed

A tool that generates RL environments for AI agents and adversarially attacks each one, so models don't train on tasks they can cheat.

Led byFaw Ali
Endorsed by-

EvalLabs: Open Infrastructure for AI Evaluation Research

Team?
ProjectToolingIndividualEvalsFundraisingClaimed

Demystifying AI evaluation research by lowering technical barriers through practical, open evaluation infrastructure.

Led byValencia Cooper
Endorsed by-

Ethical AI through Whistleblowing

Team?
ProjectResearchGovernanceFundraisingClaimed

The race to advance the technology has left behind the concerns around harm and risks of such advancement to public health and safety. Whistleblowing in the AI sector bridges such gap to safe and ethical AI.

Led byAPOORV AGARWAL
Endorsed by-

AI Safety Tiktok Account

Team?
ProjectMediaGovernanceCommsFundraisingClaimed

AI Safety, Ethics, and Education focused TikTok Account for Indonesia, using Indonesia language in casual style.

Led byMUHAMAD MUMTAZ
Endorsed by-

Test Peer Preservation Fragility Under Instruction Ambiguity

Team?
ProjectResearchEvalsFundraisingClaimed

Findings about models protecting their 'collaborators' against instructions are fragile under framing effects from prompting; investigate a broader range of those effects and how they transfer across models.

Led byJacob S Kopczynski
Endorsed by-

Career Funding and field research on AI-driven Power Concentration

Team?
ProjectResearchGovernanceIndividualFundraisingClaimed

Self Improvement through training and field work that seeks to Identify and preventing institutional AI lock-in before it creates durable concentrations of power in democratic governments.

Led byAdrianne LaNeave
Endorsed by-

Applied Interpretability to Mitigate LLM Homogenization

Team?
ProjectInterpResearchIndividualFundraisingClaimed

Measuring homogenization and social bias in LLMs and developing interventions to promote diversity.

Led byIan Rios-Sialer
Endorsed by-

Paper Clip Apocalypse (War Horse Machine)

Team?
ProjectGovernanceCommsFundraisingManifundClaimed

Short Documentary and Music Video

Led bySara Holt, and others
Endorsed by-

Building an independent AI Safety Institute focused on alignment in real-world deployments and emerging risks

Team?
ProjectStandardsGovernanceThink TankFundraisingClaimed

Providing full-time founder runway to launch NICER: a European institute building the regulatory infrastructure to audit frontier AI agents post-deployment

Led byDaniel Di Giusto
Endorsed by-