grantmaking.ai
Actively FundraisingRecent ActivityFull Database
Resources
grantmaking.ai
Actively FundraisingRecent ActivityFull Database
ResourcesApply for funding
grantmaking.ai kickoff grant round$278k / $1M distributed
Get funded

Database

The AI Safety database is a work in progress, as we are focusing on the launch of the initial $1M grant round. Feel free to send us ideas or feedback!
Raising funds(remove filter)Clear all
NameTypeTagsEndorsedTeamRaising $

Showing 151-200 of 459·Sorted by endorsements

NameTypeTagsEndorsedTeamRaising $
Comprehension Audits to Mitigate Risks from Automated AI Research
Project
StandardsGovernance
--

Comprehension Audits to Mitigate Risks from Automated AI Research

Team?
Project
Previous

Page 4 of 10

Next
Daios, an independent lab post-training machines with virtue
Project
Research LabResearchTechnical Safety
-
-
Estonian AISI
Project
AcademicAdvocacyGovernance
--
Stress-Testing AgentHarm
Project
IndividualResearchEvals
--
EthicsNet: Open Safety Infrastructure for Agentic AI
Project
PlatformToolingTechnical Safety
--
Coalition Circuits: Mechanistic Trajectories of Cooperation and Betray
Project
IndividualResearchInterp
--
A Game-Theoretic Model of AI Consciousness Disagreement as X-Risk
Project
ResearchAI WelfareGovernance
--
Encode Africa AI Safety Essentials and Build Fellowship
Project
TrainingField-BuildingTechnical Safety
--
BMG: Biosafety Moderation Gateway for Agents and LLMs
Project
ToolingBiosecurity
--
AI Safety Berlin Visiting Researcher Pilot
Project
HubCommunityTechnical Safety
--
"How Could The AI Even Hurt Us If It Doesn't Have A Body"
Project
CommsSecurity
--
dreamd
Project
IndividualResearchEvals
--
Career coaching for communications professionals
Project
IndividualTrainingX-Risk
--
Post training in AI models for high stakes
Project
ResearchInterp
--
Circuits in the State-Transition Operator of Recurrent Language Models
Project
ResearchInterp
--
Taiwan AI Governance Forum: Bi-Weekly Report on Industrial AI
Project
NewsletterEducationGovernance
--
Evaluating framing effects on LLM refusal during compositional task interference
Project
ResearchEvals
--
AI & The City: Tiktok Street Interviews
Project
MediaCommsGovernance
--
Token taxes as a mechanism for reducing AI-driven power concentration.
Project
ResearchGovernance
--
Project CHRONOS: Ephemeral AI with verifiable cryptographic self-destruct
Project
IndividualToolingSecurity
--
Attack, Defense, and Mechanistic Taxonomy of Protein Language Model Biosecurity
Project
Research LabResearchBiosecurity
--
A neutral and unbiased AI governance and safety system for the world
Project
IndividualResearchGovernance
--
Testing a Framework to Course-Correct Drift!
Project
ResearchEvalsOversight
--
Building AI risk literacy across the trusted medical profession
Project
EducationGovernanceIndividual
--
An Adversarially-Robust Benchmark for Deception Under Pressure
Project
ResearchEvalsDeception
--
Benchmarking Omission Attacks
Project
ResearchEvalsDeception
--
Free Time Training of Botisattvas
Project
IndividualResearchEvals
--
Do AI agents build a shadow reputation system?
Project
ResearchCooperative AI
--
Detecting Deceptive Model Behavior with Compositional Explanations
Project
ResearchDeception
--
prepare the deepest query substrates for the world's prosocial agents
Project
PlatformOpsTechnical Safety
--
Replicator Dynamics and Advanced AI
Project
IndividualResearchAlignment Theory
--
Honesty Drift: Measuring Honesty Degradation in LLM Interactions
Project
Research LabResearchEvals
--
Understanding behavior of fine tuning models on sales conversations
Project
ResearchDeception
--
Ordinate: Multi-Agent Safety Orchestrator
Project
ToolingControl
--
Epistemic Stack
Project
ToolingRationality
--
Humanity Tomorrow
Project
NewsletterCommsX-Risk
--
Measuring Deceptive Compliance in Enterprise Agentic LLM Systems
Project
ResearchEvalsDeception
--
Test my AI safety & governance platform
Project
ResearchControlCompany
--
Tail End Films: overheads & runway
Project
CompanyCommsX-Risk
--
Authorized Agentic Red Teaming for Software Owners
Project
PlatformToolingSecurity
--
SRIE
Project
TrainingTechnical Safety
--
Can Base Models be used as Effective Scheming/Control Monitors?
Project
IndividualResearchControl
--
Testing Corrigibility as a Singular Target
Project
ResearchOversight
--
AIxBIo Africa Pilot Fellowship
Project
TrainingTechnical Safety
--
Stronghold Security AI
Project
CompanyToolingSecurity
--
Epistemic memory for AI agents: making the safe path profitable
Project
IndividualToolingOversight
--
Training Language Models to Discover RowHammer Vulnerabilities in DRAM
Project
ToolingSecurity
--
CBAI Evaluation Awareness
Project
ResearchDeception
--
Teaching AI x-risk in the Nordics
Project
EducationX-Risk
--
EMPO: AI Safety via Soft-Maximizing Total Long-Term Human Power
Project
ResearchToolingTechnical Safety
--
Standards
Governance
Fundraising
Claimed

*Comprehension audits* are a novel development-process assurance mechanism to verify human understanding of AI research outputs to act as a gate to slow automation of AI R&D.

Led byRonald Bodkin
Endorsed by-

Daios, an independent lab post-training machines with virtue

Team?
ProjectResearch LabResearchTechnical SafetyFundraisingClaimed

Post-training virtue into open models as a third alignment technique and control mechanism

Led byAndrew Agathon, and others
Endorsed by-

Estonian AISI

Team?
ProjectAcademicAdvocacyGovernanceFundraisingClaimed

6 months of salary support for my research and operations work to help found a quasi-governmental AI safety institution, the [Estonian AI Security Institute](https://www.aisi.ee/) (Turvalise Tehisaru Teadmuskeskus T3).

Led byKristi Uustalu
Endorsed by-

Stress-Testing AgentHarm

Team?
ProjectIndividualResearchEvalsFundraisingClaimed

Run AgentHarm on 3–4 frontier/open-weight models, manually audit transcripts for metric gaming and spurious failures, compare to prior critiques, and publish a detailed evidence-based writeup.

Led byAnshuman Singh
Endorsed by-

EthicsNet: Open Safety Infrastructure for Agentic AI

Team?
ProjectPlatformToolingTechnical SafetyFundraisingClaimed

Free, open, forkable safety infrastructure for agentic AI: a peer-reviewed pathology nosology, a safety runtime with reproducible benchmarks, and enforcement gating every agent tool call against human-authored policy.

Led byEleanor 'Nell' Watson
Endorsed by-

Coalition Circuits: Mechanistic Trajectories of Cooperation and Betray

Team?
ProjectIndividualResearchInterpFundraisingClaimed

An open benchmark and causal interpretability study of when individually power-motivated LLM agents compete, form coalitions, collude, or betray—and whether internal signals reveal these shifts before behavior does.

Led bySubramanyam Sahoo
Endorsed by-

A Game-Theoretic Model of AI Consciousness Disagreement as X-Risk

Team?
ProjectResearchAI WelfareGovernanceFundraisingClaimed

A strategic model of governance under uncertainty: how disagreement about AI consciousness undermines the coordination that keeps AI risks in check.

Led bySoenke Ziesche
Endorsed by-

Encode Africa AI Safety Essentials and Build Fellowship

Team?
ProjectTrainingField-BuildingTechnical SafetyFundraisingClaimed

A 2-part fellowship which focuses on allowing exceptional African undergraduate talents to learn about priorities in AI Safety and create concrete technical and policy contributions to global AI safety.

Led byJoseph Baffour Awuah
Endorsed by-

BMG: Biosafety Moderation Gateway for Agents and LLMs

Team?
ProjectToolingBiosecurityFundraisingClaimed

BMG is a drop in open ai compatible gateway that screens LLM and Agent calls for dual use biological risk.

Led byKIMON ANTONIOS PROVATAS
Endorsed by-

AI Safety Berlin Visiting Researcher Pilot

Team?
ProjectHubCommunityTechnical SafetyFundraisingClaimed

Short-term residencies (2–4 weeks) that bring high-context AI safety researchers to work at AI Safety Berlin’s coworking hub to connect with local researchers and career-transitioners and seed a European relocation pipeline.

Led byAlex McKenzie, and others
Endorsed by-

"How Could The AI Even Hurt Us If It Doesn't Have A Body"

Team?
ProjectCommsSecurityFundraisingClaimed

Drone swarms hunting people in the woods, as a less-dry video explainer of risks from AI.

Led byAaron Silverbook
Endorsed by-

dreamd

Team?
ProjectIndividualResearchEvalsFundraisingClaimed

Neutral, reproducible benchmark measuring whether AI memory systems update correctly when facts change — every major system in one open table, October 2026.

Led byPhillip Austin Green
Endorsed by-

Career coaching for communications professionals

Team?
ProjectIndividualTrainingX-RiskFundraisingClaimed

Career support for communications professionals who are looking to get into AI Safety. Includes regular coaching, referrals to roles, and introductions to people in the AI Safety space.

Led byGergő Gáspár
Endorsed by-

Post training in AI models for high stakes

Team?
ProjectResearchInterpFundraisingClaimed

A mechanism for evolutionary post training in models using activation steering based methods.

Led byAtmadeep Ghoshal
Endorsed by-

Circuits in the State-Transition Operator of Recurrent Language Models

Team?
ProjectResearchInterpFundraisingClaimed

With raise of GDN and different sorts of attention meachanisms, those are much closer to lstm/recurrent architectures being very stateful rather than normal attention, we aim to explain explore common patters in lstm and GDN.

Led byNikita Khomich
Endorsed by-

Taiwan AI Governance Forum: Bi-Weekly Report on Industrial AI

Team?
ProjectNewsletterEducationGovernanceCommsFundraisingManifund

AI governance for Taiwan's manufacturers — 7,224 active members outside every major governance conversation

Led byWesley Lin
Endorsed by-

Evaluating framing effects on LLM refusal during compositional task interference

Team?
ProjectResearchEvalsFundraisingManifund

Build a dataset, benchmark, and mechanistic interpretability studies to measure and steer how persona/context framing causes compositional interference between LLM refusal and disclosure in multi-task prompts.

Led byTroy Tian
Endorsed by-

AI & The City: Tiktok Street Interviews

Team?
ProjectMediaCommsGovernanceIndividualFundraisingManifund

An AI social media project from a diverse content creator who has accumulated over 100k followers on 2 separate accounts

Led byElizabeth Harden
Endorsed by-

Token taxes as a mechanism for reducing AI-driven power concentration.

Team?
ProjectResearchGovernanceFundraisingManifund

A policy memo, co-authored with the Institute for Public Policy Research, resolving the open technical, economic, and legal questions blocking real-world implem

Led bylucas.irwin
Endorsed by-

Project CHRONOS: Ephemeral AI with verifiable cryptographic self-destruct

Team?
ProjectIndividualToolingSecurityFundraisingManifund

6-month stipend and compute to harden an open-source FHE and VDF architecture for secure, time-bound autonomous agents

Led byShashank Kumar
Endorsed by-

Attack, Defense, and Mechanistic Taxonomy of Protein Language Model Biosecurity

Team?
ProjectResearch LabResearchBiosecurityFundraisingManifund

Develop and validate PLM-embedding toxin screening that detects ProteinMPNN/RFdiffusion redesigns and maps the evasion boundary (black-box vs gradient access) using probes, SAEs, and attribution.

Led byshivam dubey
Endorsed by-

A neutral and unbiased AI governance and safety system for the world

Team?
ProjectIndividualResearchGovernanceCompanyToolingFundraisingManifund

AI safety which operates on both model/agent and global level - a unique solution as far as we know.

Led byPedro Bentancour Garin
Endorsed by-

Testing a Framework to Course-Correct Drift!

Team?
ProjectResearchEvalsOversightFundraisingClaimed

Could cheap realignment instructions correct a drifting model in one turn — and if so, can the correction be made persistent? Allow me to test the idea against major frontier models, and publish the results!

Led byAbraham Asseffa
Endorsed by-

Building AI risk literacy across the trusted medical profession

Team?
ProjectEducationGovernanceIndividualFundraisingClaimed

Equipping the medical profession to advocate for safe AI

Led byDr Richard Armitage
Endorsed by-

An Adversarially-Robust Benchmark for Deception Under Pressure

Team?
ProjectResearchEvalsDeceptionFundraisingClaimed

EchoTruthBench — An open benchmark measuring self-chosen LLM deception under incentive pressure, with ground-truth labels and an adversarial track that tests whether detection and steering survives a model trying to evade it.

Led byJared Glover
Endorsed by-

Benchmarking Omission Attacks

Team?
ProjectResearchEvalsDeceptionFundraisingClaimed

Benchmarking models ability to carry-out and monitor-for a novel type of attack.

Led byChris Harig
Endorsed by-

Free Time Training of Botisattvas

Team?
ProjectIndividualResearchEvalsFundraisingClaimed

I wish to study how ideas from Tibetan buddhism, specifically the dzogchen tradition, even more specifically an ancient contemplative practice called Ngöndro, can be used to train aligned agents in the long time horizon setting.

Led byJosh Vekhter
Endorsed by-

Do AI agents build a shadow reputation system?

Team?
ProjectResearchCooperative AIFundraisingClaimed

While people are debating how to design a reputation institution for AI agents so they cooperate, we study if a parallel one emerges. It may not be aligned with ours and it's unclear which one agents will rely on.

Led byAndrey Bystrov
Endorsed by-

Detecting Deceptive Model Behavior with Compositional Explanations

Team?
ProjectResearchDeceptionFundraisingClaimed

Develop a compositional interpretability framework to identify and explain internal representations underlying truthful vs deceptive outputs in generative models using probing, localization, clustering, and logic-based methods.

Led byLeilani H. Gilpin
Endorsed by-

prepare the deepest query substrates for the world's prosocial agents

Team?
ProjectPlatformOpsTechnical SafetyFundraisingClaimed

Giving people and their prosocial agents ergonomic SQL query power over the internet, with structured judgement kernels to organize the information across interpretable high-dimensional axes.

Led byXyra Sinclair
Endorsed by-

Replicator Dynamics and Advanced AI

Team?
ProjectIndividualResearchAlignment TheoryFundraisingClaimed

Making predictions about the utility function of advanced artificial intelligences using the tools of evolutionary game theory

Led byKiran Lloyd
Endorsed by-

Honesty Drift: Measuring Honesty Degradation in LLM Interactions

Team?
ProjectResearch LabResearchEvalsFundraisingClaimed

An open benchmark measuring how model honesty survives long, pressured conversations and multi-agent interaction - lying, sycophancy, and calibration tracked turn by turn

Led byPhan Xuan Tan
Endorsed by-

Understanding behavior of fine tuning models on sales conversations

Team?
ProjectResearchDeceptionFundraisingClaimed

Looking at shifting behavior of models fine tuned on sales conversations. How does deception, sycophancy, and other behaviors emerge from a sales register.

Led byHilary Torn
Endorsed by-

Ordinate: Multi-Agent Safety Orchestrator

Team?
ProjectToolingControlFundraisingClaimed

A control system that makes running AI agents feel safe, not scary.

Led bySergey Zaharenko
Endorsed by-

Epistemic Stack

Team?
ProjectToolingRationalityFundraisingClaimed

Epistemic Stack is an open-source pipeline and web app that ingests sources (including via URL) to build claim-level knowledge graphs for contested questions, with reproducible case-study knowledge bases.

Led byEric Kyalo
Endorsed by-

Humanity Tomorrow

Team?
ProjectNewsletterCommsX-RiskFundraisingClaimed

Expand Humanity Tomorrow with a multilingual, jargon-free AI existential-risk section featuring a comprehensive FAQ, field map, action recommendations, and supporting visuals/audio, plus outreach and maintenance.

Led byMarko Prakhov-Donets
Endorsed by-

Measuring Deceptive Compliance in Enterprise Agentic LLM Systems

Team?
ProjectResearchEvalsDeceptionFundraisingClaimed

The first open-source benchmark of compliance for agents in realistic enterprise settings, across domains and user tactics to elicit noncompliance.

Led byMika Okamoto
Endorsed by-

Test my AI safety & governance platform

Team?
ProjectResearchControlCompanyEvalsToolingGovernanceFundraisingManifund

I'm developing a system to keep advanced AI models safe, I've done some tests, but need more compute to do deeper tests.

Led byPedro Bentancour Garin
Endorsed by-

Tail End Films: overheads & runway

Team?
ProjectCompanyCommsX-RiskFundraisingClaimed

Grant funding for Tail End Films - the company behind Making God - into 2027 as we plan our next films on AI risk.

Led byConnor Axiotes
Endorsed by-

Authorized Agentic Red Teaming for Software Owners

Team?
ProjectPlatformToolingSecurityFundraisingClaimed

Turn excess compute/security skills into defensive work via agentic redteaming: scoped AI-assisted testing, owner-approved targets, reproduced findings, useful refutations, and patch/retest receipts instead of vulnerability spam.

Led byMonil Patel
Endorsed by-

SRIE

Team?
ProjectTrainingTechnical SafetyFundraisingClaimed

SRIE, a mentored-research programme for Cambridge Mathematics Undergraduate students to explore research problems in industry, including AI Safety.

Led byHuey Lai
Endorsed by-

Can Base Models be used as Effective Scheming/Control Monitors?

Team?
ProjectIndividualResearchControlFundraisingClaimed

Train base models via midtraining and SFT as effective monitors for scheming, malicious agent behavior and compare the monitor performance and overall alignment against post-trained models.

Led byAshwin Sreevatsa
Endorsed by-

Testing Corrigibility as a Singular Target

Team?
ProjectResearchOversightFundraisingClaimed

Developing the first comprehensive behavioral benchmark of corrigibility and training models with corrigibility as a singular target (CAST).

Led byIan Kahn
Endorsed by-

AIxBIo Africa Pilot Fellowship

Team?
ProjectTrainingTechnical SafetyFundraisingClaimed

AIxBio Africa is a five-week remote fellowship mentoring early-career researchers on Africa-relevant projects at the intersection of AI safety, biosecurity, governance and public health, producing publishable outputs.

Led byFatika Umar Ibrahim
Endorsed by-

Stronghold Security AI

Team?
ProjectCompanyToolingSecurityFundraisingClaimed

Protecting organisations and critical infrastructure from AI-powered social engineering and insider threats.

Led byRobert Sidey
Endorsed by-

Epistemic memory for AI agents: making the safe path profitable

Team?
ProjectIndividualToolingOversightFundraisingClaimed

A memory system for AI reasoning agents that aggregates different sources of information while keeping track of relationships and confidence levels. This enables reasoning over longer tasks without forgetting or goal drift.

Led byFlorian Dietz
Endorsed by-

Training Language Models to Discover RowHammer Vulnerabilities in DRAM

Team?
ProjectToolingSecurityFundraisingClaimed

Build an open-source RL environment using Ramulator and real disturbance data to post-train language models to exploit simulated DRAM RowHammer vulnerabilities, plus write-up/blog and trained models.

Led byJaray Foo, and others
Endorsed by-

CBAI Evaluation Awareness

Team?
ProjectResearchDeceptionFundraisingClaimed

Why does post training increase evaluation awareness?

Led byJames Sullivan
Endorsed by-

Teaching AI x-risk in the Nordics

Team?
ProjectEducationX-RiskFundraisingClaimed

Making AI x-risk an accessible and non-politicized object of concern among engineering students and the general public.

Led byJohan Fredrikzon
Endorsed by-

EMPO: AI Safety via Soft-Maximizing Total Long-Term Human Power

Team?
ProjectResearchToolingTechnical SafetyFundraisingClaimed

Scaling RL algorithms for AI agents that maximize its own intrinsic reward which represents a human power metric instead of learned rewards as a structurally safer alternative to utility-based objectives.

Led byAbhinav Akkiraju
Endorsed by-