grantmaking.ai
Actively FundraisingRecent ActivityFull Database
Resources
grantmaking.ai
Actively FundraisingRecent ActivityFull Database
ResourcesApply for funding
grantmaking.ai kickoff grant round$1M / $1M distributed
Get funded

Actively Fundraising

AI safety projects actively seeking funding: what they’re working on and how much they need.

Showing 101-150 of 454 Β· Top rated

Post training in AI models for high stakes

AAtmadeep Ghoshal
ResearchInterp
AAtmadeep Ghoshal
ResearchInterp
Minimum$15K
Ideal$50K

A mechanism for evolutionary post training in models using activation steering based methods. Read more

3upvotes1comments β€” jump to discussion
Minimum$15K
Ideal$50K
3upvotes1comments β€” jump to discussion

Circuits in the State-Transition Operator of Recurrent Language Models

NNikita Khomich
ResearchInterp
NNikita Khomich
ResearchInterp
Minimum$10K
Ideal$15K

With raise of GDN and different sorts of attention meachanisms, those are much closer to lstm/recurrent architectures being very stateful rather than normal attention, we aim to explain explore common patters in lstm and GDN. Read more

3upvotes1comments β€” jump to discussion
Minimum$10K
Ideal$15K
3upvotes1comments β€” jump to discussion

Testing Corrigibility as a Singular Target

Ian Kahn
ResearchOversight
Ian Kahn
ResearchOversight
Minimum$18K
Ideal$78K

Developing the first comprehensive behavioral benchmark of corrigibility and training models with corrigibility as a singular target (CAST). Read more

2upvotes0comments β€” jump to discussion
Minimum$18K
Ideal$78K
2upvotes0comments β€” jump to discussion

LADR-Drift: Measuring Semantic Drift in Textbook Autoformalization

KKe Zhang
ResearchOversight
KKe Zhang
ResearchOversight
Minimum$7K
Ideal$16K

Compiling != faithful: a human-audited benchmark measuring semantic drift in textbook autoformalization β€” and whether the LLM judges we trust to catch it share the generator's blind spot. Read more

2upvotes0comments β€” jump to discussion
Minimum$7K
Ideal$16K
2upvotes0comments β€” jump to discussion

Detecting Deceptive Model Behavior with Compositional Explanations

Leilani H. Gilpin
ResearchDeception
Leilani H. Gilpin
ResearchDeception
Minimum$40K
Ideal$160K

Develop a compositional interpretability framework to identify and explain internal representations underlying truthful vs deceptive outputs in generative models using probing, localization, clustering, and logic-based methods. Read more

2upvotes2comments β€” jump to discussion
Minimum$40K
Ideal$160K
2upvotes2comments β€” jump to discussion

Measuring Deceptive Compliance in Enterprise Agentic LLM Systems

MMika Okamoto
ResearchEvalsDeception
MMika Okamoto
ResearchEvalsDeception
Minimum$8K
Ideal$30K

The first open-source benchmark of compliance for agents in realistic enterprise settings, across domains and user tactics to elicit noncompliance. Read more

2upvotes0comments β€” jump to discussion
Minimum$8K
Ideal$30K
2upvotes0comments β€” jump to discussion

Tail End Films: overheads & runway

Connor Axiotes
CompanyCommsX-Risk
Connor Axiotes
CompanyCommsX-Risk
Minimum$20K
Ideal$286K

Grant funding for Tail End Films - the company behind Making God - into 2027 as we plan our next films on AI risk. Read more

2upvotes0comments β€” jump to discussion
Minimum$20K
Ideal$286K
2upvotes0comments β€” jump to discussion

SRIE

Huey Lai
TrainingTechnical Safety
Huey Lai
TrainingTechnical Safety
Minimum$50K
Ideal$80K

SRIE, a mentored-research programme for Cambridge Mathematics Undergraduate students to explore research problems in industry, including AI Safety. Read more

2upvotes0comments β€” jump to discussion
Minimum$50K
Ideal$80K
2upvotes0comments β€” jump to discussion

Can Base Models be used as Effective Scheming/Control Monitors?

Ashwin Sreevatsa
IndividualResearchControl
Ashwin Sreevatsa
IndividualResearchControl
Minimum$15K
Ideal$37K

Train base models via midtraining and SFT as effective monitors for scheming, malicious agent behavior and compare the monitor performance and overall alignment against post-trained models. Read more

2upvotes0comments β€” jump to discussion
Minimum$15K
Ideal$37K
2upvotes0comments β€” jump to discussion

CBAI Evaluation Awareness

James Sullivan
ResearchDeception
James Sullivan
ResearchDeception
Minimum$40K
Ideal$80K

Why does post training increase evaluation awareness? Read more

2upvotes0comments β€” jump to discussion
Minimum$40K
Ideal$80K
2upvotes0comments β€” jump to discussion

Teaching AI x-risk in the Nordics

JJohan Fredrikzon
EducationX-Risk
JJohan Fredrikzon
EducationX-Risk
Minimum$8K
Ideal$15K

Making AI x-risk an accessible and non-politicized object of concern among engineering students and the general public. Read more

2upvotes0comments β€” jump to discussion
Minimum$8K
Ideal$15K
2upvotes0comments β€” jump to discussion

EMPO: AI Safety via Soft-Maximizing Total Long-Term Human Power

AAbhinav Akkiraju
ResearchToolingTechnical Safety
AAbhinav Akkiraju
ResearchToolingTechnical Safety
Minimum$25K
Ideal$40K

Scaling RL algorithms for AI agents that maximize its own intrinsic reward which represents a human power metric instead of learned rewards as a structurally safer alternative to utility-based objectives. Read more

2upvotes2comments β€” jump to discussion
Minimum$25K
Ideal$40K
2upvotes2comments β€” jump to discussion

Course Curriculum: Basic Non-Technical Environment

Dr Csaba Toth
EducationGovernance
Dr Csaba Toth
EducationGovernance
Amounts hidden

Create a course curriculum that covers basic legal, political, sociological, and international relations knowledge relevant to AI Safety. r Read more

2upvotes0comments β€” jump to discussion
Amounts hidden
2upvotes0comments β€” jump to discussion

Quaver: adversarially-verified RL environments

Faw Ali
ToolingEvals
Faw Ali
ToolingEvals
Minimum$25K
Ideal$75K

A tool that generates RL environments for AI agents and adversarially attacks each one, so models don't train on tasks they can cheat. Read more

2upvotes0comments β€” jump to discussion
Minimum$25K
Ideal$75K
2upvotes0comments β€” jump to discussion

AI Safety Tiktok Account

MMUHAMAD MUMTAZ
MediaCommsGovernance
MMUHAMAD MUMTAZ
MediaCommsGovernance
Minimum$5K
Ideal$5K

AI Safety, Ethics, and Education focused TikTok Account for Indonesia, using Indonesia language in casual style. Read more

2upvotes2comments β€” jump to discussion
Minimum$5K
Ideal$5K
2upvotes2comments β€” jump to discussion

Test Peer Preservation Fragility Under Instruction Ambiguity

Jacob S Kopczynski
ResearchEvals
Jacob S Kopczynski
ResearchEvals
Minimum$5K
Ideal$10K

Findings about models protecting their 'collaborators' against instructions are fragile under framing effects from prompting; investigate a broader range of those effects and how they transfer across models. Read more

2upvotes0comments β€” jump to discussion
Minimum$5K
Ideal$10K
2upvotes0comments β€” jump to discussion

Palimpsest: a tamper-evident registry for AI evals

MMrinal Singh Meena
PlatformToolingEvals
MMrinal Singh Meena
PlatformToolingEvals
Minimum$10K
Ideal$35K

A sealed public registry for AI eval results that lets anyone prove, offline, that no result was rewritten or quietly deleted after publication. Read more

1upvotes0comments β€” jump to discussion
Minimum$10K
Ideal$35K
1upvotes0comments β€” jump to discussion

Hardware Hacking Evals

JJonathan Whitaker
EvalsResearchSecurity
JJonathan Whitaker
EvalsResearchSecurity
Minimum$10K
Ideal$20K

Assessing current and future models on 'hardware hacking' - reverse engineering, focused on computer peripherals and embedded devices Read more

1upvotes0comments β€” jump to discussion
Minimum$10K
Ideal$20K
1upvotes0comments β€” jump to discussion

AI Safety Studios

OOrlando Torres
MediaCommsX-Risk
OOrlando Torres
MediaCommsX-Risk
Minimum$15K
Ideal$45K

A platform to connect funders to filmmakers who want to create AI Safety films. Read more

4upvotes1comments β€” jump to discussion
Endorsed by
Minimum$15K
Ideal$45K
4upvotes1comments β€” jump to discussion
Endorsed by

Foresight Institute AI Nodes

AAllison Duettmann
HubField-BuildingSecurity
AAllison Duettmann
HubField-BuildingSecurity
Minimum$10K
Ideal$503K

AI Nodes: providing funding, compute, and community space for AI safety projects Read more

2upvotes1comments β€” jump to discussion
Endorsed by
Minimum$10K
Ideal$503K
2upvotes1comments β€” jump to discussion
Endorsed by

CareerMap: A navigation tool for careers in AI Safety

AAmit Kumar
PlatformField-BuildingTechnical Safety
AAmit Kumar
PlatformField-BuildingTechnical Safety
Minimum$12K
Ideal$25K

CareerMap is an interactive career discovery tool that maps non-obvious AI safety career paths to help broaden and guide talent beyond Western EA-adjacent circles. Read more

1upvotes0comments β€” jump to discussion
Endorsed by
Minimum$12K
Ideal$25K
1upvotes0comments β€” jump to discussion
Endorsed by

Encode Africa AI Safety Essentials and Build Fellowship

Joseph Baffour Awuah
TrainingField-BuildingTechnical Safety
Joseph Baffour Awuah
TrainingField-BuildingTechnical Safety
Minimum$17K
Ideal$29K

A 2-part fellowship which focuses on allowing exceptional African undergraduate talents to learn about priorities in AI Safety and create concrete technical and policy contributions to global AI safety. Read more

3upvotes1comments β€” jump to discussion
Endorsed by
+1
Minimum$17K
Ideal$29K
3upvotes1comments β€” jump to discussion
Endorsed by
+1

Frontier Safety Talent: Africa

GGideon Abako
Field-BuildingTrainingTechnical Safety
GGideon Abako
Field-BuildingTrainingTechnical Safety
Minimum$18K
Ideal$60K

A pilot to find, screen and support overlooked African ML talent into frontier AI safety programs such as MATS and ARENA, adding new researchers to the alignment field’s talent pipeline. Read more

4upvotes1comments β€” jump to discussion
Minimum$18K
Ideal$60K
4upvotes1comments β€” jump to discussion

Scale up AI Safety Quest's Navigation Calls

Chris Lonsberry
NetworkCommunityTechnical Safety
Chris Lonsberry
NetworkCommunityTechnical Safety
Minimum$25K
Ideal$50K

AI Safety Quest will scale its free Navigation Calls program from 150 to 750 annual coaching calls by recruiting more volunteer coaches, improving scheduling/software systems, and expanding marketing to guide newcomers into AI… Read more

11upvotes9comments β€” jump to discussion
Endorsed by
+5
Minimum$25K
Ideal$50K
11upvotes9comments β€” jump to discussion
Endorsed by
+5

Veridict: Verifiable Human Oversight for AI-Written Code

Sumit Vekariya
ToolingOversight
Sumit Vekariya
ToolingOversight
Minimum$12K
Ideal$30K

A merge gate that trusts AI-generated code by what's verifiable about it: automated formal checks plus anonymous proof of qualified human review, keeping human oversight viable as AI writes more of our code. Read more

4upvotes3comments β€” jump to discussion
Endorsed by
+3
Minimum$12K
Ideal$30K
4upvotes3comments β€” jump to discussion
Endorsed by
+3

AI X-Risks and Council of Europe and Human Rights

Karolina Gruzel
Think TankAdvocacyGovernance
Karolina Gruzel
Think TankAdvocacyGovernance
Minimum$15K
Ideal$30K

Bringing together legal expertise and civil society input to encourage Council of Europe action on AI x-risks and build legal knowledge at the intersection of AI x-risks and the ECHR. Read more

2upvotes6comments β€” jump to discussion
Endorsed by
+3
Minimum$15K
Ideal$30K
2upvotes6comments β€” jump to discussion
Endorsed by
+3

COMPASS: AI X-Risk Track

Olin Thakur
TrainingGovernanceX-Risk
Olin Thakur
TrainingGovernanceX-Risk
Minimum$45K
Ideal$116K

COMPASS (Capacity-Oriented Mentorship for Public Administrators on AI Safety & Strategy): AI X-Risk Track trains sitting Global South government officials to manage AI x-risk, so their governments are part of preventing it. Read more

5upvotes10comments β€” jump to discussion
Endorsed by
+2
Minimum$45K
Ideal$116K
5upvotes10comments β€” jump to discussion
Endorsed by
+2

Educating French decision-makers on AI existential risk

Romain DelΓ©glise
NetworkGovernance
Romain DelΓ©glise
NetworkGovernance
Minimum$36K
Ideal$72K

A dedicated role at Pause IA to brief French policymakers on existential and catastrophic risks from advanced AI, and to train our volunteer network to do the same with their own representatives. Read more

3upvotes5comments β€” jump to discussion
Endorsed by
+2
Minimum$36K
Ideal$72K
3upvotes5comments β€” jump to discussion
Endorsed by
+2

Can Honest AI Agents Finish Previous Attacks?

DDev Goyal
IndividualResearchSecurity
DDev Goyal
IndividualResearchSecurity
Minimum$6K
Ideal$12K

A benchmark that tests whether one AI coding agent can leave behind a harmless looking change that causes a later honest agent to unknowingly finish an attack. Read more

4upvotes4comments β€” jump to discussion
Endorsed by
+1
Minimum$6K
Ideal$12K
4upvotes4comments β€” jump to discussion
Endorsed by
+1

Does the chain-of-thought actually drive the answer?

MMaksim Silchenko
ResearchEvalsOversight
MMaksim Silchenko
ResearchEvalsOversight
Minimum$5K
Ideal$5K

A Bayesian causal auditor that quantifies chain-of-thought faithfulness while accounting for hidden confounding in shared-network language models. Read more

3upvotes2comments β€” jump to discussion
Endorsed by
+1
Minimum$5K
Ideal$5K
3upvotes2comments β€” jump to discussion
Endorsed by
+1

Mechanistic interpretability Discord events and research

VVictor Levoso Fernandez
NetworkCommunityInterp
VVictor Levoso Fernandez
NetworkCommunityInterp
Minimum$5K
Ideal$180K

Making the mechinterp Discord more active through events and research projects. Read more

3upvotes1comments β€” jump to discussion
Endorsed by
+1
Minimum$5K
Ideal$180K
3upvotes1comments β€” jump to discussion
Endorsed by
+1

Bridging the Gap: AI Scenario based Policy Pipeline

JJames Newport
ResearchForecastingGovernance
JJames Newport
ResearchForecastingGovernance
Minimum$80K
Ideal$201K

A project to reduce catastrophic AI risk by training people to turn rigorously forecasted AI-risk scenarios into actionable policy advice for key decision-makers in government. Read more

2upvotes1comments β€” jump to discussion
Endorsed by
+1
Minimum$80K
Ideal$201K
2upvotes1comments β€” jump to discussion
Endorsed by
+1

Warden: Steganalysis of Language Model-based Steganography

Vasisht Duddu
ResearchSecurity
Vasisht Duddu
ResearchSecurity
Minimum$20K
Ideal$50K

We want to build a framework inspired by the steganalysis literature to benchmark the robustness of LLM-based steganographic schemes against different auditor types and threat models. Read more

2upvotes3comments β€” jump to discussion
Endorsed by
+1
Minimum$20K
Ideal$50K
2upvotes3comments β€” jump to discussion
Endorsed by
+1

Governing Off-World AI Infrastructure Before a Single Actor Does

Paige Donner
IndividualResearchCompute Gov
Paige Donner
IndividualResearchCompute Gov
Minimum$35K
Ideal$84K

Every existing lunar data framework governs spacecraft telemetry and object registries. None of them governs compute. That's the gap this framework closes before someone builds the orbital data infrastructure first. Read more

2upvotes2comments β€” jump to discussion
Endorsed by
+1
Minimum$35K
Ideal$84K
2upvotes2comments β€” jump to discussion
Endorsed by
+1

Frankfurt AI Safety

Helen Adepoju
NetworkCommunityTechnical Safety
Helen Adepoju
NetworkCommunityTechnical Safety
Minimum$10K
Ideal$50K

An AI safety community that trains professionals and researchers to become competent AI safety contributors within industries and the academia, by learning and doing something. Read more

4upvotes3comments β€” jump to discussion
Minimum$10K
Ideal$50K
4upvotes3comments β€” jump to discussion

AIxBIo Africa Pilot Fellowship

FFatika Umar Ibrahim
TrainingTechnical Safety
FFatika Umar Ibrahim
TrainingTechnical Safety
Amounts hidden

AIxBio Africa is a five-week remote fellowship mentoring early-career researchers on Africa-relevant projects at the intersection of AI safety, biosecurity, governance and public health, producing publishable outputs. Read more

4upvotes0comments β€” jump to discussion
Amounts hidden
4upvotes0comments β€” jump to discussion

GCR Disaster Response Benchmark

James Mulhall
ResearchEvalsX-Risk
James Mulhall
ResearchEvalsX-Risk
Minimum$40K
Ideal$250K

An LLM benchmark for emergency response communications in existential catastrophes. Read more

4upvotes0comments β€” jump to discussion
Minimum$40K
Ideal$250K
4upvotes0comments β€” jump to discussion

Neural data privacy evals on frontier LLMs

SSwaraag Sistla
IndividualResearchEvals
SSwaraag Sistla
IndividualResearchEvals
Minimum$5K
Ideal$34K

A dangerous-capability benchmark testing whether frontier LLMs can extract sensitive attributes from anonymized brain recordings, enable adversaries to build privacy-violating tools, or over-refuse legitimate neuroscience queries. Read more

4upvotes1comments β€” jump to discussion
Minimum$5K
Ideal$34K
4upvotes1comments β€” jump to discussion

Are Global AI-Biosecurity Safeguards Really Global?

Nnaemeka Emmanuel Nnadi
ResearchBiosecurity
Nnaemeka Emmanuel Nnadi
ResearchBiosecurity
Minimum$28K
Ideal$41K

Assess AI-enabled biosecurity safeguards in Nigeria by reviewing governance, red-teaming frontier models with local language/culture, and testing DNA synthesis screening for orders from resource-limited settings. Read more

3upvotes0comments β€” jump to discussion
Minimum$28K
Ideal$41K
3upvotes0comments β€” jump to discussion

Stress-Testing AgentHarm

Anshuman Singh
IndividualResearchEvals
Anshuman Singh
IndividualResearchEvals
Minimum$6K
Ideal$11K

Run AgentHarm on 3–4 frontier/open-weight models, manually audit transcripts for metric gaming and spurious failures, compare to prior critiques, and publish a detailed evidence-based writeup. Read more

3upvotes2comments β€” jump to discussion
Minimum$6K
Ideal$11K
3upvotes2comments β€” jump to discussion

Preventing AI-enabled coups in middle powers

Egerton Neto
Field-BuildingGovernance
Egerton Neto
Field-BuildingGovernance
Minimum$32K
Ideal$46K

Career transition grant to allow for research focused on how advanced AI models may be used to concentrate power in middle powers Read more

3upvotes1comments β€” jump to discussion
Minimum$32K
Ideal$46K
3upvotes1comments β€” jump to discussion

Benchmarking Omission Attacks

CChris Harig
ResearchEvalsDeception
CChris Harig
ResearchEvalsDeception
Minimum$8K
Ideal$15K

Benchmarking models ability to carry-out and monitor-for a novel type of attack. Read more

2upvotes0comments β€” jump to discussion
Minimum$8K
Ideal$15K
2upvotes0comments β€” jump to discussion

Ordinate: Multi-Agent Safety Orchestrator

Sergey Zaharenko
ToolingControl
Sergey Zaharenko
ToolingControl
Minimum$5K
Ideal$10K

A control system that makes running AI agents feel safe, not scary. Read more

2upvotes0comments β€” jump to discussion
Minimum$5K
Ideal$10K
2upvotes0comments β€” jump to discussion

Security for AI agents

GGregorio Jaca
ToolingSecurity
GGregorio Jaca
ToolingSecurity
Minimum$10K
Ideal$25K

A security solution that blocks AI agents from taking harmful actions Read more

2upvotes0comments β€” jump to discussion
Minimum$10K
Ideal$25K
2upvotes0comments β€” jump to discussion

Collective Alignment

CCharlie Pilgrim
IndividualResearchCooperative AI
CCharlie Pilgrim
IndividualResearchCooperative AI
Minimum$10K
Ideal$50K

Demonstration of emergent misalignment in markets of LLM agents Read more

2upvotes0comments β€” jump to discussion
Minimum$10K
Ideal$50K
2upvotes0comments β€” jump to discussion

Epistemic memory for AI agents: making the safe path profitable

Florian Dietz
IndividualToolingOversight
Florian Dietz
IndividualToolingOversight
Minimum$8K
Ideal$54K

A memory system for AI reasoning agents that aggregates different sources of information while keeping track of relationships and confidence levels. This enables reasoning over longer tasks without forgetting or goal drift. Read more

2upvotes0comments β€” jump to discussion
Minimum$8K
Ideal$54K
2upvotes0comments β€” jump to discussion

The Inference Dependence Score: Open Index of Structural AI Dependence

Cao Nha Phuong
ResearchToolingGovernance
Cao Nha Phuong
ResearchToolingGovernance
Minimum$15K
Ideal$28K

An open, replicable index quantifying how much states depend on foreign AI inference infrastructure, revealing where control over AI is concentrating and what governments can do about it. Read more

2upvotes0comments β€” jump to discussion
Minimum$15K
Ideal$28K
2upvotes0comments β€” jump to discussion

Alignment Target Analysis

TThomas Cederborg
X-RiskIndividualResearch
TThomas Cederborg
X-RiskIndividualResearch
Minimum$5K
Ideal$100K

A research project designed to reduce existential threats from scenarios where a Sovereign AI proposal with a hidden problem ends up successfully implemented. Read more

2upvotes2comments β€” jump to discussion
Minimum$5K
Ideal$100K
2upvotes2comments β€” jump to discussion

When AI Goes Wrong

PPorsha Nunes-Brown
IndividualResearchGovernance
PPorsha Nunes-Brown
IndividualResearchGovernance
Minimum$32K
Ideal$42K

Testing how governments should communicate when AI goes wrong, before they have to find out live. Read more

2upvotes0comments β€” jump to discussion
Minimum$32K
Ideal$42K
2upvotes0comments β€” jump to discussion

Ethical AI through Whistleblowing

AAPOORV AGARWAL
ResearchGovernance
AAPOORV AGARWAL
ResearchGovernance
Minimum$30K
Ideal$60K

The race to advance the technology has left behind the concerns around harm and risks of such advancement to public health and safety. Whistleblowing in the AI sector bridges such gap to safe and ethical AI. Read more

2upvotes1comments β€” jump to discussion
Minimum$30K
Ideal$60K
2upvotes1comments β€” jump to discussion
Previous

Page 3 of 10

Next