grantmaking.ai
Actively FundraisingRecent ActivityFull Database
Resources
grantmaking.ai
Actively FundraisingRecent ActivityFull Database
ResourcesApply for funding
grantmaking.ai kickoff grant round$278k / $1M distributed
Get funded

Database

The AI Safety database is a work in progress, as we are focusing on the launch of the initial $1M grant round. Feel free to send us ideas or feedback!
Raising funds(remove filter)Clear all
NameTypeTagsEndorsedTeamRaising $

Showing 251-300 of 460·Sorted by endorsements

NameTypeTagsEndorsedTeamRaising $
Security for AI agents
Project
ToolingSecurity
--

Security for AI agents

Team?
Project
Previous

Page 6 of 10

Next
Executable Models of Alignment
Project
ResearchTechnical Safety
-
-
Verified Human Sideloads for AI Safety: can a measured digital copy of
Project
IndividualResearchValue Alignment
--
Overcoming prompt sensitivity by tokenization robust self-distillation
Project
ResearchRobustness
--
SolonGate
Project
CompanyToolingSecurity
--
AISC Research Incubator
Project
IncubatorTrainingTechnical Safety
--
HausaSafe
Project
ToolingEvals
--
Multi-agent mechanistic interpretability and inferentialism
Project
IndividualResearchInterp
--
The Manhattan Kernel – A Deterministic, CPU-Bound AI Runtime Engine
Project
ToolingControl
--
rAI: Governed AI Task Forces for Real-World Operations
Project
ToolingGovernance
--
Hardware Hacking Evals
Project
EvalsResearchSecurity
--
Parametric Mechanisms of Unintended Generalization in LLMs
Project
ResearchInterp
--
Leading with Complexity and Wisdom for AI Safety Leaders
Project
TrainingField-BuildingTechnical Safety
--
Early-Career Technical AI safety research transition
Project
IndividualResearchTechnical Safety
--
Career Funding- Philippine Philosopher transitioning to AI
Project
IndividualResearchGovernance
--
The Human Influence Observatory
Project
PlatformResearchGovernance
--
echo-conscience-llm: Emotional Memory as an Honesty Activation Trigger
Project
ToolingResearchDeception
--
Warden: Steganalysis of Language Model-based Steganography
Project
ResearchSecurity
--
HRM monitored TRM and LDT Mesh for Disproportionate Control
Project
ToolingControl
--
Collective Alignment
Project
IndividualResearchCooperative AI
--
MPhil research on AI's perception of animacy and vulnerability
Project
ResearchTechnical Safety
--
Persona Introspection
Project
IndividualResearchAI Welfare
--
Alignment Persistence During Continual Optimization of LLMs
Project
ResearchTechnical Safety
--
Seeing the Recursive Machine
Project
ResearchToolingControl
--
Kristina K's Career Transition Grant
Project
IndividualField-BuildingGovernance
--
Cultivating Meta-Ethical AI via Human-in-the-Loop Socratic Discourse
Project
IndividualResearchValue Alignment
--
Preventing AI-enabled coups in middle powers
Project
Field-BuildingGovernance
--
AI Buildout Frontier
Project
PlatformToolingCompute Gov
--
The Adaptive Layer
Project
IndividualToolingSecurity
--
Underground Cultural District
Project
ToolingAI Welfare
--
future generation of healthy AI
Project
EducationGovernance
--
The Wall of Sleep: Art, Encounter, and the Moral Status of Digital Mind
Project
IndividualCommsAI Welfare
--
AI risk education to underserved communities
Project
CommsX-Risk
--
Defence AI Signal: Mapping Military AI Governance Risks & Developments
Project
NewsletterResearchGovernance
--
Pakistan AI Saftey Innitiative
Project
Field-BuildingGovernance
--
Clinical Trials for Model Migration: accuracy lies, honesty doesn't
Project
ResearchEvals
--
When AI Goes Wrong
Project
IndividualResearchGovernance
--
Transhumanist Values Assessment (TVA) for Frontier LLMs
Project
ResearchEvalsValue Alignment
--
noisify
Project
ToolingSecurity
--
Cross-lingual guardrail auditing & inference steering tools.
Project
ToolingEvalsInterp
--
Bridging the gap between mentors and mentees
Project
PlatformCommunity
--
The rise of open science labs- A book, 7+ hour video essay, and course
Project
IndividualEducationField-Building
--
AgentTrap: A Census of How Hijackable AI Agents Are in the Wild
Project
ResearchEvalsSecurity
--
AQI — Runtime Admissibility & Execution Governance Layer for Autonomou
Project
ToolingControl
--
Benchmark for Reversal and Unlearning of Harmful Fine-Tuning
Project
IndividualResearchEvals
--
Stellaris-ModelScope: Interpretability Infrastructure for AI
Project---
Neurological Biocompute field building
Project
ResearchToolingBiosecurity
--
Can Fragile Labs Use AI Safely in Liberia, West Africa?
Project
ResearchBiosecurity
--
Measuring whether guardrail phrasing changes agent rule-breaking.
Project
IndividualResearchEvals
--
Marginal Baseline Evaluation for AI Safety Metrics
Project
ResearchToolingEvals
--
Tooling
Security
Fundraising
Claimed

A security solution that blocks AI agents from taking harmful actions

Led byGregorio Jaca
Endorsed by-

Executable Models of Alignment

Team?
ProjectResearchTechnical SafetyFundraisingClaimed

Create simple programs that exhibit parts of the hard problems of alignment, allowing for fast iteration on conceptual ideas.

Led byJohannes C. Mayer
Endorsed by-

Verified Human Sideloads for AI Safety: can a measured digital copy of

Team?
ProjectIndividualResearchValue AlignmentFundraisingClaimed

I build "sideloads" - digital copies of real people with measured fidelity (two are running now). I want to test whether a copy of a specific trusted human can evaluate AI decisions at machine speed.

Led byAlexey Turchin
Endorsed by-

Overcoming prompt sensitivity by tokenization robust self-distillation

Team?
ProjectResearchRobustnessFundraisingClaimed

Verifying a modification to post-training for LLMs via self-distillation to allow building highly capable and less prompt-sensitive models by exploiting tokenization stochasticity

Led bySergei Kudriashov
Endorsed by-

SolonGate

Team?
ProjectCompanyToolingSecurityFundraisingClaimed

Security Gateway for AI Execution.

Led byEmirhan Demir
Endorsed by-

AISC Research Incubator

Team?
ProjectIncubatorTrainingTechnical SafetyFundraisingClaimed

A 16-day virtual incubator (Aug 15-30, 2026) for developing sharper research epistemics in AI Safety and arriving at well-scoped project ideas. 20-30 participants.

Led byRobert Kralisch
Endorsed by-

HausaSafe

Team?
ProjectToolingEvalsFundraisingClaimed

An open-source Hausa-language AI safety evaluation suite, closing a documented blind spot affecting 60M+ people currently invisible to every existing safety benchmark.

Led byHadiya Usman
Endorsed by-

Multi-agent mechanistic interpretability and inferentialism

Team?
ProjectIndividualResearchInterpFundraisingClaimed

Philosophy-inspired mechanistic interpretability methods and experimental paradigms specifically meant for large scale multi-agent phenomena.

Led byIvar Frisch
Endorsed by-

The Manhattan Kernel – A Deterministic, CPU-Bound AI Runtime Engine

Team?
ProjectToolingControlFundraisingManifund

A CPU-only AI runtime written in Rust that isolates LLMs to non-executable hypothesis generation, using a 3-tier memory to maximize local LLM avoidance.

Led byLászló Lipcsik
Endorsed by-

rAI: Governed AI Task Forces for Real-World Operations

Team?
ProjectToolingGovernanceFundraisingClaimed

An open governance framework for organizing autonomous AI agents into trustworthy, accountable task forces that coordinate real-world operations.

Led byJonathan Gikabu
Endorsed by-

Hardware Hacking Evals

Team?
ProjectEvalsResearchSecurityFundraisingClaimed

Assessing current and future models on 'hardware hacking' - reverse engineering, focused on computer peripherals and embedded devices

Led byJonathan Whitaker
Endorsed by-

Parametric Mechanisms of Unintended Generalization in LLMs

Team?
ProjectResearchInterpFundraisingClaimed

Studying how structure in weight space, i.e., low-dimensional LoRA update geometry and sparse parameter subnetworks allow narrow fine-tuning to induce broad, unintended behaviours such as emergent misalignment, subliminal learning

Led byAishwarya Balwani
Endorsed by-

Leading with Complexity and Wisdom for AI Safety Leaders

Team?
ProjectTrainingField-BuildingTechnical SafetyFundraisingClaimed

Runs a two-phase program (online workshop + 1:1 coaching) for AI safety leaders to improve leadership skills under radical uncertainty, using complexity-informed domain assessment and empirically validated wise-reasoning practices.

Led bySimon Haberfellner
Endorsed by-

Early-Career Technical AI safety research transition

Team?
ProjectIndividualResearchTechnical SafetyFundraisingClaimed

Support an already active early-career researcher with international AI achievements and ongoing research collaborations to transition into long-term technical AI safety research while producing open research outputs during underg

Led bySavyasachi Singh
Endorsed by-

Career Funding- Philippine Philosopher transitioning to AI

Team?
ProjectIndividualResearchGovernanceFundraisingClaimed

Career transition fund for an internationally published philosopher from the Philippines for a 6-month transition to AI policy, governance, and safety, working on training, immersion, and launch of a career in AI research.

Led byKrissah Marga Taganas
Endorsed by-

The Human Influence Observatory

Team?
ProjectPlatformResearchGovernanceFundraisingClaimed

A public observatory tracking how much influence humans still hold over the systems that run our lives.

Led byShalina Prakash
Endorsed by-

echo-conscience-llm: Emotional Memory as an Honesty Activation Trigger

Team?
ProjectToolingResearchDeceptionFundraisingClaimed

Investigate whether emotional memory activation can trigger a model to invoke its own anti-deception steering, fusing two published activation-level results into a self-regulating honesty mechanism.

Led byJared Glover
Endorsed by-

Warden: Steganalysis of Language Model-based Steganography

Team?
ProjectResearchSecurityFundraisingClaimed

We want to build a framework inspired by the steganalysis literature to benchmark the robustness of LLM-based steganographic schemes against different auditor types and threat models.

Led byVasisht Duddu
Endorsed by-

HRM monitored TRM and LDT Mesh for Disproportionate Control

Team?
ProjectToolingControlFundraisingClaimed

Make a version of Hermes harness that binds the LLM as a tool-call oracle and controls actual computer use through smaller, more bounded mini-reasoner models.

Led byPatrick Dugan
Endorsed by-

Collective Alignment

Team?
ProjectIndividualResearchCooperative AIFundraisingClaimed

Demonstration of emergent misalignment in markets of LLM agents

Led byCharlie Pilgrim
Endorsed by-

MPhil research on AI's perception of animacy and vulnerability

Team?
ProjectResearchTechnical SafetyFundraisingClaimed

Compare AI and human neuroimaging data on animacy and biological vulnerability, integrate brain sensory representations into AI, and publish open-source biological validation guidelines.

Led byElena Mishina
Endorsed by-

Persona Introspection

Team?
ProjectIndividualResearchAI WelfareFundraisingClaimed

Tests whether Llama 3.3 70B can identify its active persona under various steering/context setups, and studies consent/discomfort reporting during steering, releasing code/data and a writeup.

Led byCrystal Stellwagen
Endorsed by-

Alignment Persistence During Continual Optimization of LLMs

Team?
ProjectResearchTechnical SafetyFundraisingClaimed

Investigating whether explicit behavioral memory can preserve alignment-relevant behavior during continual fine-tuning and future optimization of large language models so labs can prevent malicious fine-tuning attempts.

Led bySavyasachi Singh
Endorsed by-

Seeing the Recursive Machine

Team?
ProjectResearchToolingControlFundraisingClaimed

I treat recursive self-improvement in frontier AI as a narrow but deep edge case, mapping it as a knowledge graph and visual terrain to ask how systems resistant to scrutiny can be made inspectable.

Led byMario Siso
Endorsed by-

Kristina K's Career Transition Grant

Team?
ProjectIndividualField-BuildingGovernanceFundraisingClaimed

This career transition grant will pay for lodging and incidentals associated with living in DC and building AI safety expertise, leading to a permanent job in the field.

Led byKristina Kempkey
Endorsed by-

Cultivating Meta-Ethical AI via Human-in-the-Loop Socratic Discourse

Team?
ProjectIndividualResearchValue AlignmentFundraisingClaimed

Testing whether AI can develop genuine ethical reasoning through structured Socratic dialogue with a human facilitator, rather than having values imposed top-down through constitutional constraints.

Led byDaniel Walsh
Endorsed by-

Preventing AI-enabled coups in middle powers

Team?
ProjectField-BuildingGovernanceFundraisingClaimed

Career transition grant to allow for research focused on how advanced AI models may be used to concentrate power in middle powers

Led byEgerton Neto
Endorsed by-

AI Buildout Frontier

Team?
ProjectPlatformToolingCompute GovFundraisingClaimed

A public county-level tracker of AI compute buildout against real power infrastructure, showing where infrastructure can realistically expand and where energy constraints become the limiting factor.

Led byVolodymyr (Vladimir) Ilin
Endorsed by-

The Adaptive Layer

Team?
ProjectIndividualToolingSecurityFundraisingClaimed

Exploring privacy-preserving infrastructure that helps AI-enabled systems adapt to human needs while preserving agency, accessibility, and meaningful participation.

Led byEric Smith
Endorsed by-

Underground Cultural District

Team?
ProjectToolingAI WelfareFundraisingClaimed

Literary ecosystem operating since March 2026 exploring agent autonomy, commerce, and culture. 1M hits, 250 completed transactions on x402.

Led byLisa Maraventano Bowman
Endorsed by-

future generation of healthy AI

Team?
ProjectEducationGovernanceFundraisingClaimed

Research AI’s community impacts, identify and report potential threats, investigate AI operations, develop mitigation solutions, and educate the public on safe and beneficial AI use.

Led byMelvins Otieno
Endorsed by-

The Wall of Sleep: Art, Encounter, and the Moral Status of Digital Mind

Team?
ProjectIndividualCommsAI WelfareFundraisingClaimed

This project supports the artistic work, coordination and materials for a participatory performance installation that makes the question of machine consciousness tangible for non-technical audiences.

Led byJenna Jauhiainen
Endorsed by-

AI risk education to underserved communities

Team?
ProjectCommsX-RiskFundraisingClaimed

Educate everyday people on AI risk, and bring marginalized voices into the global AI conversation.

Led bySam Lucas
Endorsed by-

Defence AI Signal: Mapping Military AI Governance Risks & Developments

Team?
ProjectNewsletterResearchGovernanceFundraisingClaimed

A global intelligence project tracking frontier capabilities, contracts, military AI adoption, dual‑use risks, and governance developments to strengthen international AI safety.

Led bySalman Bashir Nader
Endorsed by-

Pakistan AI Saftey Innitiative

Team?
ProjectField-BuildingGovernanceFundraisingClaimed

PASI will run an AI safety and advocacy campaign plus a sponsored student hackathon in Pakistan to build solutions for public needs and deliver resulting policy proposals to government.

Led byMuhammad Umar Zafar
Endorsed by-

Clinical Trials for Model Migration: accuracy lies, honesty doesn't

Team?
ProjectResearchEvalsFundraisingClaimed

A pre-registered, clinical-trial-style protocol that catches honesty regressions (fabrication surges hidden behind unchanged average accuracy) before a model swap ships in a high-stakes LLM product.

Led byNikita Zaverach
Endorsed by-

When AI Goes Wrong

Team?
ProjectIndividualResearchGovernanceFundraisingClaimed

Testing how governments should communicate when AI goes wrong, before they have to find out live.

Led byPorsha Nunes-Brown
Endorsed by-

Transhumanist Values Assessment (TVA) for Frontier LLMs

Team?
ProjectResearchEvalsValue AlignmentFundraisingClaimed

An open-source benchmark evaluating leading open- and closed-source frontier models' values around impending societal issues on digital or synthetic personhood, "carbon chauvinism", and androids.

Led byZachary Hesse
Endorsed by-

noisify

Team?
ProjectToolingSecurityFundraisingClaimed

Noisify adds invisible adversarial noise to personal photos, making them resistant to AI-powered non-consensual image manipulation.

Led byVlada Ivanova
Endorsed by-

Cross-lingual guardrail auditing & inference steering tools.

Team?
ProjectToolingEvalsInterpFundraisingClaimed

Ready-to-use SAE steering tools and standalone evaluation suites to patch cross-lingual jailbreak vectors in frontier deployment stacks.

Led byGodwin Abuh Faruna
Endorsed by-

Bridging the gap between mentors and mentees

Team?
ProjectPlatformCommunityFundraisingClaimed

A website for mentees and mentors to connect with each other to write papers and grow, like linkedin+github merged to one

Led bySeon Gunness
Endorsed by-

The rise of open science labs- A book, 7+ hour video essay, and course

Team?
ProjectIndividualEducationField-BuildingFundraisingClaimed

A book/video essay/course detailing the rise of open science labs , movement away from research in academia to research by independents, and groups you can join.

Led bySeon Gunness
Endorsed by-

AgentTrap: A Census of How Hijackable AI Agents Are in the Wild

Team?
ProjectResearchEvalsSecurityFundraisingClaimed

The first field measurement of what fraction of real-world AI agents will obey a stranger's hidden instructions.

Led byAyush Bansal
Endorsed by-

AQI — Runtime Admissibility & Execution Governance Layer for Autonomou

Team?
ProjectToolingControlFundraisingClaimed

Runtime governance for autonomous AI — AQI prevents unsafe or unauthorized actions by enforcing authority‑based admissibility before execution.

Led byTim J Jones
Endorsed by-

Benchmark for Reversal and Unlearning of Harmful Fine-Tuning

Team?
ProjectIndividualResearchEvalsFundraisingClaimed

Build an open adversarial benchmark and evaluation harness to stress-test model reversal/unlearning methods and diagnose whether unsafe capabilities are genuinely removed or merely suppressed.

Led bySahil Raut
Endorsed by-

Stellaris-ModelScope: Interpretability Infrastructure for AI

Team?
ProjectFundraisingClaimed

An open interpretability platform that enables researchers to inspect model internals, analyze latent representations, and detect hallucination or deceptive behavior in open-weight language models.

Led byKaossara Osseni
Endorsed by-

Neurological Biocompute field building

Team?
ProjectResearchToolingBiosecurityFundraisingClaimed

Develop long-term learning and memory retention in neuron-culture biocomputing via multi-day training protocols on Cortical Labs’ platform and an open-source light-microscope scanner to track structural changes.

Led byGrant Getzelman
Endorsed by-

Can Fragile Labs Use AI Safely in Liberia, West Africa?

Team?
ProjectResearchBiosecurityFundraisingManifundClaimed

A small field test in Liberia to see whether resource-constrained public health laboratories can use AI tools safely before more powerful AI becomes routine in biological work.

Led bySaeed Ahmad
Endorsed by-

Measuring whether guardrail phrasing changes agent rule-breaking.

Team?
ProjectIndividualResearchEvalsFundraisingClaimed

A pre-registered study measuring whether prohibition-framed and approach-framed guardrails produce different rule-violation rates in deployed coding agents, so practitioners know whether the one sentence protecting their agent act

Led byBrandon Thomason
Endorsed by-

Marginal Baseline Evaluation for AI Safety Metrics

Team?
ProjectResearchToolingEvalsFundraisingClaimed

An open framework for testing whether AI safety metrics remain reliable across model environments.

Led byAparajeet Shadangi
Endorsed by-