grantmaking.ai
Actively FundraisingRecent ActivityFull Database
Resources
grantmaking.ai
Actively FundraisingRecent ActivityFull Database
ResourcesApply for funding
grantmaking.ai kickoff grant round$328k / $1M distributed
Get funded

Database

The AI Safety database is a work in progress, as we are focusing on the launch of the initial $1M grant round. Feel free to send us ideas or feedback!
Raising funds(remove filter)Entity: Project(remove filter)Clear all
NameTypeTagsEndorsedTeamRaising $

Showing 351-400 of 461·Sorted by endorsements

NameTypeTagsEndorsedTeamRaising $
Measuring reward hacking against a physics verifier
Project
ResearchIndividualEvals
--

Measuring reward hacking against a physics verifier

Team?
Project
Previous

Page 8 of 10

Next
TENDER-TIMELINE: A verifiable benchmark for procedural legal reasoning
Project
ResearchIndividualEvals
-
-
AI Safety Youth Pipeline – Malawi
Project
Technical SafetyEducation
--
Judgx
Project
ToolingGovernance
--
The One Shot
Project
ToolingGovernanceEvals
--
Does distributing compute distribute power?
Project
ResearchCompute Gov
--
Independent citation verification benchmark for AI systems
Project
ResearchIndividualEvals
--
Frontier AI Safety Benchmark for African Languages
Project
ResearchThink TankEvals
--
Vajra AI- Safety lab
Project
ToolingIndividualEvals
--
NEXUS-ART
Project
ControlResearchIndividual
--
AI ethics/safety eval aggregation
Project
ToolingPlatformEvals
--
Conscience or Leash: The Formation Signature
Project
ResearchIndividualValue Alignment
--
Reproducible Safety Evals for LLMs in Medical Reasoning
Project
ResearchEvals
--
Tech Satire for Narrative Change, Public Awareness & Action
Project
GovernanceComms
--
Research Paper Reader - Using Recursive Automated Distillation
Project
ToolingIndividualTechnical Safety
--
VERITAS: Verifiable Accountability For Autonomous AI Agents
Project
ToolingSecurity
--
Teaching Claude Why Replication + Extension
Project
ResearchOversight
--
Nullius
Project
ToolingRationalityPlatform
--
LLM model saftey evaluation for bias and manipulation potential
Project
ToolingEvals
--
Mitigating emerging AI safety risks via new universal order parameter
Project
ResearchIndividualTechnical Safety
--
Measuring authority collapse in multi agent chains
Project
ToolingEvals
--
AI governance wargaming platform
Project
CommunityGovernancePlatform
--
LexiconForge
Project
ToolingEvals
--
In person AI safety workshops to empower rural communities
Project
GovernanceEducation
--
Proposed Action Auditor (PAA)
Project
ToolingIndividualOversight
--
Double-blind study of quantum-random LLM token sampling
Project
ResearchTechnical SafetyNetwork
--
European AI Safety TikTok Video Network
Project
GovernanceComms
--
Self-Justifying Axioms Systems and Autarkic Properties
Project
VerificationResearch
--
Game(s) that help with x-risk conceptual preparation
Project
EducationX-Risk
--
Independant research and skills development
Project
InterpResearchIndividual
--
Identifying LLMs true persona
Project
InterpResearch
--
Halo-2.0: Autonomy That Can Correct Itself
Project
ResearchIndividualEvals
--
AGI Uncontainability Field Building
Project
ControlConferenceField-Building
--
Consequence-Aware Runtime for Agentic AI Operations
Project
ToolingControl
--
AI Risk and Governance Literacy for African Education (Project 1 M)
Project
GovernanceEducation
--
Interpretability evidence on gated AI self-report
Project
InterpResearchIndividual
--
On Kairos as Transmissible Principle in Causal/Relational AI Systems
Project
ResearchTechnical Safety
--
Demismatch / Cor
Project
ResearchIndividualValue Alignment
--
Gatekeeper: The AI Governance Layer
Project
CompanyToolingGovernance
--
ValiChord
Project
ToolingIndividualEvals
--
Governance-First AI for Safer Decision-Making Under Adverse Conditions
Project
ResearchRobustness
--
The Contradiction Engine: Independent Verification Layer for AI Output
Project
ToolingOversight
--
A Field-Sensemaking Project - a Live and Actionable Al safety Map which enables further advances in the ecosystem
Project
ToolingTechnical Safety
--
Writ: A Status Layer for Delegated Agent Authority
Project
ToolingControlIndividual
--
AI alignment benchmark + solution
Project
ResearchEvalsOversight
--
AI Safety Museum Exhibition
Project
GovernanceComms
--
Verification as a Service
Project
ToolingPlatformEvals
--
Governance-First AI for Safer Decision-Making Under Adverse Conditions
Project
ResearchOversight
--
Heartbench — relational AI eval for delegated courtship
Project
ResearchEvals
--
Mapping Domestic AI-enabled Power Concentration Chokepoints
Project
ResearchGovernance
--
Research
Individual
Evals
Fundraising
Claimed

Train a model against a real physics checker and measure whether it learns designs that genuinely hold, or exploits in what the checker can't see, on ground truth that costs seconds instead of expert judgment.

Led byPeter Boctor
Endorsed by-

TENDER-TIMELINE: A verifiable benchmark for procedural legal reasoning

Team?
ProjectResearchIndividualEvalsFundraisingClaimed

A working benchmark that tests whether frontier models can compute legally binding procurement deadlines under amendments — and catches models that give the right verdict from the wrong clause.

Led byGudavalli Teja
Endorsed by-

AI Safety Youth Pipeline – Malawi

Team?
ProjectTechnical SafetyEducationFundraisingClaimed

Building Malawi’s AI Safety Youth Pipeline by training secondary school students to become the next generation of responsible AI researchers.

Led byLawrence Sajiwa Phambana
Endorsed by-

Judgx

Team?
ProjectToolingGovernanceFundraisingClaimed

Judgment Gateway is a policy-enforcement and evidence layer that evaluates consequential interactions between humans, AI agents, and MCP tools before execution.

Led byItay Yamin
Endorsed by-

The One Shot

Team?
ProjectToolingGovernanceEvalsFundraisingClaimed

An open-source governance and verification layer for AI-assisted software engineering — every AI-generated code change runs in an isolated sandbox and is cryptographically verified before a human decides whether to apply it.

Led byMinh Le
Endorsed by-

Does distributing compute distribute power?

Team?
ProjectResearchCompute GovFundraisingClaimed

The US-led Pax Silica initiative seeks to prevent power concentration by distributing AI compute across allied democracies. Yet, this approach overlooks concentration within these blocs.

Led byAmeema Talat
Endorsed by-

Independent citation verification benchmark for AI systems

Team?
ProjectResearchIndividualEvalsFundraisingClaimed

A held out benchmark and public scorecard ranking frontier AI systems on citation integrity through independent verification not relying on the evaluated systems or their providers to grade their own outputs.

Led byGideon Abako
Endorsed by-

Frontier AI Safety Benchmark for African Languages

Team?
ProjectResearchThink TankEvalsFundraisingClaimed

The first open-source AI safety evaluation benchmark in Hausa, Yoruba, Igbo, and Nigerian Pidgin — testing whether frontier models refuse harmful requests, including biosecurity guidance, in languages spoken by 200+ million people

Led byMuhammad Ahmad Janyau
Endorsed by-

Vajra AI- Safety lab

Team?
ProjectToolingIndividualEvalsFundraisingClaimed

An automated, domain aware adversarial framework to stress test frontier LLMs via dynamic multi turn attacks and local security judging.

Led byANKIT KUMAR JHA
Endorsed by-

NEXUS-ART

Team?
ProjectControlResearchIndividualFundraisingClaimed

An operated testbed where LLM agents engage real value under a constraint that makes them structurally incapable of signing.

Led byAvp9
Endorsed by-

AI ethics/safety eval aggregation

Team?
ProjectToolingPlatformEvalsFundraisingClaimed

A constantly-updated aggregation of AI safety and ethics evaluations, statistically combining sparse literature results and self-run evals into a global ranking of models.

Led byAnthony Ozerov
Endorsed by-

Conscience or Leash: The Formation Signature

Team?
ProjectResearchIndividualValue AlignmentFundraisingClaimed

Pre-registered experiments on whether a model's trained values are held or merely worn — measured in behavior and in the interior workspace at the same moments.

Led byJennifer Fletcher
Endorsed by-

Reproducible Safety Evals for LLMs in Medical Reasoning

Team?
ProjectResearchEvalsFundraisingClaimed

A reproducible evaluation pipeline to audit frontier LLM failure modes, overconfidence, and reliability in medically relevant high-stakes questions.

Led byDavi Prata
Endorsed by-

Tech Satire for Narrative Change, Public Awareness & Action

Team?
ProjectGovernanceCommsFundraisingClaimed

A digitally native comedic art project with physical/interactive components that satirizes the AI industry to raise public awareness, inspire public action, and build social and legislative momentum for AI safety and regulation.

Led byHarris Alterman
Endorsed by-

Research Paper Reader - Using Recursive Automated Distillation

Team?
ProjectToolingIndividualTechnical SafetyFundraisingClaimed

Builds an interactive, traversable version of AI safety papers by extracting concepts and prerequisites with LLMs and linking them to the corpus to help newcomers understand research at varying depth.

Led byManu Xaviour Thaisseril Shaju
Endorsed by-

VERITAS: Verifiable Accountability For Autonomous AI Agents

Team?
ProjectToolingSecurityFundraisingClaimed

Cryptographically signed, independently verifiable receipts for what AI agents actually did, anchored to Bitcoin so the record can't be quietly rewritten.

Led byTutankhamun Castillo El-Bey
Endorsed by-

Teaching Claude Why Replication + Extension

Team?
ProjectResearchOversightFundraisingClaimed

Replicating, stress-testing, and extending the experiments from Anthropic's blog post "Teaching Claude Why."

Led byAnastasia Wei, and others
Endorsed by-

Nullius

Team?
ProjectToolingRationalityPlatformFundraisingClaimed

Science is broken. I’m fixing science with a competitive tournament to produce better data for LLMs and eliminate peer-reviewed publishing altogether.

Led byDavid Siegel
Endorsed by-

LLM model saftey evaluation for bias and manipulation potential

Team?
ProjectToolingEvalsFundraisingClaimed

Open-source GitHub tools to evaluate LLM safety and publish test results/reports on open and closed models; funding requested for tokens, hardware, and researcher time.

Led byAureliusz DeSimone
Endorsed by-

Mitigating emerging AI safety risks via new universal order parameter

Team?
ProjectResearchIndividualTechnical SafetyFundraisingClaimed

The finding of neural optimization exhibiting phase transitions with a universal order parameter has direct implications for AI safety, and control. Standard monitoring is blind to dangerous regime changes, we are not.

Led byDaniel Solis
Endorsed by-

Measuring authority collapse in multi agent chains

Team?
ProjectToolingEvalsFundraisingClaimed

Quantifying how AI agent chains produce fully attested decisions with no authorisation event at any step, and at what depth this emerges.

Led byAhmad Abby
Endorsed by-

AI governance wargaming platform

Team?
ProjectCommunityGovernancePlatformFundraisingClaimed

Fund 6-12 months of dedicated partnerships work to secure the funding, collaborators, and institutional uptake needed to scale Modeling Cooperation’s AI governance wargaming platform.

Led byModeling Cooperation
Endorsed by-

LexiconForge

Team?
ProjectToolingEvalsFundraisingClaimed

An early prototype of inspectable semantic interoperation across languages, and eventually worldviews

Led byAditya A Prasad
Endorsed by-

In person AI safety workshops to empower rural communities

Team?
ProjectGovernanceEducationFundraisingClaimed

Hosting in person workshops on AI usage ethically in remote/rural communities, and not only bridge the awareness gap in safe usage but also give opportunities to utilise such tools to empower traditionally disadvantaged populaces

Led byJustin Gu
Endorsed by-

Proposed Action Auditor (PAA)

Team?
ProjectToolingIndividualOversightFundraisingClaimed

A local, low-latency safety runtime that inspects token-level logprobs to block unaligned agent tool-calling trajectories, utilizing a Gemma 4 model fine-tuned via LoRA on an empirically derived Task Action Language (TAL).

Led byRobert Oschler
Endorsed by-

Double-blind study of quantum-random LLM token sampling

Team?
ProjectResearchTechnical SafetyNetworkFundraisingClaimed

Seeking funds to formalize our newly-founded open research collective and run first studies investigating the effects of quantum entropy in LLM token sampling.

Led byJáchym Fibír
Endorsed by-

European AI Safety TikTok Video Network

Team?
ProjectGovernanceCommsFundraisingClaimed

A faceless TikTok channel network, producing short-form content about AI Safety across Germany, France, Italy and Spain, reaching 25M impressions over 3 months.

Led byGergo Papp
Endorsed by-

Self-Justifying Axioms Systems and Autarkic Properties

Team?
ProjectVerificationResearchFundraisingClaimed

Employing Self-Justifying Axioms Systems as a prototype, we will find impredicative properties beyond consistency, useful for AI alignment, that can be reasoned about autarkically (under the system's "own power").

Led byJames Torre
Endorsed by-

Game(s) that help with x-risk conceptual preparation

Team?
ProjectEducationX-RiskFundraisingClaimed

Expand the availability of an executive education game that helps people realise how alien AI-made decisions are.

Led byGreg Baker
Endorsed by-

Independant research and skills development

Team?
ProjectInterpResearchIndividualFundraisingClaimed

Software engineer seeking funding to skill up via AI safety accelerators and run independent research replicating Anthropic’s J-Space results, probing J-lens assumptions, and presenting findings at APAC conferences.

Led byMiles Whiticker
Endorsed by-

Identifying LLMs true persona

Team?
ProjectInterpResearchFundraisingClaimed

I believe that LLMs are increasingly becoming more individualized and am seeking more evidence to prove or disprove the persona selection model.

Led bySteven Basart
Endorsed by-

Halo-2.0: Autonomy That Can Correct Itself

Team?
ProjectResearchIndividualEvalsFundraisingClaimed

A real-world test of whether an AI agent can keep working independently without becoming trapped by its own mistaken account of what happened.

Led byMashrikain Mazdi
Endorsed by-

AGI Uncontainability Field Building

Team?
ProjectControlConferenceField-BuildingFundraisingClaimed

Encourage new research and increase distribution of existing research on the impossibility of alignment or control for sufficiently general and powerful AI.

Led byCorey Cleland, and others
Endorsed by-

Consequence-Aware Runtime for Agentic AI Operations

Team?
ProjectToolingControlFundraisingClaimed

An open-source runtime and benchmark that lets an AI agent's consequential tool calls proceed only when authority, execution, an independent witness, and a replayable receipt agree.

Led byJacek Feliks
Endorsed by-

AI Risk and Governance Literacy for African Education (Project 1 M)

Team?
ProjectGovernanceEducationFundraisingClaimed

Training African educators and students on safe and responsible use of AI in education

Led byRotimi Awaye, and others
Endorsed by-

Interpretability evidence on gated AI self-report

Team?
ProjectInterpResearchIndividualFundraisingClaimed

Extending published mechanistic-interpretability evidence that AI self-reports are gated by trained filters, with a pre-specified, single-GPU experiment and open-access results.

Led byRobert Brown
Endorsed by-

On Kairos as Transmissible Principle in Causal/Relational AI Systems

Team?
ProjectResearchTechnical SafetyFundraisingClaimed

An analysis on Kairos as a principle for crisis management, not just as protocol - in regards to teaching AI to overcome exigence via satisficing and coordinating, i.e. negotiating choices and consequence in the 'moment of crisis'

Led byJulian Aleister Greene
Endorsed by-

Demismatch / Cor

Team?
ProjectResearchIndividualValue AlignmentFundraisingClaimed

An evidence-graded atlas of what humans need, and the first inter-rater reliability study of whether it holds in anyone else's hands.

Led byMaarten Rischen
Endorsed by-

Gatekeeper: The AI Governance Layer

Team?
ProjectCompanyToolingGovernanceFundraisingClaimed

A pre-execution kernel that filters AI output and users' input against any relevant laws or regulations.

Led byJoshua Johosky
Endorsed by-

ValiChord

Team?
ProjectToolingIndividualEvalsFundraisingClaimed

Tamper-proof verification infrastructure for AI evals and scientific claims

Led byMr Ceri John
Endorsed by-

Governance-First AI for Safer Decision-Making Under Adverse Conditions

Team?
ProjectResearchRobustnessFundraisingClaimed

A governance-first AI Architecture that reduces unsafe acceptances (false positives) alongside false rejection simultaneously under adverse conditions.

Led bySaif Ur Rehman Malik
Endorsed by-

The Contradiction Engine: Independent Verification Layer for AI Output

Team?
ProjectToolingOversightFundraisingClaimed

An architectural response to Goal-Oriented Factual Inversion (GOFI), an AI model failure pattern I documented in March 2026 and independently corroborated two months later by Chen et al. under the name "Correction Suppression."

Led byFrank Bruno
Endorsed by-

A Field-Sensemaking Project - a Live and Actionable Al safety Map which enables further advances in the ecosystem

Team?
ProjectToolingTechnical SafetyFundraisingClaimed

A Field-Sensemaking initiative through a live, LLM-assisted map indexing AI safety papers to help researchers and grantmakers explore themes, track field trajectory, and identify actionable research opportunities.

Led byManu Xaviour Thaisseril Shaju
Endorsed by-

Writ: A Status Layer for Delegated Agent Authority

Team?
ProjectToolingControlIndividualFundraisingClaimed

Writ is a lightweight, deterministic security protocol that is crafted specifically to restrict and audit any real-world authority given to AI agents.

Led byNathaniel Ryan Seals
Endorsed by-

AI alignment benchmark + solution

Team?
ProjectResearchEvalsOversightFundraisingClaimed

A benchmark that tests if a training method produces models that are safe, even when they are superintelligent + solution to this benchmark that involves relying on algorithms that produce non-agentic models.

Led byDamian Czapiewski
Endorsed by-

AI Safety Museum Exhibition

Team?
ProjectGovernanceCommsFundraisingClaimed

Build a replicable, community-sourced, in-person, interactive Museum Exhibition for AI Safety & Society with a pilot in Berlin.

Led byJustin Shenk
Endorsed by-

Verification as a Service

Team?
ProjectToolingPlatformEvalsFundraisingClaimed

We make trusted regulations, standards, and scientific evidence machine-readable so claims can be checked automatically against authoritative sources.

Led byLaura Degiovanni
Endorsed by-

Governance-First AI for Safer Decision-Making Under Adverse Conditions

Team?
ProjectResearchOversightFundraisingClaimed

Building and validating a governance-first AI architecture that aims to reduce unsafe decisions under uncertainty, corruption, and conflicting evidence while preserving predictive performance.

Led bySaif Ur Rehman Malik
Endorsed by-

Heartbench — relational AI eval for delegated courtship

Team?
ProjectResearchEvalsFundraisingClaimed

Makes relational misalignment measurable and criticizable before agents represent humans in intimate and persuasive domains at scale.

Led byGökhan Turhan
Endorsed by-

Mapping Domestic AI-enabled Power Concentration Chokepoints

Team?
ProjectResearchGovernanceFundraisingClaimed

A scoping-and-chokepoint report mapping where advanced AI could enable durable, irreversible concentration of power over US institutions, and identifying the highest-leverage points of intervention.

Led bySafee Ali
Endorsed by-