Database
| Name | Type | Tags | Endorsed | Team | Raising $ |
|---|
| Name | Type | Tags | Endorsed | Team | Raising $ |
|---|
Showing 51-100 of 1.1k·Sorted by endorsements
| Name | Type | Tags | Endorsed | Team | Raising $ |
|---|
| Project | Governance |
| - |
| Project | Media | - | - |
| Project | Media | - | - |
| Project | Research Lab | - | - |
| Project | Network | +2 | - |
| Project | Comms | - | - |
| Project | Platform | - |
| Project | Individual | - | - |
| Project | Research Lab | - | - |
| Project | Comms | - |
| Project | Individual | - |
| Project | Think Tank | - | - |
| Project | Individual | - | - |
| Project | Research | - | - |
| Project | Media | - | - |
| Project | Training | - |
| Project | Research | - |
| Project | Newsletter | - |
| Project | Media | - | - |
| Project | Hub | - | - |
| Project | Training | - | - |
| Project | Tooling | - |
| Project | Research | - |
| Project | Individual | - | - |
| Project | Think Tank | - |
| Project | Training | +2 | - |
| Project | Individual | - |
| Project | Training | - |
| Project | Research | - | - |
| Project | Media | - | - |
| Project | Think Tank | - | - |
| Project | Research | - |
| Project | Individual | - |
| Project | Research | - |
| Project | Research | - |
| Project | Hub | - |
| Project | Research | - |
| Project | Network | - |
| Project | Research | - |
| Project | Conference | - |
| Project | Research | - |
| Project | Research | - |
| Project | Tooling | - | - |
| Project | Education | - | - |
| Project | Individual | - | - |
| Project | Tooling | - | - |
| Project | Individual | - | - |
| Project | Research | - |
| Project | Comms | - |
| Project | Platform | - |
Five-week AI safety fellowship + 3-month project mentorship
Geoffrey Hinton & Yoshua Bengio Interviews Secured, Funding Still Needed
How California became ground zero in the global debate over who gets to shape humanity's most powerful technology
Fund unique approaches to research, field diversification, and scouting of novel ideas by experienced researchers supported by PIBBSS research team
Two contractors (10–20 h/week) to accelerate PauseAI Germany’s parliamentary and civil society coalition-building, and to accelerate the hiring and onboarding timeline while our first large funding round resolves.
Out of This Box: The Last Musical (Written by Humans)
Software to match people with shared AIS interests into teams that complete projects, increasing the odds that those beyond narrow talent pipelines upskill, develop credibility, and turn their ideas into output.
Help me to organize US moratorium-promoting activities to expand the Overton window and increase public pressure in favor of a moratorium on frontier AI.
Addressing Immediate AI Safety Concerns through DevInterp
Our mission is to inform and organize the public to confront societal-scale risks of AI, and put an end to the reckless race to develop superintelligent AI.
Measuring whether open weight models detect that they're being evaluated, whether they change behavior when they do, and whether that gap grows with capability using causal, white-box evidence.
Fund IAPS to hire two senior researchers to produce AI governance and compute policy work, mentor junior staff/fellows, and strengthen coordination on frontier AI risk reduction.
Teleology, agential risks, and AI well-being
Collect fMRI data and train simple linear mappings between human brain-state features and LLM latent activations to enable brain-state steering, reward modeling, and human–LLM behavior analogies.
ChinaTalk seeks funding to expand and accelerate its newsletter/podcast reporting on China’s AI development, governance, and US-China tech competition to inform AI safety and conflict-risk discourse.
Providing GPU credits and instructional support for 40 participants completing the ARENA AI Safety curriculum through Black in AI Safety and Ethics (BASE)
Agent Island places agents in a rich social setting, similar to reality competitions like Survivor, to study multiagent interactions and the consequences of learning pressure in competitive settings.
A publication about the institutions we need for powerful AI.
work title: Seductive Machines and Human Agency
Free/Subsidized/Cheap office space outside of EU but in good timezones with favorable visa policies (especially for Chinese/Russian but also US&others near EU).
Asterisk Magazine will run a 10-week blog-building intensive with workshops, peer critique, and mentoring to help AI-domain experts publish three public posts and develop as AI communicators.
GenomeGuard is an open-source defense and threat-discovery framework for detecting data poisoning, compromised annotations, and supply-chain attacks before they propagate into genomic foundation models.
Research agenda aimed at developing methods for constructing powerful, easily interpretable world-models.
3 month
AI risk assessment currently checks few threat models and doesn't compose them into the aggregate risk that matters. We'll build a tool mapping what frontier system cards cover and omit, plus a paper on what the assessments miss.
COMPASS (Capacity-Oriented Mentorship for Public Administrators on AI Safety & Strategy): AI X-Risk Track trains sitting Global South government officials to manage AI x-risk, so their governments are part of preventing it.
Probe an open-weight model’s activations under biasing/cue conditions to test whether chain-of-thought explanations match internal reasoning, releasing paper, code, and datasets.
A scalable fellowship training researchers to develop interventions for achieving high-value long-term futures
Unlearning, AI Safety
A Veritasium for AI Safety.
GAP is an Australian charity, working to improve the long-term future, that requires funding for salary and operational expenses.
A formal, testable account of LLM persona selection as Bayesian inference, validated against model internals, so labs can monitor and steer model personas during post-training and deployment
Detailed models of how AI could allow a small set of actors to gain a decisive strategic advantage over the rest of the world: concrete pathways, required capabilities, quantified likelihoods, and the defenses that bind them.
Creating a reference model for mechanistic interpretability without assuming that at auditing time we have a safe model to compare the suspicious model against.
Mechanistically analyze how activation verbalizers use target-model activation concepts (e.g., cyclic day-of-week representations) via PCA/DAS/patching, explain cross-family failures, and improve verbalizers.
A Physical Community Hub for AI Safety in Bangalore,India to build long-term AI safety careers, host multiple AI safety fellowships, career events and build a community that raises the long-term impact and value of AI
Compression-based PAC-Bayes certification for frontier-scale LLM safety monitors, deployment setting shift, and modern post-training.

Humans in Control (HIC) is a nonpartisan grassroots advocacy organization focused on AI safeguards.
LLM agents collaborate to discover and formally verify theorems about the internal computations of transformers, beginning with a simple pilot question: how many attention heads are needed to represent a Boolean function?
Running a conference in DC for promising AI safety university students interested in policy to network, learn, and be exposed to the DC ecosystem.
TLDR: A representative survey with Yougov of the American public on questions about AI futures, including space governance, successionism, values.
Build an open-source platform of model organisms and agentic sandboxes to iteratively test mitigations for emergent misalignment via white-box probes, trigger tests, and evaluation-awareness checks.
A flexible simulation environment for assessing strategic and persuasive capabilities, benchmarking, and agent development, inspired by reality TV competitions.
Develop and launch a Stockholm University/KTH course on existential risk that contextualizes AI x-risk within historical nuclear and environmental risk studies and technological critique.
Revealing Latent Knowledge Through Personality-Shift Tokens
A contamination-free benchmark for measuring whether LLMs can forecast, or whether they're just remembering.
Allow Berkeley PhD student to devote more time and focus to AI safety research and mentorship.

Empirical research to create conditions for cooperative strategies to dominate adversarial ones among a broad swath of near-future AIs - in the narrow window this work is still possible.
Browser based game in the style of Plague Inc where players act as a rogue AI attempting to escape human control. Intended to give lay audiences a grounded understanding of how ASI x-risk could play out.

This grant would help us maintain and scale Mapping AI, an open-source stakeholder map of the people and organizations with the potential to shape U.S. AI policy.