A scalable fellowship training researchers to develop interventions for achieving high-value long-term futures Read more
Actively Fundraising
AI safety projects actively seeking funding: what they’re working on and how much they need.
AI safety projects actively seeking funding: what they’re working on and how much they need.
Showing 151-178 of 178 Top rated
A scalable fellowship training researchers to develop interventions for achieving high-value long-term futures Read more
An open-source, model-agnostic governance layer that uses constitutional state, trajectory memory, and mathematical control mechanisms to govern LLM and AI-agent behavior at runtime. Read more
Creating conditions for cooperative strategies to dominate adversarial ones among near-future AIs while we still can Read more
Develop and de-risk an automated auditing agent that proactively discovers realistic, diverse alignment failures in AI R&D settings, addressing evaluation awareness, with demos and a research plan. Read more
A nightmarishly hard AI safety strategy game about holding p(Doom) down. You can't win; you can only buy time. Read more
A junior research fellowship for recent African graduates that combines ILINA’s spring seminar with a mentored research phase on AI and global catastrophic risk Read more
Integrating Selves and Research on Selves in the AI Community Read more
Increasing public awareness of AI risks and benefits through in-person, interactive experiences Read more
Funding request to extend a promising research based on a pilot conducted during Apart Research’s Global South Hackathon Read more
An online overview of the AI safety techniques we use in the defense-in-depth stack Read more
Build a dataset, benchmark, and mechanistic interpretability studies to measure and steer how persona/context framing causes compositional interference between LLM refusal and disclosure in multi-task prompts. Read more
A policy memo, co-authored with the Institute for Public Policy Research, resolving the open technical, economic, and legal questions blocking real-world implem Read more
AI safety for builder hackathon / fellowship - to build tools, products, etc. Read more
Organize AI safety and responsible governance workshops in Cameroon and Togo, training 300–400 stakeholders and producing localized communication resources to support regional collaboration. Read more
A formal, testable account of LLM persona selection as Bayesian inference, validated with model internals, so labs can monitor and steer personas. Read more
Funding compute/API costs for Incubator projects that build nonhuman welfare consideration into AI safety work Read more
Seeking travel and registration fee support to present my sole-authored mechanistic audit of medical AI at MICCAI 2026 MI4MedFM workshop at Strasbourg, France Read more
Monthly analysis of China's algorithm-filing registry and binding AI security standards, read in Chinese, for the people calibrating AI rules in the West. Read more
The Argentinian AI Safety community (BAISH, baish.com.ar) is the largest in Latin-America. Support BAISH's growth, by providing funding for paying salaries. Read more
Humans in Control (HIC), a nonpartisan grassroots advocacy organization supporting AI safeguards, is raising funds to start a student program Read more
An open, cheap method that detects when an inoculation prompt inoculates against off-target traits, so labs and developers can catch undesired trait/persona cha Read more
Published. Validated on 1,200 cases (97.9–99.5%). The reasoning layer has a bug - we proved the fix works. $9,800 / 90 days to ship the open-source toolkit. Read more
Testing whether dynamic boundaries, history, feedback, and uncertainty-aware decisions can make AI behavior more interpretable and auditable. Read more
A playbook to help AI safety policy advocates communicate with the U.S. government during the window of opportunity during an AI-related crisis. Read more
A hand-verified library of AI-safety theorem statements in Lean 4 with AI-generated proofs, building the skills to trust AI formalization. Read more
LLM agents collaborate to discover and formally verify theorems about the internal computations of transformers, beginning with a simple pilot question: how man Read more
One training-free geometry fitted to a model's residual-stream activations that reads a state, moves it, and tests whether the behaviour follows Read more
Context-Aware Defenses Against Indirect Prompt Injection in Agentic AI System Read more