Funding ask
The money is spent on (ideal ask):
Researcher time (developing the simulated AI lab, orga): $37000
LLM assistants (for research, coding, writing; no training): $6000
Compensation for ext. researcher's evaluations: $7000
No travel. No GPU. No overhead.
Breakdown by steps:
- development of the simulated AI lab for corrigibility, shutdown, deception, and evals plus the associated materials
- independent evaluation by Redwood (covers deception and evals; this is the minimum package)
- MATS use test
- BlueDot use test
- independent evaluation by CHAI or Christiano (covers shutdown & corrigibility)
- development of the tiling disambiguation in the simulation
- development of the Inner alignment disambiguation in the simulation
- addressing feedback on republish, no retest
- independent evaluation by e.g. John Wentworth on selection theorems as inner alignment
- independent evaluation by e.g. Scott Garrabrant on shutdown / corrigibility / tiling
- independent evaluation by e.g. Vanessa Kosoy
- continuing to maintain for one year (~one day a month, incorporating new research and feedback)