- How would we know? Reflections on trying to make AI go well amid deep uncertainty
- Treading the narrow corridor: a hypothesis about epistemics and the design, training, and deployment of advanced AI systems
- Applying systems thinking and evaluation science to AI safety: four demonstrations
- A score is not understanding: toward a richer toolkit for model evaluations
- Assumptions & weak signals: how to build a stronger empirical basis for reasoning about the pace of AI progress