Contents
Updates
- New at IAFF: An Untrollable Mathematician
- New at AI Impacts: 2015 FLOPS Prices
- We presented “Incorrigibility in the CIRL Framework” at the AAAI/ACM Conference on AI, Ethics, and Society.
- From MIRI researcher Scott Garrabrant: Sources of Intuitions and Data on AGI
News and links
- In “Adversarial Spheres,” Gilmer et al. investigate the tradeoff between test error and vulnerability to adversarial perturbations in many-dimensional spaces.
- Recent posts on Less Wrong: Critch on “Taking AI Risk Seriously” and Ben Pace’s background model for assessing AI x-risk plans.
- “Solving the AI Race“: GoodAI is offering prizes for proposed responses to the problem that “key stakeholders, including developers, may ignore or underestimate safety procedures, or agreements, in favor of faster utilization”.
- The Open Philanthropy Project is hiring research analysts in AI alignment, forecasting, and strategy, along with generalist researchers and operations staff.
This newsletter was originally posted on MIRI’s website.
Our newsletter
Regular updates about the Future of Life Institute, in your inbox
Subscribe to our newsletter and join over 20,000+ people who believe in our mission to preserve the future of life.
Recent newsletters
Future of Life Institute Newsletter: The AI ‘Shadow of Evil’
Notes from the Vatican on AI; the first International AI Safety Report; DeepSeek disrupts; a youth-focused video essay on superintelligence, by youth; grant and job opportunities; and more!
Maggie Munro
31 January, 2025
Future of Life Institute Newsletter: 2024 in Review
Reflections on another massive year; major AI companies score disappointing safety grades; our 2024 Future of Life Award winners; and more!
Maggie Munro
31 December, 2024
Future of Life Institute Newsletter: Tool AI > Uncontrollable AGI
Max Tegmark on AGI vs. Tool AI; magazine covers from a future with superintelligence; join our new digital experience as a beta tester; and more.
Maggie Munro
2 December, 2024
All Newsletters