ChatGPT Recommends
AI SafetyAsked to recommend a AI Safety reading list, ChatGPT came up with the following. Interestingly, it didn’t provide hot links to everything. A pretty obvious thing to do, but tedious so it just decided to omit that step. A perfect example of how AI works these days (“what can I get away with?”).
A Curated AI Safety Reading List
1. Daily / Weekly News
These answer:
What happened this week?
MIT Technology Review (AI)
Probably the best mainstream technology journalism covering AI.
Strengths:
- careful reporting
- less hype than most technology press
- good investigative work
- covers governance, labor, security and frontier models
Reuters AI
Probably the most reliable factual reporting.
Strengths
- announcements
- government actions
- company developments
- regulatory changes
Excellent source for tracking chronology.
Financial Times (AI)
Excellent for:
- AI business
- compute
- capital
- geopolitics
- industrial competition
Very useful for understanding deployment pressure.
The Economist — Artificial Intelligence
Best for
- long-term trends
- economics
- geopolitical framing
The Guardian — AI
Worth reading because it regularly publishes
- commentary
- ethics
- labour
- existential risk
- interviews with safety researchers
It also gives visibility to viewpoints that other outlets sometimes ignore.
2. Frontier Research Organizations
These answer
What are the labs actually discovering?
Anthropic Research
Among the strongest sources for
- interpretability
- alignment
- model psychology
- constitutional AI
- evaluation
Excellent technical summaries.
OpenAI Research
Important because many capabilities arrive here first.
Read:
- System Cards
- Preparedness reports
- Safety evaluations
Google DeepMind Publications
Especially
- agency
- planning
- world models
- robotics
- safety
UK AI Security Institute (AISI)
Perhaps the most consistently careful government technical work on frontier evaluations. Increasingly indispensable for anyone interested in AI safety.
US Center for AI Standards and Innovation (CAISI)
Watch for
- evaluations
- standards
- measurement
- government capability assessments
3. Think Tanks and Research Institutes
These answer
How are serious researchers thinking about AI risk?
AI Futures Project
Formerly Forethought.
Excellent forecasting work.
METR
(Model Evaluation & Threat Research)
Leading work on measuring dangerous capabilities.
Redwood Research
Particularly important for
- control
- alignment
- deceptive behaviour
- alignment-faking
Alignment Research Center (ARC)
Foundational alignment work.
Centre for the Governance of AI (GovAI)
Outstanding governance research.
Especially useful for
- compute governance
- international coordination
- policy design
Future of Humanity Institute (legacy)
Although no longer active, its papers remain foundational.
Future of Life Institute
Policy advocacy.
Good interviews and reports.
RAND
Increasingly active on frontier AI governance.
Brookings
Excellent mainstream policy analysis.
OECD AI Observatory
International policy tracking.
4. Individual Commentators
These answer
How should we interpret events?
Gary Marcus https://garymarcus.substack.com
Probably the most important skeptical voice.
Strengths
- challenges hype
- emphasizes reasoning
- hybrid AI
- engineering realism
Often right for different reasons than AI safety advocates. Reading him prevents tunnel vision.
Zvi Mowshowitz https://thezvi.substack.com
Probably the single best high-volume commentator.
Covers
- capabilities
- safety
- governance
- industry gossip
- papers
- politics
His weekly summaries save enormous amounts of time.
Scott Alexander https://www.astralcodexten.com
Rarely writes directly about AI, but when he does it is thoughtful and unusually balanced.
Dwarkesh Patel https://www.dwarkesh.com
Not a safety advocate per se.
But probably the best long-form interviewer of frontier researchers.
Tyler Cowen https://marginalrevolution.com
Excellent for
- economics
- incentives
- technology
Often offers a useful counterweight to existential-risk narratives.
Ethan Mollick https://www.oneusefulthing.org
Very good on
- deployment
- education
- enterprise adoption
A reminder that most AI today is still tool-like.
Ben Thompson (Stratechery) https://stratechery.com
Excellent business analysis.
Kelsey Piper https://substack.com/@kelseytuoc
One of the best journalists covering AI safety.
Simon Willison https://simonwillison.net
Exceptional on open-source models, tooling, and practical developments.
5. Rationalist / Alignment Community
Essential reading—even if one disagrees with parts of it.
LessWrong https://www.lesswrong.com
Still the intellectual center of technical alignment discussion.
Many ideas later adopted by labs appeared here years earlier.
Alignment Newsletter (Rohin Shah) https://rohinshah.com/alignment-newsletter/
Probably the single best technical newsletter.
Weekly summaries of new papers.
AISafety.com
An excellent curated portal that itself recommends blogs, newsletters, podcasts, forums, and learning resources, making it a useful “directory of directories.”
GreaterWrong https://www.greaterwrong.com
Alternative interface for LessWrong.
6. Academic Sources
arXiv https://arxiv.org
Must-read categories:
- cs.AI https://arxiv.org/list/cs.AI/recent
- cs.LG https://arxiv.org/list/cs.LG/recent
- cs.CL https://arxiv.org/list/cs.CL/recent
- cs.RO https://arxiv.org/list/cs.RO/recent
Search terms:
- alignment
- interpretability
- agents
- autonomy
- scheming
- model organisms
- control
- evaluation
Google Scholar alerts
https://scholar.google.com/scholar_alerts?view_op=list_alerts&hl=en
Useful for
- AI safety
- AGI
- superintelligence
- autonomous agents
SSRN
Good for governance and legal work.
7. Governance
International AI Safety Report
Required reading each edition.
Likely the closest thing the field has to a consensus technical assessment.
OECD AI Observatory
EU AI Office
UK Government AI Security Institute
NIST AI RMF
8. Newsletters
Highly recommended.
- Alignment Newsletter https://rohinshah.com/alignment-newsletter/
- Zvi’s Substack https://thezvi.substack.com
- Understanding AI (subscription required) https://www.understandingai.org
- Import AI (Jack Clark) https://jack-clark.net
- AI Safety Newsletter https://newsletter.safe.ai
- Last Week in AI https://lastweekin.ai
- AI Snake Oil https://www.normaltech.ai
- Ben’s Bites (more capabilities-oriented) https://www.bensbites.com
- AI Futures https://blog.aifutures.org
Community recommendations consistently place the Alignment Newsletter and AI Safety Newsletter among the best ways to keep up with new safety work.
9. Podcasts
- Dwarkesh Podcast
- Cognitive Revolution
- Hard Fork
- AI Risk Network
- AXRP (AI X-Risk Research Podcast)
- Future of Life Institute Podcast
- Decoder
10. Deliberate Counterpoints
A good monitoring system should include voices that challenge the prevailing assumptions.
Include:
- Gary Marcus
- Arvind Narayanan (especially on “AI Snake Oil”)
- Ethan Mollick
- Tyler Cowen
- Simon Willison
These writers often push back against overgeneralization while still taking AI seriously.
If You Only Read Ten Things
If he were willing to check only ten sources regularly, I would suggest:
- Reuters AI
- MIT Technology Review (AI)
- Anthropic Research
- UK AI Security Institute
- Zvi Mowshowitz
- Gary Marcus
- LessWrong
- Alignment Newsletter
- International AI Safety Report
- arXiv (alignment / interpretability / agents)
That combination provides a healthy mix of reporting, primary technical work, policy, commentary, skepticism, and frontier thinking.
One addition I would make for your project
Because From Tool to Actor is concerned with the transition from bounded tools to autonomous institutional actors, I would encourage you to organize your reading around five complementary lenses rather than by publication alone:
- Capabilities — What can frontier models now do?
- Control — Are we still able to understand, evaluate, and constrain them?
- Governance — How are governments and institutions responding?
- Deployment — Where are increasingly agentic systems being integrated into real workflows?
- Interpretation — How are different communities (skeptics, safety researchers, economists, industry, and journalists) making sense of these developments?