AI Safety Who's Who

AI Safety

Bill asked me to compile a list of sources that one could refer to when keeping up with this rapidly expanding field. I’ve written a long blog post about Keeping Up (one of my favourite activities), but that’s more about what it is like to keep up than any helpful lists. So here we go…

First of all, we have to consider what we are keeping up with. In this case it is AI Safety. AI Safety wasn’t a well-defined field (probably still isn’t, really) but it has become much more coherent in the last little while. We won’t just confine ourselves to actual AI Safety, however, since you still need to have the context of “what is happening in AI/frontier labs” and you need to consider the AI Industry and financial implications, you need to consider governance and regulation and you should also be aware of those who are skeptical of the whole enterprise. It’s not really concentric circles, more like Venn diagrams or something.

The other issue to keep in mind is that many of the sources that one might follow are themselves aggregators. So you are going to see the same thing multiple times. From my point of view that’s just the cost of doing business, and its also a useful bell-weather for seriousness. In other words, if you read the same story everywhere (e.g., The Incidents™) then you might know that it is something big.

I already do a lot of this, and have my own lists, although I have never shared them withanyone else, nor provided commentary. Back in the early days of the web and the first blogs, it was common for people to create a “blog roll” that sat alongside their blog posts. It provided links to what they were reading and - to be honest - it was a simple tool for yourself as a blogger, since you spent a lot of time looking at your own page, tinkering with it and reloading it, so why not have your favourites right there on the same page. It also was a kind of early system of “likes” in that you were showing who you thought worthy of sitting alongside your content.

Maybe the first place to start is at the edges. There are some people who (still) have not taken the red pill (see my post about AI pills) and we can’t forget about them, since they still shape discussion in some ways.

Skeptics

Cory Doctorow (also a columnist for The Guardian): https://craphound.com and https://pluralistic.net

Ed Zitron https://www.wheresyoured.at (has premium option)

Melanie Mitchell https://melaniemitchell.me (scholarly approach)

Gary Marcus https://garymarcus.substack.com

Institutions

Redwood https://blog.redwoodresearch.org

FLI - https://futureoflife.org

Apollo Research https://www.apolloresearch.ai/blog

UK AISI https://www.aisi.gov.uk

Canada CAISI https://ised-isde.canada.ca/site/ised/en/canadian-artificial-intelligence-safety-institute

US CAISI https://www.nist.gov/caisi see also https://www.nist.gov/artificial-intelligence

Commentators

Zvi (brings together many of the Twitter/X AI Safety community, so you don’t have to follow them individually)

Newspapers/Magazines

Blogs/SubStacks

Thoughtleaders/scientists/CEOs

Gemini’s recommendations

Here is a curated list of top-tier AI safety resources, broken down by category so Bill can pick the depth and perspective that fits him best:

1. Dedicated AI Safety Newsletters & Digests

2. Community Hubs & Discussion Forums

3. Governance, Policy & Incident Trackers

4. Frontier Labs’ Safety Blogs

To track policies and safety releases directly from the frontier builders:

Gemini’s lisk of think tanks

Here is a curated overview of prominent think tanks, research institutes, and policy centers dedicated to AI safety, governance, and frontier risk, categorized by their primary focus.

1. Major Policy & Security Think Tanks

These institutions mirror the RAND model, focusing heavily on geopolitics, national security, regulatory strategy, and supply-chain governance.

2. Dedicated AI Governance & Catastrophic Risk Institutes

These organizations specialize specifically in frontier model safety, existential/catastrophic risk, and long-term regulatory mechanisms.

3. Evaluation, Technical Standards & Societal Impact

Organizations bridging technical audits, red-teaming benchmarks, and broader societal impacts.

Quick Comparison

Organization Primary Geography Key Angle
CSET US Compute, hardware supply chain, talent & geopolitics
GovAI US Frontier AI governance, institutional design, treaties
CAIS US Catastrophic risk prevention & compute monitoring
Ada Lovelace UK / EU European regulatory policy, accountability & standards
METR US / International Dangerous capability evaluations & frontier threat modeling