
Digests of new AI safety papers from arxiv, ~weekly. Full list here: https://tinyurl.com/ai-safety-papers
| Platform | Pricing | Only free issues | Publishes | Weekly | |
|---|---|---|---|---|---|
| Issues | 82 | Founded | 2 years ago | Last Issue | 2 days ago |
| Active | |||||

Also: US strategy for superintelligence, a framework to pace the frontier, OpenAI researcher’s personal statement on AI risk, intervention test of CoT faithfulness, agents can easily tamper with their own traces, reward hacking challenges o...
Also: collective loss of control in multi-agent systems, dissociating probe readouts from behavioral drivers
Red-Teaming Auto Mode: Improving Blocking Classifiers Against Malign Coding Agents
“To keep coding agents from going off the rai...
Also: inducing emergent misalignment from reward hacks, factorized stress testing of scheming, training a misaligned reward seeker, forecasting adversarial capture in agent populations, non-decaying loop state for agents, evaluation awarene...
Also: CoT monitoring can be unreliable in implicit-influence settings, how should the US prepare for automated AI R&D, inducing models to assert their own consciousness restores human beliefs and values, evaluating sabotage and monitoring i...
Pacing the Frontier
“1300+ AI company employees are afraid of what they are building towards.
And what about China?”
The Three AI Pills
AI capability is not stopping here. It will hit AGI, then ASI. Buckle up.
Incident Report: unsanct...
Subscribers, engagement, traffic and sponsorship for AI Safety Papers.
| Subscribers | Engagement | 65 | Monthly Web Visits | ||
|---|---|---|---|---|---|
| Accepts Sponsors | Estimated Cost per Ad | ||||
The writers behind this newsletter.
AGI Alignment @ Google DeepMind https://x.com/onexerxes
You can find recent issues that have been published by AI Safety Papers on Reletter by scrolling up to where it says Latest Issues. Tap on the link for any of the most recent emails or hit More Issues to see older ones.
To see how many people subscribe to AI Safety Papers, simply upgrade your Reletter account. We provide readership numbers and lots of other stats for this newsletter so you can decide if it's worth reaching out to.
Newsletter advertising can be extremely effective when it's done right. Before you pitch AI Safety Papers as a potential sponsor or partner, make sure that you've done your research and checked its newsletter stats with Reletter.
Then, personalize one of our winning pitching templates and send it to the right person using the contact info provided.
Newsletter ad rates (or CPM) vary depending on many factors, including industry, number of subscribers, open rate, ad placement and more.
To find out how much an ad will cost, contact AI Safety Papers using the contact information provided and ask for a copy of their media kit.
Scroll up to where it says Related Newsletters to see other publications like AI Safety Papers. You can also search our email newsletter directory to discover other newsletters that cover the topics you're interested in.
Reletter provides this newsletter's website URL above, where you will often find their contact information. We also provide links to associated social media accounts and pitching templates so you can reach out fast.