Skip to content
RSS
Amplifier
Blogs
Podcasts
Music
Videos
Topics
Submit
Discover
Search
OPML
llms.txt
About
Sign up
Favorites
Account
Blog
The Pond
Last 10 notes on The Pond
turntrout.com ↗
RSS feed ↗
10 posts
Follow
ai
gaming
pond
2025-era reward hacking
ais use killer
alignment mentorship turntrout
apply alignment mentorship
convergence ai psychology
cooperativeness scalable mitigation
eval cooperativeness scalable
framework government ai
gaming modifying specification
Latest posts
Misaligned AIs Could Use Killer Robots to Take Over
Aug 11, 2026
Why I Left Google DeepMind
Jul 15, 2026
A Red Line and Oversight Framework for Government AI Contracts
Jul 15, 2026
How Robust Are Natural Language Autoencoders to Initialization?
Jul 9, 2026
Eval Cooperativeness May Be a Scalable Mitigation for Eval Gaming
May 24, 2026
Prettify Your Text In-Browser
Feb 14, 2026
No Instrumental Convergence without AI Psychology
Jan 20, 2026
Recontextualization Mitigates Specification Gaming Without Modifying the Specification
Dec 23, 2025
Apply for Alignment Mentorship From TurnTrout and Alex Cloud
Dec 23, 2025
2025-Era “Reward Hacking” Does Not Show that Reward Is the Optimization Target
Dec 18, 2025
←
Prev
✦
Random
Next
→
Visit
↗
Feed
Kagi
↗