
Predictions Over What?
Accuracy is only half the battle
Posts about AI alignment. Victory or death.
Subscribe:.rss.atom.json.md.m3u.pls
Live Last read · last published · next check

Accuracy is only half the battle

What happens if evaluations, interpretability, and monitoring succeed?

The AI Alignment ecosystem working as intended?

I am not a crackpot

A Case for Marginal Progress on Superalignment Theory

How we can get AI to set up a good future without caring about it

Economists believe in an intelligence explosion, even if they don’t realize it yet

It's time to talk about what I've been working on

That's not strictly true, but myopia is still incredibly useful

Ensuring corrigibility in an AI and ensuring corrigibility in the agents it creates are two completely distinct problems.