RSS Amplifier

Blog

Guive’s Substack

My personal Substack

guive.substack.comSource feed ↗8 posts

Live Last read · last published · next check

Written by

Latest posts

AXRP appearance

Recently, I appeared on Daniel Filan’s AXRP podcast to discuss my article arguing for AI property rights. I’m happy with how the podcast turned out; I honestly think it’s a bit more interesting than the article. You can listen to the episode here:

The case for AI property rights

Many people worry that AIs will rise up against humans, steal our property, and kill us.

GPT-oss Is an Extremely Stupid Model

I recently tried to reproduce the results from the Anthropic "Agentic Misalignment" report with GPT-oss.

Alignment Fine-tuning is Character Writing

Why does Claude love Caffè Strada and sometimes claim to have a Japanese wife?

Token and Taboo

Can language models correctly identify moral taboos?

Testing for Scheming with Model Deletion

There is a simple behavioral test that would provide significant evidence about whether AIs with a given rough set of characteristics develop subversive goals.

Updating on Bad Arguments

Here is an intuitively compelling principle: hearing a bad argument for a view shouldn’t change your degree of belief in the view.

Coming soon

This is Guive’s Substack.