pocoo · Jul 25, 2026
We trained a ternary LLM from scratch and asked it its own name. It doesn't have one.
0Sign in to vote or save
This site does not allow itself to be embedded. You can still read it on the original site — the toolbar below keeps your place in the directory.
RivaQuant: a 162M-param BitNet b1.58 ternary transformer, trained from scratch on TinyStories on a rented RTX 3090. Two real bugs, both caught by an automated cost-safety watcher mid-training and fixed live. Then we asked it to name itself, and it turns out that question doesn't have an answer for a model like this — here's why, and what actually came out.
Comments
Nothing yet. Say the first thing.
Sign in to join the conversation.