RSSAmplifier

Blog

AI Research

AI Research

blog.n.ichol.aiRSS feed ↗3 posts

Latest posts

Packaging Latent Reasoning as a Real Model

DeepSeek-V4-Flash-0731-Latent-Reasoning. A self-contained model that does thinking in latent space, NVFP4-quantized, with a production vllm form for serving runtime. https://huggingface.co/nmitchko/De

The Doctor is NOT the Mother

Adding a CoLaR Head to DeepSeek Flash v4 for Latent Reasoning with Learned Stop Criteria Published on blog.n.ichol.ai The Riddle A father and his son are driving down the road. They crash. Only the

LLM Activation Engineering: An Easy Foray

This is a recap of an old project from May 2024. Credit to Neel Nanda for llm lens and Mihaiiii for llm steer python modules. I've been playing around with steering LLM outputs by manipulating their