RSSAmplifier

Blog

KShivendu

Kumar Shivendu's blog

kshivendu.devRSS feed ↗15 posts

Latest posts

Inference-Free SPLADE: Full Quality, 13× Faster Queries

Inference-free SPLADE nearly matches full SPLADE quality at 13× lower query latency, no GPU needed at query time. Benchmarked against full SPLADE and BM25 on Qdrant and pyserini.

Token-Native Storage: Read and Write in Your Agent's Language

Agents read and write in tokens (BPE). But our DB engines are still designed for humans (UTF-8). Persist payloads as BPE token IDs with a static entropy coder and you get ~3.3× lossless compression, zero-cost persistence of LLM output, and a representation agents already speak.

Why I'm Obsessed With Search

Search is everywhere and is one of the hardest problems in CS. Here's why I'm obsessed with it.

2025: Year in Review

Reflections on a year of growth, travel, and pursuing mastery

I Reverse-Engineered Exa.ai Infrastructure Cost with Napkin Math

JSON Embeddings & Reranking

Exploring the JSON embeddings to for matching new products with existing ones

Monitoring my Life with Grafana, Prometheus, and InfluxDB

My Linux setup

Stress testing your machine

How to partially download a file

Exploring Data Structures in Redis

In-depth look at ACID properties of Postgres

Unit testing best practices in Python

Maximising productivity as a developer

FAQs related to GSoC