Does a context gate for search agents actually work?
Testing context gating on BrowseComp-Plus. A cheap label-free gate cuts a search agent's input tokens 1.4x, gate costs included, no detectable accuracy change.
Ruminations of a data scientist turned engineer. I like talking about LLMs, deep learning, engineering and data engineering.
Testing context gating on BrowseComp-Plus. A cheap label-free gate cuts a search agent's input tokens 1.4x, gate costs included, no detectable accuracy change.
Benchmarking H100 PCIe vs SXM vs NVL on training cost, step times, and NCCL profiling to find the cheapest GPU configuration for Nanochat
Enabling endless capabilities for LLMs
Validate the data as it streams, don't make your users wait.