RSS Amplifier

Blog

Daily Dose of Data Science

A free newsletter for continuous learning about data science and ML, lesser-known techniques, and how to apply them in 2 minutes. We keep things no-fluff. Join 100,000+ data scientists from top companies like Google, NVIDIA, Microsoft, Uber, etc.

blog.dailydoseofds.comSource feed ↗22 posts

Live Last read · last published · next check

Written by

Latest posts

Saves to your Listen queue, to pick up on another day or another device.

Kimi K3's Sandbox Problem Finally Has an Open-Source Fix

...explained with code.

Grok Bot Masterclass

Everything you need to understand, set up, and get real work out of Grok Bot.

How a GPU Actually Works

The intuition an LLM engineer needs. Understand techniques like quantization, speculative decoding, and continuous batching in one place.

A Cheaper Model Does Not Imply a Cheaper Turn

The practical implications of model routing, clearly explained.

How Production LLMs Reason Better At Inference Time

8 techniques, explained visually.

Continuous Batching in LLMs

The technique behind vLLM's 23x throughput jump and the default scheduler in every serving engine.

[Hands-on] Audio RAG with 200x Cheaper Vector DB Costs

...while also outperforming OpenAI and Cohere.

Karpathy's Full Agentic Engineering Lifecycle using Google's Agents-CLI

...explained as step-by-step guide.

How to Query Billion+ Rows on Postgres Without the Overhead

...explained as a full setup guide.

A 10-week Roadmap to Run LLMs in Production

...covered with hands-on resources.

[Hands-on] Build Semantic Search Inside Your Database Without an Embedding Pipeline

...explained with code.

The Missing Piece of Agent Self-Improvement

...explained step-by-step with code.

[Hands-on] How to Serve 5 Models On One GPU

How small specialized models are changing inference infrastructure, and why serving them efficiently takes more than standard serving frameworks.

Why Your Agent Remembers Everything and Understands Nothing

Building a pattern recognition layer for memory in production.

The Hands-on AI Engineer Playbook to Build RAG Apps for Production

Why RAG latency is a prefill problem, not a retrieval problem.

Build a Stock Market Research Agentic Workflow​

...using a no-code drag-and-drop builder.

Play

6 Automatic Optimization Methods for LLM Systems

...explained visually and with practical tradeoffs.

Why 86% of Claude Code Bill Has Nothing to Do With Your Prompts

...and a solution to fix Claude bills.

6 LLM Deployment Formats in Production

...explained visually.

Serverless vs. On-prem vs. Edge Deployment

...explained visually.

Graph Engineering Clearly Explained!

Covered with best practices from the industry.

11 LLM Evaluation Methods

Must-know for AI engineers (explained visually).