Daily Dose of Data Science

Daily Dose of Data Science

KV vs Prefix vs Prompt vs Semantic Caching
...explained with best practices in production.
11 hrs ago • Avi Chawla
[Hands-on] Turn Scientific Figures Into Structured Data with Mistral OCR
A full walkthrough of the extraction schema, with code.
Aug 26 • Avi Chawla
Build a Multi-Agent GTM Intelligence System
...explained with code.
Aug 25 • Avi Chawla
Preloading Knowledge Into a Model Instead of Retrieving It
How to process your corpus once, skip retrieval entirely, and serve every query from a stored cache. Three parts covering the full spectrum.
Aug 24 • Avi Chawla
How Semantic Code Navigation Cuts Agent Token Costs by up to 36%
Understanding what an agent actually does with the tokens before it writes code.
Aug 21 • Avi Chawla
What is (was?) GIL in Python?
...explained with code.
Aug 20 • Avi Chawla
Kimi K3's Sandbox Problem Finally Has an Open-Source Fix
...explained with code.
Aug 19 • Avi Chawla
Grok Bot Masterclass
Everything you need to understand, set up, and get real work out of Grok Bot.
Aug 18 • Avi Chawla
How a GPU Actually Works
The intuition an LLM engineer needs. Understand techniques like quantization, speculative decoding, and continuous batching in one place.
Aug 17 • Avi Chawla
A Cheaper Model Does Not Imply a Cheaper Turn
The practical implications of model routing, clearly explained.
Aug 16 • Avi Chawla
How Production LLMs Reason Better At Inference Time
8 techniques, explained visually.
Aug 14 • Avi Chawla
Continuous Batching in LLMs
The technique behind vLLM's 23x throughput jump and the default scheduler in every serving engine.
Aug 13 • Avi Chawla
© 2026 Avi Chawla · Privacy ∙ Terms ∙ Collection notice
Start your SubstackGet the app
Substack is the home for great culture