Daily Dose of Data Science
Subscribe
Sign in
Home
Sponsor
Premium
Archive
Leaderboard
About
Latest
Top
Discussions
Continuous Batching in LLMs
The technique behind vLLM's 23x throughput jump and the default scheduler in every serving engine.
10 hrs ago
•
Avi Chawla
12
[Hands-on] Audio RAG with 200x Cheaper Vector DB Costs
...while also outperforming OpenAI and Cohere.
Aug 12
•
Avi Chawla
12
Karpathy's Full Agentic Engineering Lifecycle using Google's Agents-CLI
...explained as step-by-step guide.
Aug 11
•
Avi Chawla
13
How to Query Billion+ Rows on Postgres Without the Overhead
...explained as a full setup guide.
Aug 10
•
Avi Chawla
9
A 10-week Roadmap to Run LLMs in Production
...covered with hands-on resources.
Aug 9
•
Avi Chawla
16
1
[Hands-on] Build Semantic Search Inside Your Database Without an Embedding Pipeline
...explained with code.
Aug 7
•
Avi Chawla
8
1
The Missing Piece of Agent Self-Improvement
...explained step-by-step with code.
Aug 6
•
Avi Chawla
15
1
[Hands-on] How to Serve 5 Models On One GPU
How small specialized models are changing inference infrastructure, and why serving them efficiently takes more than standard serving frameworks.
Aug 5
•
Avi Chawla
11
Why Your Agent Remembers Everything and Understands Nothing
Building a pattern recognition layer for memory in production.
Aug 4
•
Avi Chawla
9
1
The Hands-on AI Engineer Playbook to Build RAG Apps for Production
Why RAG latency is a prefill problem, not a retrieval problem.
Aug 3
•
Avi Chawla
9
2
Build a Stock Market Research Agentic Workflow
...using a no-code drag-and-drop builder.
Aug 1
•
Avi Chawla
6
1
8:21
July 2026
6 Automatic Optimization Methods for LLM Systems
...explained visually and with practical tradeoffs.
Jul 31
•
Avi Chawla
9
1
1
This site requires JavaScript to run correctly. Please
turn on JavaScript
or unblock scripts