Daily Dose of Data Science
A free newsletter for continuous learning about data science and ML, lesser-known techniques, and how to apply them in 2 minutes. We keep things no-fluff. Join 100,000+ data scientists from top companies like Google, NVIDIA, Microsoft, Uber, etc.
- Indexed issues, last 90 days
- 28
- Latest publication
- Sep 30, 2026
- Audience
- Checking…
- Earliest in this view
- Aug 27, 2026
Latest issues
Jev for RAG, clearly explained! (opens the original)
Read excerpt
In today’s newsletter:Agents without memory aren’t agents at all.Jev for RAG, clearly explained!Top 12 agentic use cases for Jev.Agents without memory aren’t agents at allAn LLM can appear to remember because the application keeps sending previous messages back with each request.The model itself is still stateless.Start a new session without stored history, and every preference, decision, and previous outcome disappears.<a href="http
Build a real-time hotel booking voice agent (opens the original)
Read excerpt
In today’s newsletter:vLLM closed this feature request as “not planned”[Hands-on] Build a real-time hotel booking voice agent.Where does all the VRAM go during LLM inference?vLLM closed this feature request as “not planned”:Devs have requested running several small models on one GPU through a single inference server since 2023.Workloads require serving two to five small models on the same GPU because none of them generates enough traffic to justify a
System 1 vs. System 2 Agent Harnesses, clearly explained (opens the original)
Read excerpt
In today’s newsletter:Your agent hit a wall. Beacon handoff lets another one pick up where it left off.System 1 vs. System 2 Agent Harnesses.[Interview question] How MoE routing works across GPUs?Your agent hit a wall. Beacon handoff lets another one pick up where it left off.Coding agents are getting much better at solving complex engineering tasks. But they still have a basic problem: they’re stuck in their own tool.You mi
Contrastive Language Model, clearly explained (opens the original)
Read excerpt
Your best model shouldn’t be answering every requestMost production LLM traffic isn’t equally demanding.A greeting, a one-line lookup, and a codebase refactor can all hit the same endpoint. If your best model handles all three, you’re paying more than you need to for the first two.TrueFoundry’s Auto Routing moves that decision into the AI gateway.<div clas
Build a Jev Judge (opens the original)
Read excerpt
Another brilliant course by Andrew Ng!<img alt="" class="sizing-normal" height="456" src="https://substackcdn.com/image/fetch/$s_!p-e3!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubst
Publishing over time
Last 90 days. Choose a month to open its work.
Recurring subjects
Named in the text we hold. One piece can cover several.
Audience
No verified audience measurement yet.
About this data
Counts cover the work we have indexed. Tone needs enough text and a confident classification. Excerpts and episode notes are not full articles or transcripts.
Identity or attribution wrong? Suggest a correction.