❌

Normal view

Disaggregation Is a Thousand-GPU Problem

4 September 2026 at 14:00

Three conditions that must hold before splitting prefill from decode pays off, and why chunked prefill is the right default below that threshold.

The post Disaggregation Is a Thousand-GPU Problem appeared first on Towards Data Science.

Avoiding Entity Key Drift in a Data Lake: Step 2, When Fuzzy Matching Stops Working

2 September 2026 at 15:30

I built a matcher meant to finish the cleanup that normalization left behind. Testing it against real data showed that no version of it could be made safe. What follows is the architecture that was left once the matcher was set aside.

The post Avoiding Entity Key Drift in a Data Lake: Step 2, When Fuzzy Matching Stops Working appeared first on Towards Data Science.

A RAG That Says β€œNot in This Document” Has to Show Four Kinds of Evidence

2 September 2026 at 14:00

Enterprise Document Intelligence [Vol.1 #B3] - A confident wrong answer is a bug. A bare β€œno answer” with no justification is almost as bad. Each of the four bricks has one piece of evidence to show

The post A RAG That Says β€œNot in This Document” Has to Show Four Kinds of Evidence appeared first on Towards Data Science.

Received β€” 31 August 2026 ⏭ Towards Data Science

FAQ as RAG: When You Get to Design the Corpus

31 August 2026 at 14:00

Enterprise Document Intelligence [Vol.1 #B2] - The FAQ inverts every brick of the standard RAG pipeline. Parsing is trivial, retrieval doubles as a cache, and few-shot prompting becomes a retrieval problem too

The post FAQ as RAG: When You Get to Design the Corpus appeared first on Towards Data Science.

❌