❌

Reading view

Estimating from No Data: Deriving a Continuous Score from Categories

A walkthrough of and the maths behind using low-capacity networks to acquire fine-grained scoring when only categorical labelling is available for training

The post Estimating from No Data: Deriving a Continuous Score from Categories appeared first on Towards Data Science.

  •  

Speculative Decoding on CPUs: Nearly 4x Faster Token Generation with DFlash

Speculative decoding can turn underused CPU compute into faster token generation, without changing the model's output. In our vLLM tests, DFlash delivered 3.92x the autoregressive throughput with Qwen3.5-9B on Intel Xeon 6 at concurrency 1. We break down where the speedup comes from, explain the acceptance metrics, and show what determines whether speculation pays off.

The post Speculative Decoding on CPUs: Nearly 4x Faster Token Generation with DFlash appeared first on Towards Data Science.

  •  

Estimating from No Data: Deriving a Continuous Score from Categories

A walkthrough of and the maths behind using low-capacity networks to acquire fine-grained scoring when only categorical labelling is available for training

The post Estimating from No Data: Deriving a Continuous Score from Categories appeared first on Towards Data Science.

  •  

How to Scale an Integration Pipeline Without Breaking Correctness

A production account of scaling an enterprise integration pipeline from 500 to 8,000 events per second, and the two correctness guarantees the throughput work was never allowed to trade away.

The post How to Scale an Integration Pipeline Without Breaking Correctness appeared first on Towards Data Science.

  •  

Building Enterprise Agent Systems that People can Trust, Verify andΒ Improve

5 principles that determine whether an agent system succeeds in production, explained through one I built for a $100M+Β company.

The post Building Enterprise Agent Systems that People can Trust, Verify andΒ Improve appeared first on Towards Data Science.

  •  

Can a Local LLM Run My AI Assistant?

I replayed the same 27 real production tasks through two local models, one hardware upgrade apart, to find out what it actually takes to replace Claude as the brain behind a 90-tool personal agent.

The post Can a Local LLM Run My AI Assistant? appeared first on Towards Data Science.

  •  
❌