Posts
17 Aug 2026
Improving Throughput by Optimising KV Cache Efficiency for Agentic Workloads
Boosting throughput by leveraging characteristics of agentic workloads to increase efficiency of the KV cache.
29 Jun 2026
Fast vLLM: cutting cold starts by 20x
Reducing container start-up time from 117s to 6s, the long way round.
13 Jan 2026
A Deep Dive into Diffusion Models: DDPM
An introduction to diffusion models with a focus on denoising diffusion probabilistic models (DDPM).
3 Jan 2026
Exploring the theory behind Variational Autoencoders (VAEs)
How do VAEs work?