LogoTensorsOfTheWall
Overfitted Opinions: I write things down here instead of explaining them at 2 a.m. to someone trying to sleep
vLLM: LLM Inference as a memory and scheduling problem

vLLM: LLM Inference as a memory and scheduling problem

Jul 30, 2026

Part 1 of a practical deep dive into vLLM: prefill vs decode, the KV-cache bottleneck, PagedAttention, and how modern attention backends consume paged KV.

16 min read
A Brief History of Computer Vision: Before Pixels Had Brains

A Brief History of Computer Vision: Before Pixels Had Brains

Jul 11, 2025

Part 1 of N tracing the journey of computer vision.

20 min read
Why the Hell Am I Writing a Blog?

Why the Hell Am I Writing a Blog?

Jun 21, 2025

A quick intro to why I started this blog

3 min read
Hello world!

Hello world!

Jun 20, 2025

Notion Test Blog

2 min read

This site was handcrafted with care (and mild frustration). Bugs are features, right?
© 2026 Sandesh Bharadwaj. All rights reserved.