Hey friends! 👋
Last week we dropped the article that kicked everything off: a production-grade, event-driven OCR pipeline built for Kubernetes. Decoupled model serving, autonomous workers, scale-to-zero cost efficiency, a Rust ingestion gateway — the whole blueprint.
And judging by the comments, the DMs, and the "wait, how does this part actually work?" messages piling up in our channels, you've got questions. Lots of them.
So before we dive into the week-by-week deep dives, we're doing something first.
We're going live. 😎
What this session is
This isn't a technical episode. We're not spinning up clusters or debugging YAML just yet.
Think of this as the orientation — the session where we pull back the curtain on the whole series, explain the thinking behind the architecture, and answer whatever's on your mind before the real engineering begins.
We'll walk through the article together, connect the dots between the pieces, and make sure everyone starts the journey on the same page — whether you've been shipping Kubernetes workloads for years or you've never touched a node pool in your life.
What to expect
Here's what we'll cover:
The big picture. Why document intelligence is such a massive, unsolved headache — and why "just run OCR on it" stopped being good enough. We'll frame the problem the whole series is built to solve.
A tour of the blueprint. We'll walk through the architecture at a high level: the autonomous worker model, the decoupled inference engine, the Rust gateway, and how scale-to-zero keeps the cloud bill sane. No slides full of buzzwords — just a clear map of what we’re building.
Why we made these choices. Why VLMs break naive serving setups. Why we split lightweight layout extraction from heavy text generation across different GPU tiers. Why Rust for ingestion. The trade-offs behind the design.
How the series works. What lands each week, how the codebase is structured, and how to get the most out of the Sunday office hours going forward.
Open Q&A. The main event. Bring your questions about the architecture, the tooling, the prerequisites, or anything from the article that left you curious. We'll answer live.
Who this is for
If you read the article and thought "this is exactly the kind of infrastructure I want to build, but I need to understand the map before I start walking" — this session is for you.
Come with questions. Come with skepticism. Come with your own war stories about OCR pipelines that fell over in production. That's what makes these sessions worth showing up for.
The details
📅 When: Friday, July 17th
📍 Where: Live on Substack, join here
🎟️ Access: This first session is free for all subscribers, our way of welcoming you into the series. But heads up: this one is the exception. Every live office hour after this will be for Premium Subscribers only.
So if you've been on the fence, this is the session to show up for. And if you want to follow the whole journey, now's the time to go premium!
The AI Bros are back, folks … and this time, they've brought GPUs.
See you there!





Yes!! You'll receive today! (the recordings will be available every Monday from now on 🙌)
can we please get the recording for the same.