This website uses cookies
Read our Privacy policy and Terms of use for more information.
Aug 10, 2026
•
9 min read
The Physics of Fan-In: From Silicon, to Kernel, to Global Routing
Aug 9, 2026
4 min read
What we learn from the world's largest AI Model Repo Hugging Face breach to VulnHunter: why every AI capability in your security stack needs a self-hosted fallback.
Aug 3, 2026
How RDMA, DPUs, and Kernel Bypass Get the CPU Out of the Data Path
LLM Inference
+3
Jul 27, 2026
8 min read
Weekly field notes on the silicon-level constraints of modern AI architecture.
LLM Infrastructure
+1
Jul 19, 2026
7 min read
How vLLM solves the real bottleneck in LLM serving - and why it isn't compute.
Edge AI
+4
Jul 13, 2026
Why the DeepSeek/GLM cost gap is now an architecture decision, plus this week's cloud, GenAI, IoT, and edge AI/robotics signal.
Jul 8, 2026
What AWS's Lambda-at-scale postmortem teaches about quota governance, plus this week's distributed systems, IoT, edge AI, physical AI and quantum signal.