From Silicon to Response — a sysadmin's guide to the complete AI stack
The complete map. Six layers from hardware to application — understand the whole before diving into the parts.
Read ›GPU clusters, TPUs, and why AI needs specialized compute. Think data center — except every rack is doing matrix math.
Read ›How a model learns. Supervised, unsupervised, reinforcement learning — and what RLHF means for the AI you actually use.
Read ›Inside the model. Transformers, tokenization, neural network structure — the machinery that turns text into prediction.
Read ›What actually happens when you hit Enter. The inference pipeline from prompt receipt to token generation to response delivery.
Read ›API, web UI, prompt engineering. How the outside world talks to the model — and how to talk back effectively.
Read ›From chatbots to code assistants to autonomous agents — the application layer and where AI is headed next.
Read ›These TMA pages go deeper on the topics covered in this series: