chrislabs.ai
The world of AI & infrastructure, explained for builders and adopters.
Fundamental but genuinely technical knowledge for the next generation of sales, sales engineers, and ML, data, and platform engineers - from tokens and attention to GPUs, serving, training, and the VAST Data platform.
New here? Start here.
A short path across the three core pillars - how models read text, why inference is memory-bound, and the hardware that runs it.
Foundations
How language models read, represent, and generate text.
Memory & Efficiency
Why inference is memory-bound, and how to tame it.
Systems & Infrastructure
The hardware and software stack that runs AI at scale.
NVIDIA Infrastructure
FundamentalWorkloads, AI factories, the CUDA moat - and where to go deeper.
GPUs & Racks
FundamentalThe GPU lineup and roadmap, spec sheets, NVL72 racks, and the power wall.
Networking & the Data Path
IntermediateScale up vs out, NVLink, Spectrum-X, DPUs, and GPUDirect Storage.
Inference & Serving
IntermediatePrefill vs decode, batching, and disaggregated serving.
Training & Fine-Tuning
IntermediatePretraining, LoRA/QLoRA, checkpoints, and resourcing.
NVIDIA STX & Context Memory
AdvancedBlueField-4 STX, the CMX KV-cache tier, and VAST's role.
Simulation & Digital Twins
AdvancedOmniverse, world models, and simulation as a GPU workload.
Applications
Patterns that turn models into products.
Economics & Deployment
What a token costs, and how real deployments run on today's GPUs.
Tokenomics
How a Token's Cost Is Calculated
FundamentalFrom GPU dollars per hour to dollars per million tokens - the one equation behind it all.
Why Output Costs More: Batching & Speed
IntermediateDecode is bandwidth-bound: batching, context length and the speed-vs-cost tradeoff.
What an Hour of GPU Really Costs
IntermediateCapex, depreciation, power and utilization - building the $/GPU-hour for B300, GB300 and more.
VAST Data
The data platform for the AI era.
VAST Overview
FundamentalWhy VAST matters for AI, data lakes, and agents.
VAST Architecture
IntermediateDASE: disaggregated, shared-everything architecture.
Secure Multitenancy
IntermediateTenant isolation, access control, and QoS on one shared platform.
VAST DataSpace
IntermediateOne global namespace across edge, on-prem, and every cloud.
VAST SyncEngine
FundamentalFind data across file shares, object stores and SaaS apps - and bring it onto VAST.
VAST DataBase
IntermediateA columnar table format - an Iceberg/Delta alternative, native to the platform.
DataEngine & Event Broker
IntermediateServerless functions, triggers, and Kafka-compatible streaming.
VAST VectorStore
AdvancedVector search at scale, native to the data platform.
InsightEngine (NVIDIA)
IntermediateReal-time RAG and the NVIDIA AI data-platform stack.
Context Memory (KV Cache)
IntermediateReload KV instead of recomputing it - and what a cache hit is worth in $/token.
AgentEngine & MCP
AdvancedRunning and governing AI agents next to the data.
VAST Foundation Stacks
AdvancedProduction RAG, AI-Q research, and video search on NVIDIA Blueprints.
Roadmap & Vision
FundamentalThe AI OS: Polaris, GPU SQL, and the agentic future.