Want to run the models behind your AI apps?
@nilekh.bsky.social and I are teaching a vLLM tutorial at #KubeCon + #CloudNativeCon NA 2026. We'll start on CPUs and work up to serving LLMs across multiple GPU nodes.
Nov 11, Salt Lake City
bit.ly/kcna26-suraj...