blog.atlasinference.io
Notes from the inference layer
Kernel work, measured benchmarks, and what it takes to run frontier models on hardware you own. Everything we publish is reproducible from a commit.
Engineering Featured
Seven Tenets Powering Atlas Inference Accelerated Workloads
Atlas Inference is a free and open source LLM inference engine written from scratch in Rust. These are the seven philosophical tenets we started it on, and why we left the Python vLLM stack to do it.