blog.atlasinference.io
Notes from the inference layer
Kernel work, measured benchmarks, and what it takes to run frontier models on hardware you own. Everything we publish is reproducible from a commit.
Engineering Featured
DFLASH-2: the fastest single-machine numbers Atlas has produced
66.6 tokens per second on a stock build, one DGX Spark, one stream, and every figure reproducible from a commit.