The Hallway Track

transformers

8 tracked signals on transformers.

Betting on Diffusion

No Priors · Sep 20, 2026

A startup is betting on diffusion-based LLMs because they are inherently more parallel at inference time than autoregressive transformers.

“the bitter lesson is that the more parallel solution is the one that is eventually going to win”
What building Tesla's autopilot chips taught Cognition's President about AI

LangChain · Oct 02, 2026

Transformer architecture standardization will enable extreme chip specialization and massive inference performance gains

“I think we're gonna have crazy levels of specialization because there's gonna be so much inference demand. And that's gonna unlock just like super exciting levels of performance.”
How to Optimize Transformer-Based Models for Low-Precision Training

NVIDIA Developer Blog · Jun 16, 2026

Optimizing transformers for low-precision training cuts GPU hours and speeds up experimentation and model scaling.

“Accelerating transformers is therefore not just a performance optimization, but directly affects how quickly teams can experiment and how large a model they can afford to train.”