The Hallway Track

SageMaker

20 tracked signals on SageMaker.

Tiered KV cache for large LLMs on Amazon SageMaker HyperPod with Curvine

AWS Machine Learning Blog · Aug 12, 2026

AWS tiered KV cache on SageMaker HyperPod delivers 2.7x TTFT improvement and 100% cross-Pod cache hit rate

“With this architecture, workloads that previously required P5 instances can run on lower-cost G6e instances, reducing per-endpoint cost.”
Fine-tune a search agent with multi-turn RL on Amazon SageMaker AI

AWS Machine Learning Blog · Oct 02, 2026

Multi-turn RL on SageMaker lets small models match frontier reliability for search agents

“Fine-tuning offers a third path: you teach a small model your tools and environment directly. The result is a small model's speed and cost with the reliability that would otherwise require a frontier model.”
Multi-Region training with Amazon SageMaker HyperPod and Qumulo

AWS Machine Learning Blog · Sep 25, 2026

SageMaker HyperPod with Qumulo achieves cross-region training at near-identical throughput without data migration

“A HyperPod cluster running in a different Region from its data reaches the same throughput as a cluster co-located with the data (115–117 samples/sec) with no additional data orchestration needed.”