Railway is repositioning as agent-native cloud infrastructure with 70% margins and 3M users.
“the activation energy to ship something to production should be near zero.”
Railway is repositioning as agent-native cloud infrastructure with 70% margins and 3M users.
“the activation energy to ship something to production should be near zero.”
NVIDIA and OpenAI are building the largest data center ever at gigawatt scale.
“When it's fully built out, it will be the largest data center ever built.”
Anthropic signed $1.25B/month compute deal with SpaceX's Colossus clusters through May 2029
“Cloud Services Agreements with Anthropic PBC...with respect to access to compute capacity across COLOSSUS and COLOSSUS II...the customer has agreed to pay us $1.25 billion per month through May 2029”
California high-speed rail project estimated to cost $236 billion without completed segments.
“Governor Nuomo in a private video... doesn't believe that this project will ever be done in our lifetime.”
NVIDIA is shaping infrastructure for agentic AI with a new platform called Vera Rubin.
“The value gets generated in the computing and the transformation of the actual data into a calculation.”
Meta introduces ZGateway, a proxy for ZippyDB traffic management.
A new industrial revolution is unfolding powered by AI technologies.
“No company can build this infrastructure alone.”
Memory prices up 500% in 12 months as hyperscalers lock in all 2027 DRAM production capacity
“Some are calling it the RAMpocalypse; I prefer "RAMageddon."”
Databricks launched Omnigent, an open-source 'meta-harness' layer on top of the agentic stack to make agents effective at scale.
“we call it a meta harness of harnesses, if you know what an agent harness is.”
NVIDIA's Jensen Huang frames the AI factory as the largest infrastructure buildout in history and the best enterprise investment of the next decade.
“you can't think if you don't generate words”
AWS launches AgentCore Payments enabling AI agents to autonomously execute microtransactions via stablecoins
“Amazon Bedrock AgentCore payments is the first managed service within Amazon Bedrock AgentCore that helps AI agents autonomously execute microtransaction payments for paid APIs, MCPs, and content with a few lines of code.”
AWS AgentCore Runtime Instances enable GPU-backed, 14-day multi-agent sessions on persistent EC2
“serverless sessions that cap at a few hours don't cut it”
Hyperscalers are investing all short-term operating cash flow into AI capacity as demand outpaces supply
“Hyperscalers are investing all of their short-term operating cash flow into building this capacity to meet demand that continues to outpace supply in almost every case we see”
Databricks launches Lakebase Search with full-text and vector search natively in Postgres for AI agents
NVIDIA releases open reference platform for continuous hardware-level AI agent safety monitoring
Google DeepMind adds private, server-side memory to Private AI Compute for personal AI.
“Introducing private, server-side memory to Private AI Compute for personal AI.”
Stripe engineer demos autonomous shopping agent as agentic commerce infrastructure matures
“the infrastructure for agentic transactions has been laid down by companies like Google, OpenAI, and Stripe”
NVIDIA NVLink Fusion enables NVHBM memory for custom XPU accelerators at hyperscale
Agents will use the web 1,000x more than humans, requiring reinvented search infrastructure
“We started parallel with the bet that agents would do it a thousandx more than humans ever have.”
OpenAI CFO frames full-stack chip-to-product compounding as path to cheaper, scalable intelligence
Meta open-sourced MetaRoCE, a new RDMA transport protocol built for million-GPU AI clusters on Ethernet.
“The fabric sees packets, but the NIC sees intent.”
Agent infrastructure is now commoditized by cloud platforms, making context the new competitive frontier
“They're all taxes one has to pay in order to get an agent out there to play the game.”
Modal proposes decoupling RL rollout workers from trainer clusters to use distributed GPU capacity across datacenters.
“IO wants all four of these at the same time. Enough GPU, same region, fast fabric, and available now. Any of these like is manageable, but all four of them that are pretty hard to get at the same time.”
Baseten raised $13B Series F as inference engineering emerges as a critical AI discipline
“How do you turn those weights from training into a product that is fast, reliable, and affordable at scale?”
Meta doubled GEM ads model training efficiency to 20-25% MFU while scaling FLOPs 4x in 12 months
Y Combinator is seeking founders to build offshore AI compute flotillas on the ocean
“It sounds crazy, but we think part of the answer may be to move compute offshore.”
Most production ML security breaches stem from basic infrastructure mistakes, not exotic AI attacks
“almost everything that is breaking in the production ML security isn't some exotic AI attack. It's the same boring infrastructure mistakes that we supposedly fixed years ago.”
As AI agents move to production, the core challenge shifts from model intelligence to running probabilistic agents on deterministic infrastructure.
“These systems are fundamentally probabilistic. Infrastructure is not allowed to be.”
Nebius, backed by $2B Nvidia investment, offers full-stack open LLM inference from silicon to service
“Typically, most AI teams are stuck between choosing two bad options. Closed APIs are very easy to get started with, but you often hit a ceiling very quickly.”
NVIDIA releases domain-specific DOCA Agent Skills for BlueField infrastructure development
“AI agents are becoming a standard part of development workflows, but general-purpose agents weren't built with specialized infrastructure software such as NVIDIA DOCA in mind.”
NVIDIA launches cuObject and SCADA Server SDK for faster AI storage access
NVIDIA VSS Blueprint 3.3 reduces cost of building production-scale visual AI agents
AWS achieves 40% throughput gain for MoE reinforcement learning using EKS, EFA, and DeepEP
Salesforce launches Mission Force, an agentic AI platform bringing private-sector tools to public-sector and first responders.
“Mission force is more than technology. It's trust in the moments when decisions matter most.”
AI agent payments lack controls, requiring new infrastructure beyond legacy payment systems
“the missing infrastructure layer for AI payments”
NVIDIA Dynamo's shadow engine recovery restores LLM inference capacity in seconds instead of minutes
NVIDIA Spectrum-X Ethernet reengineers networking for giga-scale AI GPU clusters
NVIDIA BlueField-4 DPU enables dedicated networking for agentic AI factory infrastructure
“Agentic AI factories connect diverse users, agents, applications, data sources, and storage systems to massively accelerated compute at multi-terabit bandwidth per server, making dedicated DPU processing essential for line-rate networking, storage, and security.”
Warp built a cloud agent platform after hitting limits of local laptop-based AI coding agents.
“we realized that we had kind of reached the limits of what we could do on our laptops, and we wanted agents to do work that was more long-running, that was adapted to different constraints”
LLMs are shifting recommender systems from embedding-similarity to generative next-action prediction
“The advent of LLMs has inspired a shift from the traditional embedding-similarity-based objective to a generative one, where the goal is to predict the next action or item in a large catalog given a sequence of user histories.”