The Hallway Track

AI Engineer World's Fair

agents engineering production

Dates
2026-06-29 → 2026-07-02
Location
San Francisco, CA
Ecosystem
production engineering
Importance
9/10

Official site →

Related coverage & signals

[AINews] SpaceXAI Grok 4.6 and Grok @Bot

Elon Musk · Latent Space Blog · Aug 13, 2026

SpaceXAI launches Grok 4.6, a 1.5T model targeting knowledge work agents

“builds on Grok 4.5 with a particular focus on long-running agents and more ambitious interactive and visual work”
[AINews] Google I/O 2026: Gemini 3.5 Flash, Omni (NanoBanana for Video), Spark (background agents), and Antigravity 2.0

Latent Space Blog · May 20, 2026

Google I/O 2026 repositioned Gemini as consumer AI surface and developer agent platform with three major launches

“Google used I/O to reposition Gemini as both a consumer AI surface and a developer/agent platform, with three core technical announcements: Gemini 3.5 Flash for fast agentic/coding workloads, Gemini Omni for multimodal generation/editing starting with video, and a broader Antigravity agent stack spanning desktop/CLI”
Quoting John Gruber

Simon Willison · Sep 25, 2026

Meta's Muse is the first consumer-accessible agentic AI with persistent Linux VMs per user

“It's the first consumer-accessible agentic AI system, and Meta has truly done an amazing job with that. But it's a genuinely open question whether consumers have any understanding what this means.”
The New Physics of Business — Garry Tan, Y Combinator

AI Engineer · Jul 17, 2026

YC's Garry Tan claims 400x personal coding productivity gain with AI agents.

“One person does what used to take a thousand people. And I don't mean that as a metaphor. I mean that mechanically this year, the people in this room will do this.”
The Agent for Your Agent.

LangChain · Jun 23, 2026

LangChain launched Engine, an agent that autonomously investigates traces and drafts PRs to improve other agents.

“We're working towards a future where agents improve themselves.”
GLM-5.2 is the step change for open agents

Nathan Lambert · Interconnects · Jun 22, 2026

Z.ai's open-weight GLM-5.2 marks a step-change for open agentic models, rivaling top labs.

“minor version numbers can have AI models crossing meaningful user experience thresholds”
Satya Nadella: Why Humans Still Create Value

Satya Nadella · No Priors · Jun 08, 2026

AI lets enterprises finally capture human capital and tacit knowledge, but humans stay valuable by finding gaps.

“Every company is going to have the human capital that is still going to be super valuable because humans and their ability to find the gaps that exist at all times is going to be the way we all will create value”
Developer Keynote (Google I/O '26) - Audio Described

Google Developers (Google I/O) · May 26, 2026

Google launches Anti-gravity agentic platform and Gemma 4 hits 100M downloads in first month

“It's our smartest open model yet. It's purposebuilt for advanced reasoning, agentic workflows, and the response has been incredible. 100 million downloads in the first month and it's pushing Gemma downloads past half a billion.”
How to Build a Model Router in the Harness

LangChain · Oct 06, 2026

LangChain achieved 64% cost reduction in their coding agent via model routing with no quality loss

“we were able to see a 64% reduction in median cost per thread with no measurable change in quality”
Quoting Felix Rieseberg

Simon Willison · Oct 05, 2026

Claude Cowork moves model inference and VM execution to the cloud for mobile and battery improvements

“The "new" version of Cowork runs model inference and the VM in the cloud. Each session gets its own sandbox, not sharing state with other sessions.”
Grok 4.7 is now available on Amazon Bedrock

AWS Machine Learning Blog · Sep 28, 2026

xAI's Grok 4.7 lands on Amazon Bedrock with 500K context and self-verification for agents

“A model that checks its own output before continuing tends to fail less catastrophically on long trajectories, where an early mistake otherwise compounds through every later step.”
How to go from your agent's traces to a fine-tuned model in one workflow

LangChain · Sep 24, 2026

LangChain launches LangSmith fine-tuning in public beta with SmithTune, a CLI to post-train models from agent traces.

“today we're launching LangSmith fine-tuning in public beta with SmithTune, a CLI to allow you to post-train models from your LangSmith traces in one workflow”
Evaluate skill-equipped agents with Strands Evals and Amazon Bedrock AgentCore

AWS Machine Learning Blog · Sep 22, 2026

Strands Evals and Amazon Bedrock AgentCore add skill-focused evaluators to measure agent skill selection and instruction following.

“A skill is a reusable set of instructions, usually stored in a SKILL.md file, that teaches an agent a domain-specific task like redacting a contract, reconciling an invoice, or following a team’s pull-request conventions.”
Why I still haven’t bought into true RSI

Nathan Lambert · Interconnects · Sep 19, 2026

Lambert argues true recursive self-improvement won't arrive soon; current AI-safety anxiety reflects scaled agents, not imminent superintelligence.

“they’ll turn out to be directionally correct (relative to the expectations of almost anyone not linked to the community) but factually wrong.”
Shutting Off AI Would Be Anarchy

No Priors · Sep 04, 2026

AI verification and validation is crucial in chip development.

“it's like being in the 1990s, you have the internet, but they tell you that you can only use it from two to four o'clock.”
The Evolution of the Agent Harness

Latent Space Blog · Aug 22, 2026

Models absorbing harness capabilities into weights is reshaping agent architecture toward human-attention scaffolding

“The change last winter, last Christmas — it's a little hard to pin down. I mean, the harness changed and a little post-training changed and then new pre-trained models came… but it felt like a big jump which is not that easy to pin down what did it.”
From Primitives to Production: How Anthropic Builds Agents

Databricks · Aug 19, 2026

Anthropic defines agents by leaning on model intelligence with minimal core primitives in a loop.

“Agents in anthropic are actually very simply defined where essentially you want to give the model and lean in on model intelligence as much as possible.”