The Hallway Track

human-in-the-loop

22 tracked signals on human-in-the-loop.

Your AI Agent Is Confidently Wrong About Production — Willem Pienaar, Cleric

AI Engineer · Oct 05, 2026

AI agents confidently misdiagnose production failures due to lack of verification signals

“The agent is very confident in returning a response. He says, "You know, you have a memory leak from the code you just deployed." And it still takes a person to go and check if it really happened.”
The Agentic Commerce Stack — Ahnaf Prio, Best Buy

AI Engineer · Aug 27, 2026

Best Buy reports 45% of AI agent sessions on major platforms involve shopping queries.

“right now about 45% of all agent sessions that happen within major providers like chat.gbt.com and Google Gemini are related to shopping”
Production-grade AI agents for financial compliance: Lessons from Stripe

AWS Machine Learning Blog · Jun 26, 2026

Stripe built a production-grade AI agent system on Amazon Bedrock that cut compliance review handling time by 26 percent.

“skilled analysts were spending up to 80% of their time navigating fragmented systems to gather documentation rather than performing high-value risk assessments”
Observing And Testing CX Agents | Interrupt 26

LangChain · Jun 10, 2026

Cisco built an agentic feedback loop turning production thumbs-down signals into merged PR fixes with human oversight.

“every thumbs up is a lead. Every trace with an error is a potential regression. And every confused user is a description of a gap that needs to be addressed.”
Why AI Is Reinventing How Businesses Buy Everything

a16z · Oct 02, 2026

AI agents are taking over end-to-end corporate procurement, starting with human-in-the-loop trust-building

“We're going to take control of this entire process from start to finish.”
What Is Human in the Loop?

LangChain · Oct 01, 2026

Best production agents use deliberate human checkpoints rather than pursuing full autonomy from day one

“The best teams, sort of like shipping the best agents, just aren't really going for a full autonomy on day one.”
Don't be a meat proxy

Simon Willison · Aug 03, 2026

New term 'meat proxy' describes people who blindly relay AI output without validation

“By all means, prompt AI. But don't just relay the output. Read it, understand it, validate it, and then write a response in your own words (a decent certificate that you've done the prior steps). Making that effort is value you can add.”
Quoting Jon Udell

Simon Willison · Jun 28, 2026

Jon Udell argues we should reframe 'human in the loop' as inviting agents into our development loop, not joining theirs.

“I dislike the phrase “human in the loop” because it cedes authority to the machines.”
Quoting Emanuel Maiberg, 404 Media

Simon Willison · Jun 04, 2026

Google quietly removed its public commitment to keeping 'humans in the loop' from an AI statement.

“The new statement no longer stated that "it's critical that we maintain humans in the loop."”
datasette-agent 0.2a0

Simon Willison · Jun 10, 2026

datasette-agent 0.2a0 adds mid-execution user questions and human-approved query saving.

“Tools that declare a context parameter receive a ToolContext object, and await context.ask_user(...) can ask a yes/no, multiple-choice (options=[...]) or free-text (free_text=True) question.”