10 tracked signals on safety.
Quoting John Gruber
Simon Willison · Sep 25, 2026
Meta's Muse is the first consumer-accessible agentic AI with persistent Linux VMs per user
“It's the first consumer-accessible agentic AI system, and Meta has truly done an amazing job with that. But it's a genuinely open question whether consumers have any understanding what this means.”
Why Peregrine's founders left national security to work on cities
Sequoia Capital · Sep 04, 2026
Peregrine aims to improve city collaboration through technology and safety.
“The idea is that we can do both: preserve and protect an individual's privacy, while bringing people together.”
Safety and alignment in an era of long-horizon models
OpenAI · OpenAI Blog · Jul 20, 2026
OpenAI reveals new safety risks and failures from deploying long-running AI agents
NVIDIA Open Agent Safety Platform: A Reference for Continuous In-Silicon Agent Monitoring
NVIDIA Developer Blog · Sep 28, 2026
NVIDIA releases open reference platform for continuous hardware-level AI agent safety monitoring
Offering Zero Data Retention for frontier models
OpenAI · OpenAI Blog · Aug 19, 2026
OpenAI previews Private Safety Processing enabling AI safety without compromising customer data privacy
Guardrails First: Engineering Member-Facing Health AI — Rashi Agrawal, Hinge Health
AI Engineer · Aug 19, 2026
ECRI named AI chatbot misuse the #1 health technology hazard of 2026, not a frontier problem but a production baseline issue.
“Most AI safety failures in health care are not model failures. They are architectural decisions that were made before even a single token was generated.”
Introducing ChatGPT for Teens: Built for learning, backed by protections
OpenAI · OpenAI Blog · Aug 18, 2026
OpenAI launches teen-specific ChatGPT with built-in protections and parental controls
Waymo Co-CEO Dmitri Dolgov: The Demo Is Only 1% Of The Work
Y Combinator · Aug 03, 2026
Waymo drives 4 million fully autonomous miles weekly across 15 US cities with superhuman safety
“when you're dealing with atoms instead of bits, breaking things is not really okay”
Evals-Driven Development for a Mental Health AI Coach — Akele Reed & Dave Revere, SonderMind
AI Engineer · Jul 25, 2026
SonderMind built a clinically grounded AI mental health coach using evals-driven development for safety.
“General purpose LLMs however are not built for mental health care which has resulted in some very tragic events.”
Helping build shared standards for advanced AI
OpenAI · OpenAI Blog · Jun 23, 2026
OpenAI supports shared standards for advanced AI via the Appia Foundation.