A teenage person sitting alone on a park bench looking at their smartphone in daylight

Mental Health Chatbots Now Face a Safety Test They Often Fail

Mental health and youth-facing chatbots are no longer plausibly treated as neutral software tools. Recent research and new legislation point in the same direction: these systems often fail basic safety and privacy checks in predictable ways, and regulators are starting to require guardrails that many deployments still do not meet. Research is finding repeatable failure…

Read More
A doctor consulting with a patient using AI-assisted technology displayed on a computer screen in a clinical setting.

Google DeepMind’s AI Co-Clinician Is Strongest as a Supervised Teammate, Not an Autonomous Doctor

Google DeepMind’s AI Co-Clinician matters because it pushes medical AI beyond a chatbot or back-office assistant, but its real advance is narrower than some headlines suggest: it works best as a supervised clinical teammate inside the consultation, not as a doctor substitute. The system combines multimodal inputs and multi-agent checks to support decisions in real…

Read More
a building with glass windows

OpenAI Buys Promptfoo to Build Security and Compliance Into Enterprise AI Agents

OpenAI’s acquisition of Promptfoo is less about adding another AI feature and more about moving security testing and compliance checks into the core of enterprise agent deployment. The practical change is that OpenAI wants automated red-teaming, vulnerability detection, reporting, and traceability to sit inside Frontier, its enterprise AI agent platform, instead of being treated as…

Read More
Robots performing tasks in a robotics competition arena with engineers observing in the background.

The DARPA Robotics Challenge Mattered Most as a Deployment Test, Not Proof Humanoid Robots Were Ready

The 2015 DARPA Robotics Challenge was valuable because it tested whether disaster-response robots could keep working through real operating constraints, not because it proved humanoid robots were ready for field deployment. By forcing teams to complete eight sequential mobility and manipulation tasks under degraded communications and without physical resets, the challenge exposed where supervised autonomy…

Read More
Professor studies complex formulas on a blackboard.

Google’s Bayesian Teaching Upgrade Gives LLMs a Better Way to Update Beliefs

Google Research’s Bayesian Teaching work matters because it targets a specific weakness in current LLMs: they often stop learning anything useful about a user after the first exchange. Instead of fine-tuning models to reproduce final correct answers, Google trains them to imitate a Bayesian assistant’s step-by-step probability updates, so the model learns how to revise…

Read More
A spacecraft orbiting Jupiter with visible glowing radiation belts around the planet, illustrating the intense magnetosphere environment.

Cassini Changed the Risk Map: Why Jupiter Radiation Planning Now Depends on the Electrons Below the Peak

Jupiter mission planning is no longer based on a simple assumption that its radiation belts are just a larger, harsher version of Earth’s. For ESA’s Juice mission, the decisive shift is that newer measurements, especially from Cassini’s 2000 flyby, point to a more uneven and in some ways more dangerous electron environment than older expectations…

Read More
woman carrying white and green textbook

If the Pilot Cuts Teacher Workload, ATL Saathi Could Change How India’s Tinkering Labs Actually Operate

Google DeepMind and India’s Atal Innovation Mission have launched ATL Saathi as a teacher-facing AI assistant for Atal Tinkering Labs, not as a chatbot for students. That distinction matters because the 100-school pilot is testing a specific deployment thesis: whether Gemini-based support can turn scattered training material and ad hoc project guidance into a practical…

Read More