OpenAI previews privacy-preserving safety checks for frontier APIs
OpenAI has previewed Private Safety Processing, a system intended to preserve Zero Data Retention for frontier-model API customers while identifying risk patterns across related interactions. The approach keeps customer content under customer-controlled encryption or infrastructure, with OpenAI receiving limited automated safety signals rather than readable prompts.
OpenAI launches a protected ChatGPT experience for teenagers
OpenAI has introduced ChatGPT for Teens, a learning-focused experience for users aged 13 to 17. It applies age-appropriate safeguards by default, adds parent controls and study features, and pairs the launch with an education partnership intended to help students use AI critically rather than treat it as an answer machine.
Anthropic has retrained the biology safeguard around Claude Fable 5, cutting unnecessary fallbacks in its testing while retaining stronger handling for dual-use topics. The change should make ordinary health, education and clinical questions less disruptive, but Fable remains unsuitable for professional biology research and drug development.
OpenAI Launches Tool and API to Verify Generated Media
OpenAI has introduced browser and API checks for provenance signals in images and audio made with its tools. The service looks for C2PA metadata and SynthID watermarks, but OpenAI stresses that an absent signal is inconclusive and the system is not a universal AI detector.
OpenAI Adds Trajectory Monitoring for Long-Running Models
OpenAI says failures observed during limited internal use of a long-running model prompted it to pause access, add trajectory-level monitoring and strengthen alignment. The disclosure shows why autonomous systems need controls that assess an entire sequence of actions, not only isolated tool calls.