Articles Tagged “Responsible AI”
7 articles found

Claude Text Watermarking: What App Builders Should Know
Anthropic is embedding invisible watermarks in Claude text output and C2PA metadata in image files, with a detection API confirmed as on the roadmap.

Fable 5 Biology Safeguards Cut False Positives by 85%
Anthropic retuned Fable 5's biology classifier, cutting biology fallbacks roughly 85% and total fallbacks 67% on Claude.ai while keeping dual-use limits.

GPT-Live Audio Gets SynthID Watermarks and a Verify API
OpenAI now embeds Google DeepMind's SynthID watermark in all GPT-Live audio and opened a verification API so any team can check provenance automatically.

Claude Reflect Helps You Use AI More Mindfully
Anthropic's new Reflect dashboard, launched July 9, shows your Claude usage over 1 to 12 months and adds quiet hours and break reminders.

Google DeepMind's AI Control Roadmap Charts a Safer Path for AI Agents
Google DeepMind published its AI Control Roadmap on June 18, 2026 — a defense-in-depth framework for safely deploying AI agents that keeps systems secure even when alignment is imperfect.

Cognizant Launches Secure AI Services — A Build-Time and Run-Time Trust Platform for Agentic Enterprise AI
Cognizant launched Secure AI Services on May 7, 2026 — a new integrated offering that combines a Secure Agent Development Lifecycle, Neuro Cybersecurity, and Responsible AI to govern and scale enterprise agentic systems.

Microsoft Lays Out a Pre-Deployment Playbook for Frontier AI Security
Microsoft published a detailed pre-deployment AI security playbook on May 1, 2026 — Brad Smith and Natasha Crampton's blueprint for how frontier AI developers, governments, and deployers should secure the next generation of agentic models together.
