Anthropic
4 reportsCoverage places evaluation methods; computing infrastructure; and deployment and safety practices within the wider context of Anthropic. For deployment and safety practices, the primary record comes from technical reports, with reproducible benchmarks providing a separate test; company demonstrations and marketing claims may not reflect performance in broader settings.
Anthropic to Mark Claude AI Text with Invisible Watermarks Globally
Anthropic will introduce machine-readable watermarks to all text generated by new Claude models from August 2, 2026, aiming to meet EU AI Act transparency rules and provide a technical signal for identifying AI-generated content worldwide
Public Claude AI Chats Indexed by Google, Exposing Sensitive Data
A technical lapse allowed Google to index publicly shared Claude AI conversations, making sensitive user data-including medical and business information-searchable until the links were removed from results
New Genie Coefficient Proposed to Measure AI Misinterpretation Risk
A new metric called the Genie coefficient aims to quantify the gap between user intent and AI agent actions, addressing the persistent challenge of AI systems misreading underspecified instructions in real-world tasks
Moonshot AI Releases Kimi K3, a 2.8 Trillion Parameter Open Model
Moonshot AI has introduced Kimi K3, an open-source model with 2.8 trillion parameters and a one-million-token context window, targeting complex scientific and coding workflows. The company claims performance gains, but key limitations remain