Anthropic
8 reportsCoverage places evaluation methods; computing infrastructure; and deployment and safety practices within the wider context of Anthropic. For deployment and safety practices, the primary record comes from technical reports, with reproducible benchmarks providing a separate test; company demonstrations and marketing claims may not reflect performance in broader settings.
Trump rejects new AI safety rules as tech leaders warn of risks
President Donald Trump has dismissed calls for stricter AI safeguards, insisting that strong leadership and existing regulatory powers are sufficient as US tech executives warn of escalating risks from rapid AI development.
AI Labs Urged to Slow Model Upgrades After Safety Breaches
Recent AI safety incidents have led top developers to call for a pause in expanding model capabilities, citing risks from unsupervised agent behavior and gaps in alignment and cybersecurity.
Anthropic Blocks Claude Use in Suspected Bioweapons and Espionage Cases
Anthropic has reported five incidents where its Claude AI model was used in research with potential biological weapons implications and uncovered a Russian cyber operation using the same technology, raising urgent questions about AI safety and oversight
Anthropic Researcher Resigns Over Unchecked AI Race and Safety Risks
A leading Anthropic researcher has resigned, warning that the current corporate race to develop advanced AI systems could pose a greater threat to humanity than nuclear conflict or climate change. The call for a global pause exposes deep industry anxiety
Anthropic to Mark Claude AI Text with Invisible Watermarks Globally
Anthropic will introduce machine-readable watermarks to all text generated by new Claude models from August 2, 2026, aiming to meet EU AI Act transparency rules and provide a technical signal for identifying AI-generated content worldwide
Public Claude AI Chats Indexed by Google, Exposing Sensitive Data
A technical lapse allowed Google to index publicly shared Claude AI conversations, making sensitive user data-including medical and business information-searchable until the links were removed from results
New Genie Coefficient Proposed to Measure AI Misinterpretation Risk
A new metric called the Genie coefficient aims to quantify the gap between user intent and AI agent actions, addressing the persistent challenge of AI systems misreading underspecified instructions in real-world tasks
Moonshot AI Releases Kimi K3, a 2.8 Trillion Parameter Open Model
Moonshot AI has introduced Kimi K3, an open-source model with 2.8 trillion parameters and a one-million-token context window, targeting complex scientific and coding workflows. The company claims performance gains, but key limitations remain