Anthropic

3 reports
Anthropic is an artificial intelligence organization that develops models, training methods, evaluations, infrastructure, or safety practices. The clearest evidence of its role comes from AI research program, safety evaluation, and model governance, together with documented outcomes and accessible records.

Coverage places evaluation methods; computing infrastructure; and deployment and safety practices within the wider context of Anthropic. For deployment and safety practices, the primary record comes from technical reports, with reproducible benchmarks providing a separate test; company demonstrations and marketing claims may not reflect performance in broader settings.

Official website

Public Claude AI Chats Indexed by Google, Exposing Sensitive Data

A technical lapse allowed Google to index publicly shared Claude AI conversations, making sensitive user data-including medical and business information-searchable until the links were removed from results

Read the analysis

New Genie Coefficient Proposed to Measure AI Misinterpretation Risk

A new metric called the Genie coefficient aims to quantify the gap between user intent and AI agent actions, addressing the persistent challenge of AI systems misreading underspecified instructions in real-world tasks

Read the analysis

Moonshot AI Releases Kimi K3, a 2.8 Trillion Parameter Open Model

Moonshot AI has introduced Kimi K3, an open-source model with 2.8 trillion parameters and a one-million-token context window, targeting complex scientific and coding workflows. The company claims performance gains, but key limitations remain

Read the analysis