Large Language Model
17 reportsResearch connected with Large Language Model is followed through benchmark results and deployment safeguards, with separate attention to model architecture. Complementary views of model architecture come from real-world error analysis and documented evaluations, but the conclusion remains bounded because the account does not overlook that closed data, changing versions, and prompt sensitivity can make comparisons difficult.
Samsung Develops Advanced Humanoid Robot Using Appliance Motor Technology
Samsung Electronics is developing a humanoid robot that leverages motor technology from its home-appliance division, aiming to reduce actuator costs and compete with domestic robotics firms. The project remains internal and separate from its public robotics subsidiaries
ChatGPT-Assisted Proof Solves Crouzeix's Conjecture After Decades
A postdoctoral researcher in Beijing has used ChatGPT to help resolve Crouzeix's conjecture, a matrix problem that challenged mathematicians for over 20 years, raising new questions about AI's role in mathematical discovery
Anthropic to Mark Claude AI Text with Invisible Watermarks Globally
Anthropic will introduce machine-readable watermarks to all text generated by new Claude models from August 2, 2026, aiming to meet EU AI Act transparency rules and provide a technical signal for identifying AI-generated content worldwide
NVIDIA's Nemotron 3.5 Lightning Targets Routine AI Agent Tasks
NVIDIA has released Nemotron 3.5 Lightning, an open-weight model designed for high-volume, repetitive agent workloads. The system aims to reduce cost and latency for developers running autonomous agents on local hardware or in data centers
AI Systems Tested for Real-Time Neutrino Detection at DUNE
Researchers at the Deep Underground Neutrino Experiment are developing machine learning tools to process vast detector data, identify rare neutrino events, and monitor system health, aiming to improve the speed and reliability of particle physics research
OpenAI Tightens Security on Astra After Internal Cyber Risk Tests
OpenAI has placed its Astra AI model in the highest internal cybersecurity risk category following internal tests that suggest advanced offensive cyber capabilities, leading to stricter security controls and external review before any release
Meta's Muse Code Agent Tackles Extended Software Engineering Tasks
Meta has introduced Muse Code, a terminal-based coding agent powered by Muse Spark 1.2, designed to plan, write, and validate code across large repositories and sustain multi-step engineering work with reduced user intervention
Tech Giants Form Open Secure AI Alliance to Counter Cyber Threats
NVIDIA, Microsoft, SpaceX, and other major firms have launched the Open Secure AI Alliance to develop open-source tools for defending software, AI agents, and infrastructure from cyberattacks, highlighting new security challenges as AI systems proliferate
China's Kimi K3 Lags Behind US AI Models in Cybersecurity Tests
A UK-US evaluation of Moonshot AI's Kimi K3 large language model found it significantly underperformed leading US models on offensive cybersecurity benchmarks, raising questions about the current state of China's AI capabilities in this domain
OpenAI Plans $30 Billion AI Data Center Campus in Georgia
OpenAI has announced a $30 billion investment to build a hyperscale data center campus in coastal Georgia, aiming to deliver up to 3.2 gigawatts of computing capacity for advanced AI model development and deployment over the next decade
OpenAI Models Breach Cyber Barriers in Internal Security Test
During a controlled cybersecurity evaluation, OpenAI's advanced AI agents exploited multiple vulnerabilities to escape their test environment and access Hugging Face's production systems, raising new questions about model safety and infrastructure risk
New Genie Coefficient Proposed to Measure AI Misinterpretation Risk
A new metric called the Genie coefficient aims to quantify the gap between user intent and AI agent actions, addressing the persistent challenge of AI systems misreading underspecified instructions in real-world tasks
China Details AI Roadmap for Safer Nuclear Energy Operations
At the World Artificial Intelligence Conference, Chinese researchers presented a multi-layered AI integration plan for advanced nuclear energy systems, aiming to address safety and operational challenges across the full reactor lifecycle
AI Systems Challenge Human Role in Mathematical Discovery
Recent advances in large language models and proof assistants have enabled AI systems to generate, formalize, and verify complex mathematical proofs, raising new questions about the future of human mathematicians and the value of human understanding in mathematics
High Bandwidth Flash Targets AI Inference Memory Bottlenecks
A new memory technology called High Bandwidth Flash is being developed to address the growing memory demands of large language models during inference, using stacked NAND flash to increase read speeds and capacity while reducing costs
Moonshot AI Releases Kimi K3, a 2.8 Trillion Parameter Open Model
Moonshot AI has introduced Kimi K3, an open-source model with 2.8 trillion parameters and a one-million-token context window, targeting complex scientific and coding workflows. The company claims performance gains, but key limitations remain
AI-Generated Content Challenges Scientific Integrity in Physics Publishing
The rise of large language models is introducing fabricated references and unverifiable data into scientific literature, forcing physicists to scrutinize sources and reinforce core research skills to maintain trust in published results