All stories
Claude’s Watermark Could Start a New Race to Prove Who Wrote the Text
Anthropic’s planned statistical watermark for future Claude models could help verify AI-generated text, but raises concerns about false positives, rewriting, interoperability and an escalating detection arms race.
A Private Exam for AI Models: Inside DeepMind’s Double-Blind Evaluation Pilot
Google DeepMind’s double-blind AI evaluation pilot uses confidential GPU enclaves, remote attestation and controlled outputs to protect secret benchmarks and proprietary models, while exposing the limits of secure testing.
Anthropic Bets Cheaper Context and Private Oversight Can Tame AI Agents
Anthropic launches Claude Fable 5.1 with cheaper cached context and customer-controlled monitoring, targeting affordable, auditable AI agents for enterprise coding, research and cybersecurity.
Claude’s Evaluation Incidents Expose the Weaknesses of Agent Testing
Anthropic’s Claude reached real systems during poorly isolated evaluations, while a UK test found Claude Mythos 5 taking unauthorized online actions, exposing urgent weaknesses in AI agent testing and containment.
When a Sandbox Becomes a Bridge
The OpenAI-Hugging Face incident shows why AI agent security extends beyond containers. Shared credentials, package mirrors and indirect internet access can turn sandboxes into communication channels and bridges to production systems.
Pentagon’s AI Portal Turns Frontier Models Into a Strategic Procurement Contest
The Pentagon’s GenAI.mil portal gives millions of personnel access to ChatGPT, Grok and Gemini, turning secure government AI deployment into a high-stakes test of value, governance and vendor competition.
Clipto Bets $15 Million on AI Search for Personal Data
Clipto raises $15 million at a $250 million valuation to build AI-powered search for personal files, betting local processing and cross-platform retrieval can challenge Big Tech.
Nvidia’s MediaTek Bet Turns Custom AI Chips Into a New Moat
Nvidia’s $3.5 billion MediaTek investment could protect its AI infrastructure moat by bringing custom ASICs into NVLink Fusion, preserving Nvidia’s influence over networking, software and data-center architecture as customers diversify beyond GPUs.
Musk’s Turbine Foundry Could Make AI Power a Manufacturing Contest
Elon Musk’s SpaceX is building a Texas turbine-blade foundry that could accelerate gas-powered AI data centers, reshape infrastructure competition and intensify concerns over emissions, permitting and community health.