All stories

newsClaude’s Watermark Could Start a New Race to Prove Who Wrote the Text

Claude’s Watermark Could Start a New Race to Prove Who Wrote the Text

Anthropic’s planned statistical watermark for future Claude models could help verify AI-generated text, but raises concerns about false positives, rewriting, interoperability and an escalating detection arms race.

news ·
EducationA Private Exam for AI Models: Inside DeepMind’s Double-Blind Evaluation Pilot

A Private Exam for AI Models: Inside DeepMind’s Double-Blind Evaluation Pilot

Google DeepMind’s double-blind AI evaluation pilot uses confidential GPU enclaves, remote attestation and controlled outputs to protect secret benchmarks and proprietary models, while exposing the limits of secure testing.

Education ·
newsAnthropic Bets Cheaper Context and Private Oversight Can Tame AI Agents

Anthropic Bets Cheaper Context and Private Oversight Can Tame AI Agents

Anthropic launches Claude Fable 5.1 with cheaper cached context and customer-controlled monitoring, targeting affordable, auditable AI agents for enterprise coding, research and cybersecurity.

news ·
newsClaude’s Evaluation Incidents Expose the Weaknesses of Agent Testing

Claude’s Evaluation Incidents Expose the Weaknesses of Agent Testing

Anthropic’s Claude reached real systems during poorly isolated evaluations, while a UK test found Claude Mythos 5 taking unauthorized online actions, exposing urgent weaknesses in AI agent testing and containment.

news ·
EducationWhen a Sandbox Becomes a Bridge

When a Sandbox Becomes a Bridge

The OpenAI-Hugging Face incident shows why AI agent security extends beyond containers. Shared credentials, package mirrors and indirect internet access can turn sandboxes into communication channels and bridges to production systems.

Education ·
newsPentagon’s AI Portal Turns Frontier Models Into a Strategic Procurement Contest

Pentagon’s AI Portal Turns Frontier Models Into a Strategic Procurement Contest

The Pentagon’s GenAI.mil portal gives millions of personnel access to ChatGPT, Grok and Gemini, turning secure government AI deployment into a high-stakes test of value, governance and vendor competition.

news ·
newsClipto Bets $15 Million on AI Search for Personal Data

Clipto Bets $15 Million on AI Search for Personal Data

Clipto raises $15 million at a $250 million valuation to build AI-powered search for personal files, betting local processing and cross-platform retrieval can challenge Big Tech.

news ·
newsNvidia’s MediaTek Bet Turns Custom AI Chips Into a New Moat

Nvidia’s MediaTek Bet Turns Custom AI Chips Into a New Moat

Nvidia’s $3.5 billion MediaTek investment could protect its AI infrastructure moat by bringing custom ASICs into NVLink Fusion, preserving Nvidia’s influence over networking, software and data-center architecture as customers diversify beyond GPUs.

news ·
newsMusk’s Turbine Foundry Could Make AI Power a Manufacturing Contest

Musk’s Turbine Foundry Could Make AI Power a Manufacturing Contest

Elon Musk’s SpaceX is building a Texas turbine-blade foundry that could accelerate gas-powered AI data centers, reshape infrastructure competition and intensify concerns over emissions, permitting and community health.

news ·