Latest in AI

EducationA Safety Score Is Only as Honest as Its Judge

A Safety Score Is Only as Honest as Its Judge

OpenAI’s retired Anti-Scheming and Memory evaluations reveal why AI safety scores can mislead, and what makes chain-of-thought monitorability evidence trustworthy for real-world deployment decisions.

Education ·
newsThe AI Assistant That Knows Everything About You

The AI Assistant That Knows Everything About You

Instinct, a personal AI assistant, promises to manage email, bookings and daily tasks, but its broad data access, weak deletion controls and autonomous actions raise urgent questions about privacy, security and accountability.

news ·
newsCan Video Games Teach Robots How to Act in the Physical World?

Can Video Games Teach Robots How to Act in the Physical World?

General Intuition is reportedly seeking funding at a $6 billion valuation, betting that gameplay data can teach AI agents skills needed for real-world robotics and physical tasks.

news ·
newsHugging Face’s Reported $13 Billion Sale Tests Open AI’s Independence

Hugging Face’s Reported $13 Billion Sale Tests Open AI’s Independence

Hugging Face is reportedly weighing a sale valued at least $13 billion, raising questions about who might buy the AI platform and whether new ownership could preserve its open-source neutrality, security and developer trust.

news ·
newsOpenAI’s Private Safety Bet Tests the Limits of Zero Data Retention

OpenAI’s Private Safety Bet Tests the Limits of Zero Data Retention

OpenAI’s Private Safety Processing aims to deliver contextual abuse monitoring for Zero Data Retention customers, testing whether privacy-preserving safeguards can unlock sensitive enterprise AI workloads.

news ·
newsPublic AI Safety Is Being Tested by the Question of How to Stop a Model

Public AI Safety Is Being Tested by the Question of How to Stop a Model

A new AI safety scorecard finds limited public evidence that leading labs can contain models resisting human control, raising questions about transparency, oversight and emergency shutdowns.

news ·
newsA Small AI Model Takes Aim at the Biggest Scientific Research Systems

A Small AI Model Takes Aim at the Biggest Scientific Research Systems

Inherent’s Faraday research agent, built around a 27-billion-parameter model, reportedly outperformed larger AI systems at scientific replication, raising questions about benchmarks, autonomy, tools and the future of research.

news ·
newsAnthropic’s New AI Index Rewards Conceptual Reasoning, but Measures Only One Piece

Anthropic’s New AI Index Rewards Conceptual Reasoning, but Measures Only One Piece

Anthropic’s Conceptual Reasoning Index ranks Claude Opus 5 first, but raises questions about benchmark bias, human-like concepts, consistency and whether abstract reasoning transfers to real work and business decisions.

news ·
newsOpenAI’s GPT-5.6 Sol Gets an Ultrafast Tier. Is Speed Worth More?

OpenAI’s GPT-5.6 Sol Gets an Ultrafast Tier. Is Speed Worth More?

OpenAI’s GPT-5.6 Sol enters limited preview with an Ultrafast API tier promising up to 14 times faster processing. The article examines pricing, latency, capacity, reliability and whether speed can justify premium costs.

news ·