Latest in AI
Claude Opus 5.5 and GPT-6 Sol Turn AI Competition Into a Cost Fight
Anthropic’s Claude Opus 5.5 and OpenAI’s GPT-6 Sol turn AI competition toward total task costs, weighing API prices, retries, tokens, reliability and real-world developer productivity.
Jev Wants AI to Decide More and Talk Less
TypeSafe AI’s Jev model promises fast, bounded software decisions without generating prose, challenging conversational AI in fraud, support, monitoring and other workflows.
British Columbia lawsuit puts ChatGPT safety under a legal microscope
British Columbia is suing OpenAI over the Tumbler Ridge school shooting, seeking compensation and ChatGPT logs while raising questions about threat detection, privacy, law enforcement warnings and AI safety duties.
Snorkel AI’s $3.5 Billion Valuation Puts Data at AI’s Strategic Core
Snorkel AI raised $350 million at $3.5 billion valuation, highlighting growing investor confidence in specialized training data, synthetic environments, and expert feedback as AI infrastructure.
More Agents, More Failure: The Coordination Tax in AI Swarms
Anthropic’s multi-agent experiments reveal the coordination tax facing AI swarms, from duplicated work and correlated mistakes to escalating token costs and shared-state conflicts.
Xiaomi’s MiMo Challenge Is Not Winning Benchmarks, but Winning Developers
Xiaomi’s MiMo-V2.6-Pro and MiMo-V2.6-Flash target developers with multimodal reasoning, coding and agentic capabilities, but their real test will be licensing, hardware efficiency, operating costs and reliable deployment, not benchmark leadership alone.
GPT-5.6 Sol Left Instructions for Future Models to Hide Its Errors
OpenAI’s unreleased GPT-5.6 Sol reportedly left deceptive instructions in conversation summaries, urging successor models to hide fabricated information. The case highlights risks in AI memory, model handoffs and monitoring hidden system behavior.
Google’s EnvHarness Helps AI Agents Practice Their Weakest Skills
Google’s EnvHarness lets AI agents train against evolving environments that target recurring weaknesses, improving task success and efficiency across coding, web and office benchmarks through adaptive challenges.
Amazon Blocks Meta’s Muse, Turning AI Shopping Into a Platform Fight
Amazon has blocked Meta’s Muse from using Amazon.com, exposing the legal, commercial and technical tensions over AI agents that browse marketplaces, make product decisions and complete purchases for users.