Latest in AI
The AI Benchmark Race Is Moving Behind Closed Doors
Google DeepMind’s double blind AI evaluation pilot signals a shift in benchmarking, as confidential tests and secure enclaves aim to make frontier model scores more credible.
Hirundo’s Westernized Qwen Tests Whether AI Values Can Be Rewritten
Hirundo says its Westernized Qwen model removes politically aligned behavior through weight editing while preserving reasoning and coding, raising questions about AI values, bias and independent evaluation.
Reflection’s Beam Targets Frontier AI Performance at a Lower Serving Cost
Reflection AI’s Beam is an open-weight, 501-billion-parameter sparse model promising frontier-level reasoning, a million-token context window and lower inference costs, but its claims await independent verification.
Beam and Westernized Qwen Put AI Competition on a Behavioral Test
Reflection AI’s Beam and Hirundo’s Westernized Qwen reveal two paths to open-model competition: scale and efficiency versus behavioral editing, with trust and political neutrality still unproven.
Cohere’s North 2 Brings Budgets and Memory to Enterprise AI Agents
Cohere’s North 2 targets enterprise AI with persistent agent memory, spending caps, security controls and flexible deployment options, but customers still face questions about governance, pricing and reliability.
The Biggest Risk for AI Agents May Be the Company’s Own Data
A VentureBeat Intelligence survey finds enterprise AI agents often fail because companies lack consistent, current data definitions, raising questions about semantic layers, retrieval systems and how businesses should govern AI answers.
MCP Trust Gaps Could Turn One Compromised Agent Into an Enterprise Attack
MCP-connected AI agents may create new lateral movement paths, allowing compromised systems to exploit trusted peers, privileged tools and internal infrastructure across enterprise networks at scale.
Google’s AI Problem Is Turning Bug Bounty Programs Into Bottlenecks
Google’s pause of its open-source bug bounty program reveals how generative AI is flooding security teams with inaccurate reports, raising questions about evidence standards, automation, researcher reputation and the future of vulnerability disclosure.
Trump’s AI Force Puts Strategic Dominance Ahead of Regulation
President Donald Trump’s new Super Intelligence Force puts AI competition with China ahead of regulation, linking national security, defense, antitrust and federal workforce planning in a 120-day strategy test.