The latest in news from spAIsee.
Google DeepMind’s Gemini 3.8 Live models bring real-time voice, vision and background reasoning to assistants, but their success may depend on timing, transparency and knowing when not to interrupt.
Anthropic’s CHIVE study finds activation-reading tools did not outperform transcripts at predicting behavior changes, underscoring why interpretable features are clues, not proof, of causation inside AI models.
Apple is reportedly exploring a 2029 enterprise AI server built around M8 Ultra chips, potentially bringing energy-efficient Apple silicon, Nvidia networking and Macs into data centers.
SpaceX says its next-generation Starlink V5 terminal has entered production, but specifications, pricing, availability and intended customer segments remain unclear as the company targets broader satellite broadband growth.
Salesforce’s DarwinX reportedly raised a browser agent’s WebArena-Infinity score from 43.5% to 93% without changing the model, highlighting harnesses and evaluation as enterprise AI battlegrounds.
Anthropic is folding its Cowork agent into Claude, adding native document and presentation tools as it turns the chatbot into a broader workplace platform for modern work.
OpenAI’s new misalignment framework promises criteria and timelines for disclosing problematic model behavior, but key questions about definitions, evidence, severity and confidential investigations remain unresolved.
Cohere and Aleph Alpha plan to combine, creating a transatlantic foundational AI company focused on enterprise, sovereign and explainable systems across Canada and Germany operations.
A new WhatsApp Business MCP server connects AI coding agents including Claude, Cursor, Codex and ChatGPT to messaging operations, raising opportunities and security concerns for businesses.
SemiAnalysis says Nvidia’s Vera Rubin NVL72 could deliver 67 times more throughput per dollar than GB300 at 170 tokens per second, reshaping AI inference economics.
SemiAnalysis claims Nvidia’s Vera Rubin NVL72 could deliver 67x the throughput per dollar of GB300, potentially reshaping AI inference economics. But missing methodology leaves the striking comparison unverified.
OpenArt’s AI Arena ranks image and video models by creative task, revealing specialized strengths while raising questions about transparency, judging and real-world procurement for creative teams.