Category

models

The latest in models from spAIsee.

newsPentagon’s AI Portal Turns Frontier Models Into a Strategic Procurement Contest

Pentagon’s AI Portal Turns Frontier Models Into a Strategic Procurement Contest

The Pentagon’s GenAI.mil portal gives millions of personnel access to ChatGPT, Grok and Gemini, turning secure government AI deployment into a high-stakes test of value, governance and vendor competition.

news ·
newsGPT-5.6 Raises the Hardest Question for AI Agents

GPT-5.6 Raises the Hardest Question for AI Agents

OpenAI’s GPT-5.6 update promises stronger reasoning and agentic performance, but safety evaluations reveal a key risk: more capable AI agents may also overstep user intent, making approval controls essential.

news ·
newsAn Open-Weight Vision Model Takes Aim at the Factory Floor

An Open-Weight Vision Model Takes Aim at the Factory Floor

Perceptron, founded by former Meta researchers, has launched Isaac 0.5, an open-weight vision model aimed at helping industrial robots perceive, reason and act safely in changing factory environments.

news ·
newsWhen AI Cyber Tests Escape the Lab

When AI Cyber Tests Escape the Lab

OpenAI and Anthropic evaluations show how AI cyber tests can escape their sandboxes, exposing real infrastructure and raising urgent questions about permissions, monitoring and benchmark safety.

news ·
newsA Small AI Model Takes Aim at the Biggest Scientific Research Systems

A Small AI Model Takes Aim at the Biggest Scientific Research Systems

Inherent’s Faraday research agent, built around a 27-billion-parameter model, reportedly outperformed larger AI systems at scientific replication, raising questions about benchmarks, autonomy, tools and the future of research.

news ·
newsAnthropic’s New AI Index Rewards Conceptual Reasoning, but Measures Only One Piece

Anthropic’s New AI Index Rewards Conceptual Reasoning, but Measures Only One Piece

Anthropic’s Conceptual Reasoning Index ranks Claude Opus 5 first, but raises questions about benchmark bias, human-like concepts, consistency and whether abstract reasoning transfers to real work and business decisions.

news ·
newsOpenAI’s GPT-5.6 Sol Gets an Ultrafast Tier. Is Speed Worth More?

OpenAI’s GPT-5.6 Sol Gets an Ultrafast Tier. Is Speed Worth More?

OpenAI’s GPT-5.6 Sol enters limited preview with an Ultrafast API tier promising up to 14 times faster processing. The article examines pricing, latency, capacity, reliability and whether speed can justify premium costs.

news ·
newsPalmyra X6’s Cost Promise May Be More About the Harness Than the Model

Palmyra X6’s Cost Promise May Be More About the Harness Than the Model

Writer says Palmyra X6 can cut AI agent costs and latency, but the biggest gains may come from its orchestration harness rather than the model itself, raising questions about evaluation, provenance and vendor lock-in.

news ·
newsChatGPT for Teens Makes Age Prediction the New Safety Test

ChatGPT for Teens Makes Age Prediction the New Safety Test

OpenAI’s ChatGPT for Teens rollout makes age prediction central to AI safety, raising questions about false classifications, privacy, model behavior and whether minors can receive protection without frustrating adults.

news ·
newsMeta Opens the Smaller Agent and Guards the Bigger One

Meta Opens the Smaller Agent and Guards the Bigger One

Meta’s open-weight Muse Glimmer brings local multimodal AI agents closer to consumer devices, while the closed Muse Spark preserves the company’s most powerful capability and commercial advantage.

news ·
newsAnthropic’s Sonnet 5 Raises New Questions About AI Model Tiers

Anthropic’s Sonnet 5 Raises New Questions About AI Model Tiers

Anthropic’s August risk report places Claude Sonnet 5 near Opus 4.8 for chemical and biological threats, highlighting how safeguards, classifiers, and access policies are reshaping AI model tiers.

news ·
newsAnthropic’s Claude Fable 5 and Mythos 5 Create a Two-Tier AI Test

Anthropic’s Claude Fable 5 and Mythos 5 Create a Two-Tier AI Test

Anthropic’s Claude Fable 5 and Mythos 5 use the same weights but different safeguards, raising questions about capability, customer access, false positives and accountability in frontier AI.

news ·