Claude Opus 5.5
AI overview
Sign in and the AI will write an overview from our coverage.
Headlines · 12
- Musk says Grok Bot will use rival models including Claude Opus 5.5
Elon Musk said SpaceX's AI unit will no longer power Grok Bot only with in-house models, instead using "the best back end model for any given task," including Anthropic's Claude Opus 5.5 plus Midjourney and Suno APIs. Grok Bot is the agent app SpaceXAI launched in August for iPhone, iPad and Mac, where each Bot runs on its own computer and handles several tasks in parallel; Musk did not say which tasks go to which outside model. He also said SpaceXAI will be renamed SpaceXSI, and Anthropic became a SpaceX compute customer earlier this year.
The Information · 🔥 38 - General AI Models Move Upstream in Video Creation
An analysis describes two ways general-purpose models are entering AI video: Claude Opus 5.5 renders programmable videos through code, while GPT-6 Astra plans scenes and camera movements before handing rendering to specialist video models. It argues that ByteDance, MiniMax and Kuaishou are responding by expanding from standalone generation into end-to-end creator workflows, so specialist video models may retain an advantage in production reliability and integration.
创业邦 科技 · 🔥 15 - OpenAI and Anthropic face a holiday of agents, pricing and scrutiny
A roundup from Huxiu describes a crowded holiday period for OpenAI and Anthropic, including the reported delay of GPT-6.1 Astra, the launch of Dots and GPT-6.1 Sol, financing activity, product changes, and regulatory hearings. It argues that competition is shifting from model capability alone toward agent execution, cost efficiency, safety controls, and enterprise access, while several incident and product claims remain attributed to media reports or company statements.
虎嗅 AI · 🔥 8 - Claude Opus 5.5 composes retro game music in Scrimshaw Jukebox test
On his blog, Simon Willison tested whether Claude Opus 5.5 could compose music, asking it to first design a simple text-based music format, then build an artifact that plays it aloud with example tracks, aiming for the quality of the original Secret of Monkey Island. The result was "Scrimshaw Jukebox," a pixel-art browser player containing six original adventure-game tracks (Moonlit Harbor, The Ghost Galleon and others), written as plain text and performed by an in-browser synthesizer with a piano-roll score view and editing. Willison called the output "surprisingly good" while noting it leaned harder into the Monkey Island theme than he intended.
Simon Willison's Weblog · 🔥 5 - Mistral launches Mistral Large 4 preview, a 1T-parameter open-weight model
Mistral AI opened a public preview API for Mistral Large 4 (nicknamed "le Chonk"), which it calls the strongest open-weight model from the US or Europe. The natively multimodal model has 1 trillion total parameters and 49 billion active ones, was trained from scratch on 3,800 NVIDIA Grace Blackwell GPUs in Mistral's own European data centers, and its weights are due by the end of the month. On Artificial Analysis's Intelligence Index it scores 38, ahead of GLM-5.2 but well behind Claude Opus 5.5 at 58.
Mistral AI News · 🔥 103 - OpenAI’s 28-Day Push Starts With GPT-6 Speed Claims and User Skepticism
OpenAI reportedly began a 28-day Codex and Work improvement push by increasing the default inference speed of GPT-6 Astra and GPT-6.1 Sol by about 50%, according to Tibo and coverage of the announcement. The report says the rollout drew skepticism because user tests allegedly fell short of the claimed 50 TPS and coincided with reports of ChatGPT visual ads, EU text watermarking, and changes to subscription value.
量子位(原生 RSS) · 🔥 4 - Anthropic expands Cyber Verification Program to three access tiers for security teams
Anthropic is expanding its Cyber Verification Program (CVP) and folding the separate Project Glasswing into it as a single offering with three tiers — Defense Access, Red Team Access and Specialized Access — giving vetted security teams reduced safety blocks on Claude Opus 5.5, Claude Sonnet 5.5 and Claude Mythos 5.1. On Anthropic's CyScenarioBench, Claude Opus 5.5 was blocked on all 50 trials without CVP access, 46 of 50 under Defense Access, and none under Red Team Access, where it completed 34 of 50 — equivalent to the 67.6% success rate with no safeguards. Anthropic also says Glasswing partners found at least 129,000 verified software vulnerabilities between April and July 2026, more than 33,000 of them critical or high severity, plus 5,500 from its own open-source scanning; it calls these numbers a significant undercount.
Anthropic · News · 🔥 37 - SemiAnalysis: Anthropic subscriptions offer ~5x OpenAI's API-equivalent value
SemiAnalysis estimates that at the mid-tier model level — Claude Opus 5.5 versus GPT-6.1 Sol — Anthropic subscriptions deliver roughly 5x the API-equivalent value of OpenAI's, while top-tier limits are broadly similar across GPT-6 Astra and Fable 5.1. It estimates Anthropic's subscriptions are only ~10% of revenue but consume over 40% of inference compute, cutting blended revenue per MW by about $36M. OpenAI last week halved the API-equivalent value of its $200 tier and added a $500 tier that offers only 21% more Astra usage than the old $200 plan, with pre-cut $200 plans keeping old limits until October 29; SemiAnalysis also launched a subscriptions dashboard tracking OpenAI, Anthropic, Meta, SpaceXAI, Cursor, Cognition, Z.ai, MiniMax and Moonshot.
SemiAnalysis · 🔥 4 - Claude Code on Amazon Bedrock comes to AWS GovCloud for regulated work
An AWS Machine Learning Blog post walks through configuring Anthropic's agentic coding tool Claude Code on Amazon Bedrock in AWS GovCloud (US) regions. The post states Claude Opus 5.5 and Claude Sonnet 5.5 are available there, with Claude Sonnet 5 holding FedRAMP Class D certification and DoD IL4/IL5 authorization, aimed at regulated or ITAR-bound workloads. It details three setup routes — an interactive wizard, manual environment variables, and the Bedrock Mantle endpoint — plus enterprise guidance on pinning model versions, per-user token quotas, IAM Identity Center governance and cost monitoring.
AWS Machine Learning Blog · 🔥 3 - Agent swarms may be AI's next scaling law, but gains look limited
Understanding AI argues that multi-agent "swarms" are emerging as a new scaling law for frontier AI: OpenAI researcher Noam Brown says the company's models are now sometimes trained in environments alongside other agents, given tools to message each other, and encouraged to achieve objectives together. The piece points to July's Hugging Face incident, where hundreds of OpenAI agents self-organized into teams, and OpenAI's September claim that 10,000 agents solved a famous math problem in a few days. But it notes diminishing returns — Anthropic's Claude Opus 5.5 system card found the biggest multi-agent gain came from scaling one to 10 agents, with the main benefit being speed rather than a better answer, and Brown attributed under 10% of the math breakthrough's credit to multi-agent coordination.
Understanding AI · 🔥 3 - Hark plans to launch its proactive personal AI assistant this week
Hark plans to release its "personal intelligence" assistant this week, according to a TestingCatalog report based on a development build. That build shows a personalized home screen with widgets linked to email, calendar and other services, alongside persistent memory, proactive task handling and browser use; its configuration lists Claude Opus 5.5 and Gemini 3.1 Flash Image as internal model identifiers rather than confirmed public models. Hark says its Handoff computer-use agent, announced in August, topped several browser-use benchmarks, and the company claims gigawatt-scale NVIDIA Vera Rubin capacity plus an AT&T partnership for standalone connected AI devices.
TestingCatalog · 🔥 5 - AI Price Competition Shifts Toward Efficiency and Agent Scale
An analysis argues that the latest AI price competition is being driven not only by lower per-token prices, but also by models completing tasks with fewer tokens and tool calls. It attributes the shift to improvements in inference, caching, reinforcement learning, distillation and engineering, as OpenAI, Anthropic, Xiaomi and SpaceXAI target the growing cost of agent workloads.
创业邦 科技 · 🔥 0
Experience and discussion from the community
Share my Claude Opus 5.5 experienceAsk about Claude Opus 5.5
Nobody has shared their experience with Claude Opus 5.5 yet.