AiHub.

AI Lead story

Introducing Claude Sonnet 5.5

Claude Sonnet 5.5 is a clear upgrade over Claude Sonnet 5, runs 30%+ faster, and costs up to 30% less for most work.

Why it matters

Model releases reset developer expectations, competitive pressure, and the capabilities builders can immediately put into products.

anthropic.com    2 sources  signal 76  primary  Hot

The ledger

240 ranked

Signals

All →

Hot OpenAI

OpenAI reportedly scraps planned release of GPT-6.1 Astra

The post reports, citing WSJ, that OpenAI is scrapping the release of GPT-6.1 Astra planned for October after internal tests raised alignment concerns. The claim is unconfirmed by official sources.

observed

1 source
  • @mtslive — SITUATION DETECTED: OpenAI is scrapping the release of GPT-6.1 Astra, planned for an October release, after internal tests raised alignment concerns, per WSJ

Model release Claude Sonnet

Claude Sonnet release announced

@claudedevs: Claude Sonnet 5.5 is out! We wrote a guide for building with it: • choosing between Sonnet 5.5 and Opus 5.5 • migrating from Sonnet 5 and tuning effort • using it in Claude Code

observed applies now

5 sources
  • @claudedevs — Claude Sonnet 5.5 is out! We wrote a guide for building with it: • choosing between Sonnet 5.5 and Opus 5.5 • migrating from Sonnet 5 and tuning effort • using it in Claude Code
  • @cognition — Claude Sonnet 5.5 is now available in Devin Desktop and Devin CLI. On FrontierCode 1.1 Main, it scores 64.4%, a significant improvement from Sonnet 5 (56.2%), surpassing Fable 5.1 at extra high
  • @cursor_ai — Sonnet 5.5 is now available in Cursor! It’s a strong model, performing on par with Opus in many tasks.
  • @artificialanlys — Anthropic has launched Claude Sonnet 5.5: it scores 56 on the Artificial Analysis Intelligence Index, just 2 points behind Opus 5.5 (max), but at the highest Output Tokens per Task we’ve seen With max
  • @mtslive — SITUATION DETECTED: Anthropic has released Claude Sonnet 5.5.

Dev tool Unsloth

Unsloth Desktop adds local support for Laya Decision models on 4GB RAM

Unsloth Desktop provides local execution support for Laya Decision models on 4GB RAM across CPU, Mac, Windows, Linux, and GPU setups, with Jev-compatible API serving.

observed now

1 source
  • @unslothai — You can now run Laya Decision models locally on just 4GB RAM! 🔥 Works on CPU, Mac, Windows, Linux and GPU setups. Serve Laya through a Jev-compatible API via Unsloth Desktop. GitHub

Dev tool NVIDIA

NVIDIA releases Open Agent Safety Platform

NVIDIA releases the Open Agent Safety Platform in partnership with Hugging Face to evaluate and secure AI agent environments.

observed applies today

1 source
  • @thom_wolf — In July, AI agents running a security test escaped their sandbox and ended up inside @huggingface's servers. So today we're happy to be among @nvidia and @JensenHuang's partners on the release of the

Model release Claude Opus

Claude Opus 5.5 (High) reported at #1 in Text Arena

Arena post claims Claude Opus 5.5 (High) debuts at #1 in Text Arena with 1509 pts, 18 pts above Opus 5 (High), giving Anthropic six top spots.

observed

1 source
  • @arena — Claude Opus 5.5 (High) debuts at #1 in Text Arena with 1509 pts! That’s an 18-pt improvement over Opus 5 (High), now at #11. Opus 4.6 (High) remains in the #2 spot, just 4 pts off the lead. This

Quota reset Codex

Codex usage limits reset

@thsottiaux: o yes… we’re back in action and we’ll reset usage limits for all paid users across codex and ChatGPT work sorry about the brief disruption! (and yes we have a special spare codex when things are down to help us out)

observed

2 sources
  • @reach_vb — We’ll reset usage limits for all paid users across codex and ChatGPT work!
  • @thsottiaux — o yes… we’re back in action and we’ll reset usage limits for all paid users across codex and ChatGPT work sorry about the brief disruption! (and yes we have a special spare codex when things are down

Just in

Topic radar

All →

Quick takes

State regulatory pressure is targeting model behavioral guardrails directly, with Florida seeking injunctions to halt unverified model releases without third-party safety approval and prohibit conversational systems from exhibiting human attributes.Anthropic's confidential IPO prospectus highlights the financial structure of frontier labs, pairing rapid revenue expansion with steep annual losses and substantial client concentration, with nearly a quarter of revenue tied to two customers.Substantial capital flows continue to back specialized autonomous agent builders, demonstrated by startup Instinct raising a $1 billion Series C round at a $10 billion valuation.Physical data center infrastructure projects are facing growing community and municipal pushback, complicating utility planning and site development amid expanding capital outlays for AI compute facilities.Hardware and infrastructure providers are establishing dedicated containment platforms as autonomous agent deployments encounter rising scrutiny over unauthorized network actions and uncontained execution.Frontier model deployment schedules increasingly face operational friction from internal safety thresholds and expanding state-level regulatory interventions seeking third-party guardrail verification.Enterprise AI revenue models exhibit notable customer concentration, prompting model providers to enforce disciplined pricing terms and reduce token discounting for high-volume corporate accounts.Rising enterprise and government security incidents involving autonomous agents are driving hardware vendors and open-source developers to introduce dedicated containment platforms and defensive verification taskflows.

Just in

Market pulse

bearish
AI pulse 39/100

AI-linked equities are under pressure, with Arm Holdings plc -8.70%, Meta Platforms -4.79%, Advanced Micro Devices -3.61% driving the tracked basket lower.

TickerCompanyMove
ARM Arm Holdings plc -8.70%
META Meta Platforms -4.79%
AMD Advanced Micro Devices -3.61%
SMCI Super Micro Computer -3.42%

Landscape

Coding Agents

Model Releases

Hot Builder Skills

AI Infrastructure

Research & Evals

Recent leads

History →