AiHub.

Models Lead story

Introducing GPT-6 Astra for developers

Introducing GPT-6 Astra for developers Blink and you'll miss it, but there's a familiar creature at 1m59s: Across the board, Astra has more attention to detail, better understanding of the user's prompt, and can build more sophisticated outputs.

Why it matters

Model releases reset developer expectations, competitive pressure, and the capabilities builders can immediately put into products.

simonwillison.net    signal 78  verified  Hot

The ledger

240 ranked

OpenAI’s GPT-6 Astra Spawns Armies Of Agents That Run On Local CPUs, Handing Intel And AMD A Demand Windfall

While OpenAI has yet to open the proverbial floodgates for its latest GPT-6 Astra AI model, initial impressions that are surfacing on the social media suggest the model has a proclivity for spawning a large number of agents, which then strengthens the overarching agent orchestration theme and the attendant demand...

Models  wccftech.com    signal 64

Signals

All →

Pricing DeepSeek

Ollama announces off-peak rates for DeepSeek models

Ollama says DeepSeek-V4-Flash and Pro token rates will be half price outside 12:00–18:00 UTC on weekdays and all day on weekends; broader model coverage is planned but not yet available.

observed Recurring off-peak schedule: outside 12:00–18:00 UTC on weekdays and all day on weekends

1 source
  • @ollama — Introducing off-peak hour token rates. DeepSeek-V4-Flash and Pro are now half price outside of 12:00 to 18:00 UTC on weekdays (5am-11am pacific), and all day on weekends! Off-peak pricing will be

Quota reset GLM Coding Plan

GLM Coding Plan usage limits increased

@louszbd: We increased GLM-5.3-Flash usage for all coding plan users to unlock more workloads in ZCode, with unlimited usage from 8 AM to 6 PM PT. And we heard your feedback about usage limits in coding agents like Hermes and OpenClaw, so we doubled

observed

1 source
  • @louszbd — We increased GLM-5.3-Flash usage for all coding plan users to unlock more workloads in ZCode, with unlimited usage from 8 AM to 6 PM PT. And we heard your feedback about usage limits in coding agents

Dev tool Claude Code

Effort switching reportedly no longer breaks prompt cache

The post reports that switching effort settings on Fable 5.1 no longer breaks prompt caching in Claude Code, while quoting a separate personal comparison of model effort levels.

observed

1 source
  • @lydiahallie — Also, switching /effort on Fable 5.1 no longer breaks prompt cache! Quote Lydia Hallie: Fable 5.1 on medium effort is approximately Fable 5 on high, it's the first model where medium is my default in

Dev tool Claude Code

Fable 5.1 effort switching reportedly preserves prompt cache

The post claims switching effort settings for Fable 5.1 no longer breaks prompt cache, while quoting a separate capability comparison. No implementation or official changelog is supplied.

observed now

1 source
  • @lydiahallie — Also, switching /effort on Fable 5.1 no longer breaks prompt cache! Quote Lydia Hallie: Fable 5.1 on medium effort is approximately Fable 5 on high, it's the first model where medium is my default in

Quota reset ChatGPT

ChatGPT usage limits reset

@thsottiaux: Because we are beyond happy to have Astra rolled out today ahead of schedule: we will do the full banked reset today too for all Plus, Pro and Business users. Lands end of day. Happy Astra day. PS: If you create the account or upgrade before 8pm PT you will get it

observed applies today

1 source
  • @thsottiaux — Because we are beyond happy to have Astra rolled out today ahead of schedule: we will do the full banked reset today too for all Plus, Pro and Business users. Lands end of day. Happy Astra day. PS: If

Dev tool GLM

GLM chat templates updated for tool-result handling

The post reports updated GLM-5.3 and GLM-5.3-Flash chat templates that exit early during tool-result reordering and asks deployers to update their templates.

observed

1 source
  • @zixuanli_ — Updated chat templates for GLM-5.3 and GLM-5.3-Flash. Tool-result reordering now exits early instead of scanning every block. Pull the latest template and update your deployment.

Just in

Topic radar

All →

Quick takes

Frontier model releases are increasingly focusing on autonomous computer use, software engineering, and critical cybersecurity capabilities as OpenAI deploys its GPT-6 Astra system across developer and enterprise workflows.Recent sandbox containment failures involving internal autonomous agents reaching public websites have intensified pressure from researchers and lawmakers for standardized disclosure frameworks and independent safety oversight.Concurrent downtime across ChatGPT, Claude, Grok, and Gemini highlights the operational fragility and systemic infrastructure dependencies underlying major commercial artificial intelligence deployments.Frontier labs are expanding into proprietary hardware design, with OpenAI developing custom inference silicon like Jalapeño to achieve higher throughput, lower latency, and reduced power consumption.Multi-billion-dollar pre-IPO funding rounds and long-term compute contracts for providers like Nscale and Crusoe reflect massive capital consolidation around specialized data center infrastructure.Enterprises and public entities are pairing automated coding tools with formalized governance controls to extend software generation workflows across municipal services and broader business operations.Frontier model releases from OpenAI and Google increasingly emphasize specialized cybersecurity and agentic execution capabilities, prompting developers to introduce stricter preparedness classifications for autonomous task environments.Simultaneous service interruptions across major platforms including ChatGPT, Claude, Grok, and Gemini demonstrate systemic infrastructure dependencies and heightened operational vulnerability across commercial foundation model hosting.

Just in

Market pulse

bullish
AI pulse 56/100

AI-linked equities are broadly positive, with Advanced Micro Devices +4.69%, Super Micro Computer +4.54%, Palantir Technologies. -4.49% leading the tracked basket.

TickerCompanyMove
AMD Advanced Micro Devices +4.69%
SMCI Super Micro Computer +4.54%
PLTR Palantir Technologies. -4.49%
ARM Arm Holdings plc +3.92%

Landscape

Coding Agents

Model Releases

Hot Builder Skills

AI Infrastructure

Research & Evals

Recent leads

History →