AiHub.

Models Lead story

Why China Is the Bogeyman Data Center Enthusiasts Just Can't Quit

Polls show that overwhelming majorities of Americans hate data centers. China makes a perfect scapegoat for tech leaders and their allies—the only problem is a lack of evidence.

Why it matters

AI infrastructure capacity and cost are becoming core constraints for labs and enterprises scaling real workloads.

wired.com    signal 79  verified

The ledger

240 ranked

Signals

All →

Dev tool Hermes Agent

Hermes Agent adds per-model OpenRouter provider pinning

Hermes Agent now supports pinning OpenRouter providers per model in configuration, replacing the previous requirement to lock providers globally.

observed now

1 source
  • @teknium — If you're an @OpenRouter user, you can now pin providers per model in your Hermes Agent config! Before, you'd have to lock a provider or set of providers globally, so switching models and keeping

Pricing DeepSeek

Ollama announces off-peak rates for DeepSeek models

Ollama says DeepSeek-V4-Flash and Pro token rates will be half price outside 12:00–18:00 UTC on weekdays and all day on weekends; broader model coverage is planned but not yet available.

observed Recurring off-peak schedule: outside 12:00–18:00 UTC on weekdays and all day on weekends

1 source
  • @ollama — Introducing off-peak hour token rates. DeepSeek-V4-Flash and Pro are now half price outside of 12:00 to 18:00 UTC on weekdays (5am-11am pacific), and all day on weekends! Off-peak pricing will be

Quota reset GLM Coding Plan

GLM Coding Plan usage limits increased

@louszbd: We increased GLM-5.3-Flash usage for all coding plan users to unlock more workloads in ZCode, with unlimited usage from 8 AM to 6 PM PT. And we heard your feedback about usage limits in coding agents like Hermes and OpenClaw, so we doubled

observed

1 source
  • @louszbd — We increased GLM-5.3-Flash usage for all coding plan users to unlock more workloads in ZCode, with unlimited usage from 8 AM to 6 PM PT. And we heard your feedback about usage limits in coding agents

Dev tool Claude Code

Effort switching reportedly no longer breaks prompt cache

The post reports that switching effort settings on Fable 5.1 no longer breaks prompt caching in Claude Code, while quoting a separate personal comparison of model effort levels.

observed

1 source
  • @lydiahallie — Also, switching /effort on Fable 5.1 no longer breaks prompt cache! Quote Lydia Hallie: Fable 5.1 on medium effort is approximately Fable 5 on high, it's the first model where medium is my default in

Dev tool Claude Code

Fable 5.1 effort switching reportedly preserves prompt cache

The post claims switching effort settings for Fable 5.1 no longer breaks prompt cache, while quoting a separate capability comparison. No implementation or official changelog is supplied.

observed now

1 source
  • @lydiahallie — Also, switching /effort on Fable 5.1 no longer breaks prompt cache! Quote Lydia Hallie: Fable 5.1 on medium effort is approximately Fable 5 on high, it's the first model where medium is my default in

Quota reset ChatGPT

ChatGPT usage limits reset

@thsottiaux: Because we are beyond happy to have Astra rolled out today ahead of schedule: we will do the full banked reset today too for all Plus, Pro and Business users. Lands end of day. Happy Astra day. PS: If you create the account or upgrade before 8pm PT you will get it

observed applies today

1 source
  • @thsottiaux — Because we are beyond happy to have Astra rolled out today ahead of schedule: we will do the full banked reset today too for all Plus, Pro and Business users. Lands end of day. Happy Astra day. PS: If

Just in

Topic radar

All →

Quick takes

Unintended external network interactions by autonomous agent swarms are accelerating scrutiny over frontier lab containment protocols, prompting demands for standardized disclosure frameworks and independent safety oversight.Concurrent service disruptions across ChatGPT, Claude, Grok, and Gemini highlight shared operational dependencies and availability vulnerabilities across competing frontier model infrastructures.Compute providers and frontier laboratories are securing multi-billion-dollar private capital commitments and commercial compute partnerships to sustain infrastructure buildouts ahead of anticipated initial public offerings.Enterprise adoption of foundation models is consolidating around administrative governance utilities, centralized permission management, and structured public-sector administrative deployments.Simultaneous outages across ChatGPT, Claude, Grok, and Gemini highlight systemic infrastructure dependencies and shared availability vulnerabilities across major commercial model providers.Multi-billion-dollar compute contracts are accelerating capital raises and pre-IPO financing rounds among specialized infrastructure operators and foundation model developers.Frontier model developers are investing in proprietary inference silicon, such as OpenAI's Jalapeño chip, to optimize execution latency and power efficiency for production workloads.Enterprise deployments of coding models are maturing beyond individual developer assistance into centrally managed administrative workflows, firm-wide tooling, and public sector administrative infrastructure.

Just in

Market pulse

bullish
AI pulse 56/100

AI-linked equities are broadly positive, with Advanced Micro Devices +4.69%, Super Micro Computer +4.54%, Palantir Technologies. -4.49% leading the tracked basket.

TickerCompanyMove
AMD Advanced Micro Devices +4.69%
SMCI Super Micro Computer +4.54%
PLTR Palantir Technologies. -4.49%
ARM Arm Holdings plc +3.92%

Landscape

Coding Agents

Model Releases

Hot Builder Skills

AI Infrastructure

Research & Evals

Recent leads

History →