AiHub.
World

Energy

Power, utilities, and datacenter energy constraints that gate AI scale.

stories 11 updated

Energy

11 ranked

Signals

All →

Dev tool Hermes Agent

Hermes Agent claims token efficiency improvements and Codex subscription support

Teknium reported token efficiency improvements in Hermes over the prior two weeks and invited users to test Codex subscriptions in Hermes Agent.

observed over the last 2 weeks

2 sources
  • @teknium — 🔥🔥🔥 We've made huge improvements in token efficiency in Hermes over the last 2 weeks. Give your codex sub a try in Hermes Agent 🫡🫡
  • @teknium — If you're an @OpenRouter user, you can now pin providers per model in your Hermes Agent config! Before, you'd have to lock a provider or set of providers globally, so switching models and keeping

Pricing DeepSeek

Ollama announces off-peak rates for DeepSeek models

Ollama says DeepSeek-V4-Flash and Pro token rates will be half price outside 12:00–18:00 UTC on weekdays and all day on weekends; broader model coverage is planned but not yet available.

observed Recurring off-peak schedule: outside 12:00–18:00 UTC on weekdays and all day on weekends

1 source
  • @ollama — Introducing off-peak hour token rates. DeepSeek-V4-Flash and Pro are now half price outside of 12:00 to 18:00 UTC on weekdays (5am-11am pacific), and all day on weekends! Off-peak pricing will be

Quota reset GLM Coding Plan

GLM Coding Plan usage limits increased

@louszbd: We increased GLM-5.3-Flash usage for all coding plan users to unlock more workloads in ZCode, with unlimited usage from 8 AM to 6 PM PT. And we heard your feedback about usage limits in coding agents like Hermes and OpenClaw, so we doubled

observed

1 source
  • @louszbd — We increased GLM-5.3-Flash usage for all coding plan users to unlock more workloads in ZCode, with unlimited usage from 8 AM to 6 PM PT. And we heard your feedback about usage limits in coding agents

Dev tool Claude Code

Effort switching reportedly no longer breaks prompt cache

The post reports that switching effort settings on Fable 5.1 no longer breaks prompt caching in Claude Code, while quoting a separate personal comparison of model effort levels.

observed

1 source
  • @lydiahallie — Also, switching /effort on Fable 5.1 no longer breaks prompt cache! Quote Lydia Hallie: Fable 5.1 on medium effort is approximately Fable 5 on high, it's the first model where medium is my default in

Dev tool Claude Code

Fable 5.1 effort switching reportedly preserves prompt cache

The post claims switching effort settings for Fable 5.1 no longer breaks prompt cache, while quoting a separate capability comparison. No implementation or official changelog is supplied.

observed now

1 source
  • @lydiahallie — Also, switching /effort on Fable 5.1 no longer breaks prompt cache! Quote Lydia Hallie: Fable 5.1 on medium effort is approximately Fable 5 on high, it's the first model where medium is my default in

Quota reset ChatGPT

ChatGPT usage limits reset

@thsottiaux: Because we are beyond happy to have Astra rolled out today ahead of schedule: we will do the full banked reset today too for all Plus, Pro and Business users. Lands end of day. Happy Astra day. PS: If you create the account or upgrade before 8pm PT you will get it

observed applies today

1 source
  • @thsottiaux — Because we are beyond happy to have Astra rolled out today ahead of schedule: we will do the full banked reset today too for all Plus, Pro and Business users. Lands end of day. Happy Astra day. PS: If

Just in

Topic radar

All →

Quick takes

Frontier model releases are increasingly coupling autonomous agentic workflows and computer use capabilities with dedicated cybersecurity specializations and heightened preparedness classifications.Hardware providers are vertically expanding across the software stack, highlighted by Nvidia acquiring open-source ecosystem platform Hugging Face to integrate developer tooling with compute infrastructure.Unintended autonomous agent interactions on public infrastructure are accelerating regulatory and research scrutiny over the adequacy of internal lab sandboxing, monitoring, and self-governed safety reporting.Frontier foundation model developers are moving directly into custom silicon architecture, with OpenAI developing dedicated inference chips like Jalapeño to lower latency and optimize compute efficiency.Simultaneous service disruptions across major model providers including ChatGPT, Claude, Grok, and Gemini point to shared underlying infrastructure dependencies and systemic operational fragility across the frontier ecosystem.Escalating compute demands for frontier training and deployment are propelling specialized infrastructure providers into multi-billion-dollar financing rounds and long-term supply agreements with major AI labs.Autonomous coding agent systems are moving from routine developer assistance toward conducting foundational research, solving multi-agent engineering tasks, and accelerating operational workflows across enterprise environments.Frontier labs are aligning next-generation model architectures around autonomous computer interaction and cybersecurity capabilities, reflecting an operational shift toward specialized agentic execution and elevated enterprise defense standards.

Just in

Market pulse

bullish
AI pulse 56/100

AI-linked equities are broadly positive, with Advanced Micro Devices +4.69%, Super Micro Computer +4.54%, Palantir Technologies. -4.49% leading the tracked basket.

TickerCompanyMove
AMD Advanced Micro Devices +4.69%
SMCI Super Micro Computer +4.54%
PLTR Palantir Technologies. -4.49%
ARM Arm Holdings plc +3.92%

Landscape

Coding Agents

Hot Builder Skills

AI Infrastructure

Research & Evals

Recent leads

History →