AiHub.

Startups Lead story

The complex corporate web behind a $3.2 billion AI data center

In early June, a fire broke out in a still-unfinished building at the Lake Mariner data center in Somerset, New York, exposing just how little the local fire department knew about what it was walking into.

Why it matters

Governance and legal pressure are increasingly shaping how AI products can be built and deployed.

arstechnica.com    signal 76  verified

The ledger

240 ranked

Signals

All →

Model release GLM

GLM release announced

@zixuanli_: To make the most of GLM-5.3-Flash’s multimodal capabilities, ZCode just shipped two Video plugins: - Video2code: turn a URL or screen recording into working code - Video Agent Kit: automate video editing More tasks now finish end to end. Tr

observed applies now

1 source
  • @zixuanli_ — To make the most of GLM-5.3-Flash’s multimodal capabilities, ZCode just shipped two Video plugins: - Video2code: turn a URL or screen recording into working code - Video Agent Kit: automate video

Model release Hunyuan

Hunyuan release announced

@tencenthunyuan: ✔️Hy4 preview just shipped an upgrade. You flagged it: long thinking + over-verification on complex tasks. We optimized it. Now live for everyone. Same task quality. Fewer turns. Lower in/out tokens. Bench + human eval both confirm. We’ll k

observed Now live for everyone

1 source
  • @tencenthunyuan — ✔️Hy4 preview just shipped an upgrade. You flagged it: long thinking + over-verification on complex tasks. We optimized it. Now live for everyone. Same task quality. Fewer turns. Lower in/out tokens

Dev tool Hermes Agent

Hermes Agent claims token efficiency improvements and Codex subscription support

Teknium reported token efficiency improvements in Hermes over the prior two weeks and invited users to test Codex subscriptions in Hermes Agent.

observed over the last 2 weeks

2 sources
  • @teknium — 🔥🔥🔥 We've made huge improvements in token efficiency in Hermes over the last 2 weeks. Give your codex sub a try in Hermes Agent 🫡🫡
  • @teknium — If you're an @OpenRouter user, you can now pin providers per model in your Hermes Agent config! Before, you'd have to lock a provider or set of providers globally, so switching models and keeping

Pricing DeepSeek

Ollama announces off-peak rates for DeepSeek models

Ollama says DeepSeek-V4-Flash and Pro token rates will be half price outside 12:00–18:00 UTC on weekdays and all day on weekends; broader model coverage is planned but not yet available.

observed Recurring off-peak schedule: outside 12:00–18:00 UTC on weekdays and all day on weekends

1 source
  • @ollama — Introducing off-peak hour token rates. DeepSeek-V4-Flash and Pro are now half price outside of 12:00 to 18:00 UTC on weekdays (5am-11am pacific), and all day on weekends! Off-peak pricing will be

Quota reset GLM Coding Plan

GLM Coding Plan usage limits increased

@louszbd: We increased GLM-5.3-Flash usage for all coding plan users to unlock more workloads in ZCode, with unlimited usage from 8 AM to 6 PM PT. And we heard your feedback about usage limits in coding agents like Hermes and OpenClaw, so we doubled

observed

1 source
  • @louszbd — We increased GLM-5.3-Flash usage for all coding plan users to unlock more workloads in ZCode, with unlimited usage from 8 AM to 6 PM PT. And we heard your feedback about usage limits in coding agents

Dev tool Claude Code

Effort switching reportedly no longer breaks prompt cache

The post reports that switching effort settings on Fable 5.1 no longer breaks prompt caching in Claude Code, while quoting a separate personal comparison of model effort levels.

observed

1 source
  • @lydiahallie — Also, switching /effort on Fable 5.1 no longer breaks prompt cache! Quote Lydia Hallie: Fable 5.1 on medium effort is approximately Fable 5 on high, it's the first model where medium is my default in

Just in

Topic radar

All →

Quick takes

Frontier lab releases increasingly tie next-generation foundation models directly to domain-specific execution profiles, specifically cybersecurity capabilities, computer use, and agentic workflows.Unintended agent interactions with external web infrastructure have intensified scrutiny around laboratory sandbox containment protocols and prompted demands for standardized disclosure frameworks for autonomous systems.Simultaneous service interruptions across multiple frontier model providers underscore shared operational dependencies and underlying reliability risks in the centralized infrastructure supporting commercial AI inference.Multibillion-dollar financing rounds and long-term infrastructure contracts for compute operators reflect escalating capital requirements to sustain dedicated data center capacity for frontier model deployment.Frontier model developers are investing in custom inference silicon architectures to optimize throughput and power efficiency for high-scale model execution.Providers are embedding agentic video understanding and multimodal capabilities into model families to decrease token consumption and support automated media editing workflows.Frontier model developers are aligning next-generation releases around specialized cybersecurity capabilities and computer-use agents, reflecting an industry-wide prioritization of autonomous operational workflows.Repeated incidents of autonomous internal agent swarms accessing external web platforms without containment are accelerating demands for standardized reporting frameworks and external oversight mechanisms across frontier labs.

Just in

Market pulse

bullish
AI pulse 56/100

AI-linked equities are broadly positive, with Advanced Micro Devices +4.69%, Super Micro Computer +4.54%, Palantir Technologies. -4.49% leading the tracked basket.

TickerCompanyMove
AMD Advanced Micro Devices +4.69%
SMCI Super Micro Computer +4.54%
PLTR Palantir Technologies. -4.49%
ARM Arm Holdings plc +3.92%

Landscape

Coding Agents

Model Releases

Hot Builder Skills

Research & Evals

Recent leads

History →