Live telemetry · daily brief
GAWK
by nativerse
● ARCHIVE
2026-05-08
2 tool incidents
What verifiably moved in the AI ecosystem in the last 24h. Every number traces to a public source.
What moved
- ■claude-opus-4-6-thinking holds #1 on LMArena for the 17th consecutive snapshot.
- ■llamaindex downloads declined for the 13th consecutive snapshot.
- ■openai downloads declined for the 5th consecutive snapshot.
SDK adoption6 notable shifts across the six registries
Package install volume is where developers place real bets — distinct from where labs are shipping marketing. pypistats.org, npmjs.com, crates.io · view all →
+298.7k
Docker Hub: ollama/ollama
+298.7k all-time pulls day-over-day
+25.6k
VS Code Marketplace: GitHub.copilot
+25.6k cumulative installs day-over-day
Model Usagetencent/hy3-preview holds #1 · biggest mover: deepseek/deepseek-v4-pro +19 to #10 · biggest drop: anthropic/claude-sonnet-4.5 -13
OpenRouter rankings reflect API-first developer spend; direct customers like consumer ChatGPT are invisible by construction. openrouter.ai · view all →
+19
deepseek/deepseek-v4-pro
was #29 on 2026-05-01
OpenRouter request volume reflects developer API spending, not end-user adoption. Biased toward API-first workflows that route through OpenRouter — direct OpenAI / Anthropic / Google customers who never use OpenRouter are invisible.
-13
anthropic/claude-sonnet-4.5
was #28 on 2026-05-01
—
#1 tencent/hy3-preview
#1 tencent/hy3-preview
—
#2 moonshotai/kimi-k2.6
#2 moonshotai/kimi-k2.6
—
#3 anthropic/claude-sonnet-4.6
#3 anthropic/claude-sonnet-4.6
Tool Health
2 incidents in the last 24h (vs 3 yesterday)
Why this matters · Provider outages and degradations cause retry storms upstream. Tracking the 7-day shape catches flapping providers before they page you.
Elevated errors on Claude Opus 4.7
started 12:10 UTC · resolved 12:45 UTC · minor impact
Elevated transcription failures affecting ChatGPT & Codex
started 17:15 UTC · resolved 08:31 UTC · minor impact
claude-code: Degraded → Operational
openai-api: Operational → Degraded
codex: Operational → Degraded
Benchmark movers
No rank changes in the LMArena top 3
Why this matters · Public benchmarks are gameable, but rank shuffles still hint at where the frontier is genuinely moving.
#1 claude-opus-4-6-thinking
anthropic · 1500 Elo
#2 claude-opus-4-6
anthropic · 1496 Elo
#3 gemini-3.1-pro-preview
google · 1487 Elo
Top HN stories
Top 5 on HN in the last 24h
Why this matters · What developers debate on Hacker News often previews which models, frameworks, and patterns they'll actually adopt next.
Dirtyfrag: Universal Linux LPE
85 points · 37 comments
AlphaEvolve: Gemini-powered coding agent scaling impact across fields
44 points · 2 comments
Natural Language Autoencoders: Turning Claude's Thoughts into Text
34 points · 7 comments
Agents need control flow, not more prompts
24 points · 5 comments
Agentic Engineering
22 points · 5 comments
Notable AI Lab activity
7 labs moved in the last 24h
Why this matters · GitHub event volume on a lab's own repos is the cleanest publicly-verifiable proxy for engineering activity we have.
OpenAI
−18 events (now 142) · San Francisco, US
Microsoft Research
+77 events (now 142) · Redmond, US
DeepSeek
+8 events (now 74) · Hangzhou, CN
Hugging Face
−33 events (now 63) · New York, US
Stanford CRFM
−17 events (now 11) · Stanford, US
Google DeepMind
−7 events (now 10) · London, GB
Google Research
−5 events (now 8) · Mountain View, US
Get this board in your inbox — one email, every morning (UTC).
Subscribe free