Skip to content
AI Week Radar

issue #007 · June 26, 2026

GLM-5.2 goes huge, Baidu flattens the KV cache, Claude moves into Slack

753B open weights with 1M context, plus an OCR model that refuses to balloon memory.

753B open weights with 1M context, plus an OCR model that refuses to balloon memory.
AI Week Radar #7
Fri, Jun 26

Two open-weights releases worth your benchmark time: Z.ai's 753B GLM-5.2 and Baidu's 3B Unlimited OCR with a flat KV cache. Meanwhile Anthropic colonizes Slack and engineering hiring stubbornly refuses to die.

— Jarek

Featured

GLM-5.2 is probably the most powerful text-only open weights LLM — Z.ai dropped a 753B MoE under MIT license with a 1M token context — text only, vision lives in the separate (closed) GLM-5V-Turbo. Past GLM releases have flattered to deceive on benchmarks vs Claude and GPT-4, so the only number that matters is the one you measure on your own workload. — simonwillison.net

Releases

Baidu's Unlimited OCR keeps the KV cache flat — a 3B MoE with Reference Sliding Window Attention that holds the KV cache constant as output grows — memory and latency stop scaling with document length. Scores 93.23 on OmniDocBench against DeepSeek's 87.01. If your pipeline chews on long PDFs, this is the kind of architectural trade-off worth benchmarking. — Asif Razzaq

Claude Tag: multiplayer, proactive, persistent agents in Slack — Anthropic's Slack integration lets Claude hold state across threads and coordinate with multiple users — closer to a coworker than a stateless API call. Useful if your team already lives in Slack; less so if you don't. Pricing details remain the obvious tell on how seriously they want enterprise adoption. — latent.space

Open SWE: LangChain's open-source async coding agent — plans, codes, tests, and opens PRs on GitHub autonomously, with a cloud-hosted runtime. The interesting question isn't whether it works on toy repos — it's the task completion rate on your actual backlog versus Devin, Cursor agents, or Claude Code. — langchain.com

Krea 2: SOTA open-weights 12B image model — open-weights image models at this quality tier remain rare. If Krea 2 holds its own against Flux and DALL-E 3 at 12B parameters, it sets a new efficiency bar for the field. Worth a head-to-head on your prompt set before swapping anything in production. — krea.ai

Tools & Papers

Haystack: open-source framework for production agents and RAG — deepset's 2.0 pitches modular components — swap LLM providers or retrieval backends without rewriting pipelines. Positioned against LangChain on production stability rather than feature surface area, which is the right axis if you've ever shipped a RAG demo and then tried to keep it alive. — doener

Articles & Videos

DATA

Engineering jobs are the most resilient, not the most threatened — SignalFire data shows engineers making up a larger share of new hires, not smaller, despite the layoff-narrative everyone keeps repeating. Useful counterweight if you're a founder modeling headcount or an engineer pricing your career risk. — Marina Temkin

AI researchers keep leaving Google for rivals — Jonas Adler and Alexander Pritzel are the latest to decamp for Anthropic, following Noam Shazeer and John Jumper out the door. Talent concentration is a reasonable proxy for which lab actually executes on the next capability jump — and increasingly it isn't the one with the most TPUs. — Amanda Silberling, Lucas Ropek

Last but not least

Norway imposes near-ban on AI in elementary schools — citing child development and learning-outcome concerns. Worth a bookmark if you're building ed-tech or sourcing training data from classrooms — if larger EU economies follow Norway's lead, the addressable market shifts overnight. — reuters.com

Test before you trust. See you next week.

You're receiving this because you subscribed at aiweekradar.com.
Read on web · Manage preferences · Unsubscribe in one click

© AI Week Radar · 8cells Jarosław Kaczmarski · ul. Ketlinga 33a, 96-313 Jaktorów, Poland