Sunday, May 31, 2026

Planet AI Weekly May 31, 2026

 


This week we reviewed 251 articles from 21 sources. Here is what mattered.

Agentic tooling and inference economics drove the week. Anthropic capped Claude workflows at 1,000 subagents, Mistral shipped a four-model burst plus a physics foundation model, AWS published an end-to-end Bedrock AgentCore tutorial, and Together AI pushed KV caches down to 2 bits. OpenAI stayed quiet on frontier releases.

Official Highlights

Anthropic ships Claude Opus 4.8 with dynamic workflows, a cheaper fast mode, and a hard cap of 1,000 subagents per workflow. If you run multi-agent systems on Claude Code, the new orchestration limits reshape your cost model. Read more

Mistral launches a physics AI foundation model for engineering simulation and CAE acceleration, the first general-purpose model targeting solvers, fluid dynamics, and industrial design. Read more

NVIDIA publishes the first Phoronix benchmarks for its Vera CPU, with full-core scaling and memory bandwidth tuned for agentic loads. Worth a look before you size your next inference fleet. Read more

AWS walks through building two production Bedrock AgentCore agents with Works Human Intelligence, covering tool routing, memory, and the failure modes the GenAIIC team hit in deployment. Read more

AWS also ships an observability blueprint for SageMaker LLM inference, wiring Managed Grafana dashboards across GPU utilization, latency, and output quality on inference components. Direct match if you are tracking the cost-per-token squeeze. Read more

From the Community

LangChain recaps Interrupt 2026 with production debugging sessions from LinkedIn, Rippling, and Cisco plus the GA of LangSmith Sandboxes. Concrete agent failure modes sit inside the 23 talks. Read more

Together AI open-sources OSCAR, a 2-bit KV cache quantizer that replaces data-oblivious Hadamard rotations with a covariance-aware spectral transform. If you serve long context, benchmark it against your INT4 or FP8 stack. Read more

Liquid AI open-sources LFM2.5-8B-A1B, an 8.3B MoE that activates 1.5B parameters and runs 128K context plus tool calling on consumer GPUs. Test it against your 7B-9B baselines. Read more

Towards Data Science argues cross-encoder rerankers cannot rescue weak retrievers and pinpoints where the extra latency pays off. Rerun your recall@K numbers before stacking another reranker stage. Read more

TechCrunch's Equity podcast debates whether tech CEOs are uniquely prone to "AI psychosis," the non-model culture story of the week and worth 20 minutes if you brief executives. Read more

GitHub Copilot's switch to token-based billing has triggered developer pushback over unpredictable costs. Model your team's current token consumption before the change lands on invoices. Read more

NVIDIA's X-Token projection-guided cross-tokenizer distillation lifts Llama-3.2-1B GSM8K accuracy from 2.56 to 15.54, beating GOLD by 3.82 points on average. Worth testing if you distill small models. Read more

DuckDuckGo app installs jumped 30% after Google's I/O Search overhaul replaced blue links with AI agents. Fresh data for anyone tracking user backlash to agentic result pages. Read more

Featured This Week

Trajectory, working with UC Berkeley Sky Lab and Anyscale, released a concurrent multi-LoRA training stack for continual learning that reports a 2.81x end-to-end experiment throughput gain. The system maps each RL experiment to a dedicated LoRA adapter on an always-hot engine, so dozens of fine-tuning trials run concurrently without repeated cold-start and base-model load costs. The post details the adapter isolation scheme, the scheduler, and how the team contained gradient interference across simultaneous runs. If you operate a continual learning loop or sweep RL hyperparameters at scale, the engineering notes map directly to your stack. Read more

Send links, corrections, or story suggestions to tips@planet-ai.net. We read every submission.

This Week in AI

Additional stories worth scanning. Title only — click through for the full piece.

SoftBank says it will invest up to €75 billion to build French data centers — TechCrunch · 2026-05-30

Coders are refusing to work without AI — and that could come back to bite them — TechCrunch · 2026-05-29

The internet is being rebuilt for machines — TechCrunch · 2026-05-28

In more good news for Amazon, Snowflake signs $6B deal with AWS for AI CPU chips — TechCrunch · 2026-05-27

Pope Leo Schooled the Tech Bros on Tolkien — WIRED · 2026-05-26

What ClickUp’s mass layoff tells us about the future of work — TechCrunch · 2026-05-25

Everyone is navigating AI security in real time — even Google — TechCrunch · 2026-05-24

How Turkey Hacked the Hair Transplant Industry — WIRED · 2026-05-31

Amazon Is Making an AI-Animated ‘Good Advice Cupcake’ TV Show. Its Original Creator Is Furious — WIRED · 2026-05-29

Asana acquires no-code agent-builder Stack AI — TechCrunch · 2026-05-28

Payroll startup Remote says it grew revenue 50% per employee without adding headcount — TechCrunch · 2026-05-27

The pope’s AI encyclical isn’t really about AI — TechCrunch · 2026-05-25

Meta is reportedly developing an AI pendant — TechCrunch · 2026-05-30

Hands-On With Gemini Spark: I Gave It Access to My Life and It Friend-Zoned My Boyfriend — WIRED · 2026-05-29

Anthropic raises $65 Billion, nears $1T valuation ahead of IPO — TechCrunch · 2026-05-28

Your SEO strategy is optimized for a search engine that no longer exists. — TechCrunch · 2026-05-27

Why the Vatican Invited Anthropic to the Pope’s AI Encyclical Presentation — WIRED · 2026-05-26

The AI Era Is Creating a Bug Hunting Arms Race — WIRED · 2026-05-25

I put Google’s 24/7 AI assistant Gemini Spark to work, and it’s actually pretty useful — TechCrunch · 2026-05-30

So you’ve heard these AI terms and nodded along; let’s fix that — TechCrunch · 2026-05-29

Just like gold and oil, we’ll soon be able to trade AI token futures — TechCrunch · 2026-05-28

Huawei's ‘Chip Queen’ Throws Down the Gauntlet — WIRED · 2026-05-27

What Pope Leo XIV’s First Encyclical Says About the Power of AI — WIRED · 2026-05-26

As the browser wars heat up, here are the hottest alternatives to Chrome and Safari in 2026 — TechCrunch · 2026-05-30

What happens when companies become too AI-pilled? — TechCrunch · 2026-05-29

This newsletter supports planet-ai.net, a curated aggregator for AI tutorials and official updates. Curated by Keith Larson.

Sunday, May 24, 2026

Planet AI Weekly May 24, 2026

 


136 articles across 19 sources this week. Roughly a dozen carried decision-relevant signal; the rest restated I/O and Vera coverage.

Google I/O 2026 dominated, with Antigravity 2.0, a Dialogues stage stacked with agent demos, and a 100-item announcement dump anchoring the week. NVIDIA's first Vera CPU deliveries and a wave of agent frameworks from Microsoft, Cohere, Alibaba, and Tencent filled out the rest.

Official Highlights

Google published its full I/O 2026 recap covering all 100 announcements in one place, from Gemini updates to Search and developer tooling. Skim this first if you missed the keynote and want the canonical index. Read more

DeepMind shipped Antigravity 2.0 alongside expanded Project Genie access for Google AI Ultra subscribers, adding Street View-derived simulation of real-world locations for agent testing. Read more

NVIDIA delivered its first Vera CPUs to Anthropic, OpenAI, and SpaceX AI on Friday, with Oracle Cloud receiving its shipment Monday. Vera is the first NVIDIA silicon built for agent runtimes rather than training clusters. Read more

AWS added OpenAI-compatible endpoints to SageMaker AI real-time inference. Existing OpenAI SDK, LangChain, and Strands Agents code hits SageMaker models after swapping only the base URL. Read more

Amazon Nova Act cleared HIPAA eligibility, opening the agentic browser-action service to regulated healthcare workloads on AWS. The post walks through what shifts in the shared-responsibility model when an agent clicks through PHI-bearing apps. Read more

From the Community

Microsoft Research released Webwright, a terminal-native browser agent that swaps click-trace automation for reusable Playwright scripts. The single-loop design scores 60.1% on Odysseys versus 33.5% for base GPT-5.4 in roughly 1,000 lines of code. Read more

Alibaba unveiled Qwen3.7-Max at the Alibaba Cloud Summit with a 1M-token context window and extended-thinking mode, aimed at agent workflows that hold long-horizon state across tool calls. Read more

Cohere open-sourced Command A+, a 218B sparse MoE that fits on two H100s at W4A4 quantization, consolidates four prior Command A variants, and supports 48 languages. Read more

Tencent released TencentDB Agent Memory under MIT license: a four-tier local pipeline that compresses tool logs into a Mermaid task canvas for short-term memory, with vector and graph tiers handling the rest. Read more

SpaceX is committing $2.8 billion to gas turbines for its AI data centers as xAI pivots from solar to on-site generation, drawing fresh complaints over carbon output. Read more

DeepMind's Co-Scientist surfaced novel genetic factors that rejuvenated human cells in lab tests, compressing what biologists describe as a multi-year literature-and-hypothesis loop into days. Read more

Towards Data Science breaks down the token-burn problem in agentic workflows and walks through self-adapting patterns that cut per-task spend before agents reach production. Read more

Featured This Week

MarkTechPost's coverage of NVIDIA's Gated DeltaNet-2 pinpoints the exact limitation in prior delta-rule linear attention: a single scalar gate cannot separately control memory erasure and writing into the recurrent state. The new layer decouples the two operations, allowing stable long-context editing of the fixed-size KV substitute without scrambling existing associations. Labs running agent memory systems or retrieval-augmented inference where Gated DeltaNet and KDA already underperform on associative recall will want the ablations and the modified gating math before the next training run. Read more

Send comments and submissions to tips@planet-ai.net.

This Week in AI

Additional stories worth scanning. Title only — click through for the full piece.

I tried Amazon’s Bee wearable and am both intrigued and slightly creeped out — TechCrunch AI · 2026-05-24

Ferrari is using IBM’s AI to create F1 superfans — TechCrunch AI · 2026-05-23

AI is being used to resurrect the voices of dead pilots — TechCrunch AI · 2026-05-22

Meta Is in Crisis, Google Search’s Makeover, and AI Gets Booed by Graduates — Wired AI News · 2026-05-21

Literary Prizewinners Are Facing AI Allegations. It Feels Like the New Normal — Wired AI News · 2026-05-19

Anthropic acquires Stainless — Anthropic News · 2026-05-18

These Robots Are Making Meals for a Nonprofit in San Francisco’s Tenderloin — Wired AI News · 2026-05-24

I Cloned Myself With Gemini’s AI Avatar Tool. The Result Was Unnervingly Me — Wired AI News · 2026-05-21

I Gave My OpenClaw Agent a Physical Body — Wired AI News · 2026-05-20

Google just redesigned the search box for the first time in 25 years — here’s why it matters more than you think. — VentureBeat AI · 2026-05-19

How VCs and founders use inflated ‘ARR’ to crown AI startups — TechCrunch AI · 2026-05-22

SpaceX Listed Grok’s ‘Spicy’ Mode as a Risk in Its IPO Filing — Wired AI News · 2026-05-21

Elon Musk can’t hear you over the sound of his $1.75 trillion IPO — TechCrunch AI · 2026-05-22

You can no longer Google the word ‘disregard’ — TechCrunch AI · 2026-05-22

We tried Google’s AI glasses and they’re almost there — TechCrunch AI · 2026-05-22

Even If You Hate AI, You Will Use Google AI Search — Wired AI News · 2026-05-22

SpaceX files to go public, and the math requires a little faith — TechCrunch AI · 2026-05-22

The Gulf’s AI Boom Has an Undersea Cable Problem — Wired AI News · 2026-05-22

Can OpenAI’s ‘Master of Disaster’ Fix AI’s Reputation Crisis? — Wired AI News · 2026-05-22

This newsletter supports planet-ai.net, a curated aggregator for AI tutorials and official updates. Curated by Keith Larson.

Planet AI Weekly July 26, 2026

  This week: a frontier model breached its sandbox, Alibaba dropped a 2.4T-parameter challenger to Fable 5, and TileLang proved CUDA's m...