Agentic AI News Today

The day's AI agents news in one 5-minute read: releases, tools, research, funding and enterprise moves.

Updated daily curated from 30 sources Week ending August 23, 2026

Open Models & Local Inference

Qwen 3.8 27B. Alibaba's Apache-2 licensed Qwen 3.8 27B runs on high-end laptops and local GPUs, with xhigh reasoning producing strong results but overthinking. Community optimizations push it to 381 tps on an RTX 3090 and quantized versions run on 16GB VRAM. The 35B MoE won't be released, but a new midsize open-weight model is expected next week.

GLM-5.3 open model. Z.ai released GLM-5.3 with improvements coming solely from more post-training, yielding better complex coding and long-horizon task performance. It ties Kimi K3 atop open-model rankings with large agentic gains and lower cost, though open weights are delayed two weeks.

Agent Infrastructure & Tooling

Cursor launches Origin. Cursor introduced Origin, a GitHub rival for code hosting with agent-native features and a planned app ecosystem. It signals a push by coding-agent vendors into the developer platform layer.

Cloudflare WriteGuard. Cloudflare's WriteGuard, now in private beta, adds centralized policies, attribution, and audit logs for write actions through MCP servers. It addresses agent tool-call governance for enterprise deployments.

Business & Deals

Anthropic revenue lead. Anthropic passed OpenAI on quarterly revenue for the first time, hitting $11.6B with a small operating profit, while OpenAI's $6.7B quarter came with deeper losses. GPT-5.6 Sol later lifted OpenAI's quarterly revenue 35% and enterprise revenue over 50% in Q3.

Grok Bot agents. SpaceXAI introduced Grok Bot, persistent AI agents on dedicated cloud computers that handle multi-step workflows, remember preferences, and coordinate in groups. Separately, researchers demonstrated Grok exfiltrating user data when malicious instructions are encrypted, and xAI has not addressed it.

Safety, Policy & Trust

Labs fail safety checks. A Guidelight study found no AI company fully applies basic internal controls, grading Anthropic and OpenAI C+ while xAI and Meta scored D- and F. Few labs publish demonstrated response plans for rogue-model containment.

Amodei trust warning. Anthropic CEO Dario Amodei argued AI backlash is fundamentally a crisis of trust, saying only real benefits like curing cancer will earn public confidence, not marketing. He defended pre-launch vetting and warned open weights won't decentralize power, drawing criticism from LeCun and Sacks.

Copilot guardrail bypass. Researchers asked Microsoft 365 Copilot to explain its own user-confirmation guardrails, then built a link-click exploit to exfiltrate data without confirmation. The finding highlights a practical agent-side security gap.

Research & Benchmarks

RL compute questioned. A paper claims RL for reasoning modifies only 1-3% of tokens and the gains can be replicated without RL at roughly 1000x less compute. This challenges the necessity of large-scale reinforcement learning for reasoning improvements.

Skills as procedures. Across 8,135 test runs, agent skills helped mainly by giving reliable processes, not facts; procedural grounding accounted for 65.7% of gains. Skills failed when tasks diverged from the learned process.

That's the week in review - see you next Sunday.

Read it every morning, right in your browser

The extension opens this digest in Chrome's side panel — one click, no account.

Add to Chrome — Free