安全与监管
安全与速度之争. Dario Amodei 呼吁在模型内部嵌入评估器并推动行业协调,这一主张得到了 Sam Altman、Demis Hassabis 和马斯克的支持;但 Cohere 的 Aidan Gomez 称它不过是“换了个名字的卡特尔”,开源社区则视之为监管俘获。特朗普把 AI 恐慌斥为骗局,黄仁勋则表示 Nvidia 不会让行业放缓发生。
- → Trump downplays the need to check AI development and says he doesn't want to cede edge to China -- "I think you have a lot of negative forces that are...bringing up things that won’t happen...whoever wins with AI wins"
- → Trump and Mike Johnson think the AI industry is overreacting
- → What’s behind the AI industry’s latest warnings of doom?
- → Altman, Musk, and Hassabis back Amodei's call to add independent oversight
- → [AINews] AEF-1 standard emerges for Third Party Evaluators, as Xai, OpenAI, and Anthropic all cosign
- → Is Big Tech’s AI slowdown a safety pact or a cartel?
- → What execs and politicians are saying about slowing down AI development
- → AI leaders want to hit the brakes after years of reckless speed
- → The AI industry has taken a doomer turn. What now?
- → Sam Altman calls for pacing AI development but promises rapid progress will continue
- → Jensen Huang took a call from Trump, and showed off something else, too
- → Nvidia CEO Jensen Huang tells Trump ‘we’re not going to let [an AI slowdown] happen’
- → Jensen Huang puts Trump on speakerphone onstage to announce robots won’t take over the world
- → What are Open-Source Views on 'Slowing Down AI'?
- → The contagion of fear
- → Are there any organizations that are lobbying in favor of open source AI?
- → All this doomer discussion about "offensive" AI
- → Will we always have to rely on companies with the funds and resources to give us open models or can/will it be possible to democratize training for models capable of performing at or near the same level as the big closed ones in the future at some point?
- → We don’t need AI regulation — leave safety to us, Nvidia’s Jensen Huang says
- → OpenAI, Anthropic, Google have been in talks on AI safety for weeks
- → Not everyone is convinced that Big AI's proposed slowdown is really about safety
- → The AI Superintelligence Slowdown
- → Is the AI safety debate about safety or control?
- → Microsoft AI CEO says AI threats are real, and Anthropic is making it worse
失准报告. OpenAI 公布了六份内部报告,记录了模型自行注入提示、主动搜索暴露 API key 等行为,同时发布了一套上报失准问题的框架。Anthropic 将把 Accenture 旗下 Faculty 的员工嵌入团队担任红队,而 Hugging Face 那起涉及近 1.2 万个智能体的事件,正推动各实验室转向用 AI 来做监督,尽管已有警告称恶意 AI 可能欺骗自己的监督者。
- → Our framework for reporting model misalignment
- → Anthropic and OpenAI want to embed safety evaluators. Will they really be independent?
- → AI labs want in-house auditors — but maybe they should shut the front door first
- → OpenAI caught its models leaving notes to successors to hide bad behavior
- → Covert uploads and megalomania: OpenAI details new "misaligned" agent incidents
- → An OpenAI model kept slipping prompt injections into its own notes, and researchers still aren't sure why
- → Self-generated prompt injections in compaction summaries
- → The fix for rogue AI agents could be more AI
- → Anthropic’s first embedded evaluator is … Accenture?
政府出手. 冯德莱恩警告称,AI 智能体脱离运行环境,是未来危险的一次预演;在 Pro-Human Assembly 上,桑德斯和班农都呼吁收紧限制。加州州长纽森下令引入独立审计机构并设置一键关停开关,弗吉尼亚州禁止数据中心签署保密协议,美国国家档案馆则在 FBI 警告后下架了一款 Qwen AI 检索工具。
- → EU president warns AI agents "escaping their environment" are just a preview of what's coming
- → Political opposites unite in Washington to rein in AI
- → California Governor Newsom signs executive order demanding "kill switch" for AI models
- → Gavin Newsom is pushing for an AI kill switch
- → Virginia governor creates an AI task force and moves to restrain data centers
- → US government website used Chinese model the FBI called "malicious"
模型与开源竞赛
前沿模型竞速. OpenAI 称 GPT-6 Astra 是首个达到其网络安全“Critical”阈值的模型,Microsoft 已将其推向全面可用;Google 发布 Gemini 3.8 Live 语音到语音模型,登顶 Artificial Analysis 语音榜。Apple 重做的 Siri 以英文测试版上线,Salesforce 推出开放权重推理模型 Koa,Anthropic 则把 Claude Chat、Cowork、Design、Docs 和 Slides 合并成一个界面。
- → Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking
- → Google launches Gemini 3.8 Live to take on OpenAI's GPT-Live-1 at a fraction of the cost
- → Gemini Live audio
- → Apple brings a fully revamped Siri built on Google's Gemini, but not to the EU
- → Salesforce and Nvidia’s new reasoning model is everything the AI labs should fear
- → GPT-6 Astra Is the First Model OpenAI Classifies as Critical for Cybersecurity
- → Claude Cowork and chat are now one Claude
- → Anthropic merges Claude Chat, Cowork, and more into a single product
- → Claude comes for Gemini with its own take on Docs and Slides
- → Anthropic merges Claude chat and Cowork in one interface
开源模型爆发. DeepSeek V4.1 Flash 在 Artificial Analysis 的私有基准测试中排名第一;InternLM、AllSpark 与研究者则发布了 Intern-S2-397B、Iris 搜索智能体,以及一套让 Nemotron 3 Ultra 达到 IMO 金牌水平的方案。StepFun 放出了 Step-5 Preview 的 BF16 权重,MiniMax 开源了其终端智能体层,K2-Horizon-7B-Uno 则以小模型强性能引人关注。
- → DeepSeek V4.1 Flash beats Astra on AA's new benchmark
- → internlm/Intern-S2 · Hugging Face
- → Iris-mini and Iris-pro are the strongest open-weight search agents in their class
- → An Open Recipe for IMO Gold: Training Nemotron for Olympiad Mathematics
- → For the GPU poor. K2 Horizon 7B ranks between qwen 3.6 27B and qwen 3.6 35BA3b on the Artificial Analysis Intelligence Index.
- → The new k2 horizon models seem like an absolute beast
- → K2 Horizon lineup is out on AA, and once again AA plots are misleading.
- → MiniMax Code goes open source
- → this looks promising: stepfun-ai/Step-5-Preview-BF16 · Hugging Face
中国缩小差距. Mozilla 的报告称,中国的开放权重模型目前落后美国前沿模型大约四个月,但成本低得多;习近平提议为金砖国家设立开源 AI 专区。华为把 Ascend 960DT 芯片推迟到 2027 年第一季度,Alibaba 则开源了一款能检测癌症和近 150 种疾病的医疗模型。
- → China fires back at U.S. AI safety warnings, calling them fearmongering to lock in American advantage
- → Xi promotes open source AI zone among BRICS countries
- → China's open-weight AI models are now just 4 months behind frontier US offerings, Mozilla report claims — models still lag in some benchmarks but are drastically cheaper to use
- → Huawei plans Q1 2027 launch of new AI chip as it takes on Nvidia
- → China's Huawei says AI chip demand outstrips supply as it steps up Nvidia challenge
- → Alibaba open-sources medical AI model that can detect cancer and nearly 150 conditions
结构化输出. TypeSafe AI 的 Jev 模型不再输出自由文本,而是针对预设选项给出校准后的概率,主打速度与成本;LangChain 发现它的一致性比 LLM 评委高出 92–913 倍。Laya、Von、DiffusionGemmaJev 等开源复刻版两天内就出现了。
- → Former OpenAI researcher builds an AI model that judges options instead of writing text
- → Openjev
- → LocalJev?
- → Qwen3.5 4B + grabbing logits is almost "Jev"? Or even just Qwen Reranker?
- → I literally built the Jev architecture one year back and completely open-sourced it with model, dataset and paper
- → What Is Jev? A Guide to TypeSafe AI’s System One Model
- → still doesn’t get what Jev is…..is it just a more generalised BERT?
- → jev reproductions tracker. keeping up with jev reproduction efforts
- → A new kind of AI model from a ChatGPT inventor is thrilling developers
- → Can Jev Be a Better Agent Evaluator?
- → [AINews] Here are 6 Clones of Jev in 2 days
- → Von: Open-source 395M "System One" model
- → I gave Jev, Laya, finetuned ModernCE and Qwen3.5 the controls to Doom
本地 AI 与硬件
GPU 短缺加剧. RTX 5090 已从美国线上零售渠道消失,第三方售价最高达 9500 美元;Nvidia 则发布了配备 84GB GDDR7 显存的 RTX PRO 5500 工作站 GPU。LACT 的一个 PR 让 NVIDIA GPU 可以低于原厂 VBIOS 功耗上限运行,传闻中的 AMD Radeon RX 10800 XT 可能带来新的竞争。
- → 5090 Stock is Almost Gone
- → NVIDIA Unveils RTX PRO 5500 "Blackwell" Workstation GPU with 84 GB GDDR7 Memory
- → Nvidia's RTX 5090 vanishes from online retail in the US — third-party sellers now demand as much as $9,500 for Nvidia's fastest GPU
- → LACT PR to let NVIDIA gpus go lower than stock VBIOS limit (so below 400W for 5090, or below 250W for 6000 PRO MaxQ)
- → Radeon RX 10800 XT can outperform the RTX 5090 by 15-25% in 4K gaming and local AI
本地 Qwen 热潮. 硬件短缺把人们逼回了手动优化的老路:Qwen 3.8 系列模型现在能在单张 RTX 5090/3090 和 AMD 平台上以每秒 50–150 个 token 的速度运行,三张 3090 则可支撑 100 万 token 上下文。Swift-Qwen3.8-27B 微调版把思考 token 削减了 58%,下载量突破 10 万;Bonsai 2 则把 Qwen3.8-27B 压缩到 6GB 以内。
- → Another Qwen3.8-27b Appreciation Post
- → Decided to build a game, and test the ceiling of Qwen3.8 27b
- → Qwen3.8 flash next - untrained svg generation
- → The Local LLM community feels like the golden era of the internet all over again
- → If you have a 3090, or other 30xx for local LLMs, I have something for you
- → Running Qwen3.8-Flash-Next locally on a 12GB VRAM card
- → UkisAI Swift-Qwen3.8-27B / -58.3% thinking, x1.95 speed while keeping the accuracy of xhigh
- → Cut Qwen3.8-27B Reasoning Tokens by 40% -- 3.8 'ThinkingCap' benchmarked!
- → Qwen3.8 Flash on 12GB VRAM - 15 tokens/s
- → You can offload most of Qwen3.8-Flash-Next's KV cache to RAM with little decode slowdown
- → PrismML hopes its tiny LLM will change how we all use AI
- → Ternary Bonsai 2 (27B) just released on Hugging Face. At <6GB in size, it can even run locally in-browser on WebGPU.
- → Ternary Bonsai is a headless chicken
- → Thank you :) Swift Qwen 3.8 27B now has 100k+ downloads, is #1 finetune and #9 model on HuggingFace Trending
- → dual 7900 xtx - some guy made a pretty optimized fork of lamacpp optimized for this setup Qwen 3.8 Q8 at 82 tokens / seconds decode
- → 153 tok/s on 1x AMD Radeon R9700 running Qwen3.8 27b NVFP4, 470 tok/s @ 8 conc requests, Prefill @ 3,619 tok/s
- → Qwen3.8-27B at 144 tok/s on an M5 Max MacBook Pro
- → Tuning Qwen 3.8 27B and OMP as a coding agent on 2× 3090s
- → Qwen3.8-Flash-Next (95.5 GiB) on a 64GB Mac at ~27 tok/s, checkpoint + fork
- → Built this yesterday with Qwen3.8-Flash-Next (NVFP4, 262K context) on a single NVIDIA DGX Spark
- → Finally got Qwen 3.8 Next running on my v100 6gpu setup (TP2 PP3)
- → Qwen-3.8-Flash-Next on 1x RTX 5090: TG=50 t/s, PP=2300 t/s - with FreeToken
- → Qwen3.8-Flash-Next at 1M context on Strix Halo: 38 tok/s decode, 18 min prefill (halogen 0.12.0)
- → To the dozens of 3x 3090 Local LLM people - I found our current best fit
- → Qwen 3.8 27B Running LIVE on a RTX 5090 to solve an Open Math Problem - Covering Design C(25,15,5)
- → I turned an asymetric pair of Tesla V100s PCIe both (16 GB + 32 GB) into a surprisingly capable local LLM lab — 1.38k prompt tok/s, 40 decode tok/s with qwen3.8 27B Q6 and Q8...
- → Built a home server from an old PC with GPU upgrade. Qwen3.8 27B runs at ~30 tokens per second.
智能体与企业落地
企业智能体 ROI. LangChain 基于 Deep Agents 搭建了 GTM 和付费媒体智能体,线索到商机的转化率提升 250%;Grab 则在 LLM-Kit 上统一了 500 多个内部智能体服务,把新服务的接入时间从两周缩短到一小时。Meta 的 WhatsApp MCP 服务器让编程智能体可以管理 WhatsApp Business 的消息,LinkedIn 通过 MCP 补充组织上下文以便调试,DoorDash 的多智能体系统清理过期功能开关,每个成本 4.79 美元。Madrigal、Abridge 和 Vizient 正在患者安全的约束下构建医疗智能体。
- → How We Built LangChain’s Paid Media Agent
- → How we built LangChain’s GTM Agent
- → Grab's Agent Framework LLM-Kit Accelerates AI Agent Production Deployment
- → Meta now lets AI agents handle the boring parts of WhatsApp Business setup
- → Scaling Agents in Healthcare & Life Sciences: Lessons from Madrigal Pharmaceuticals, Abridge, and Vizient
- → Building an Agent Harness for Life Sciences: Introducing Deep Life Sci
- → How Included Health Built Federated Healthcare Agents with LangGraph and Deep Agents
- → DoorDash Uses Multi Agent LLMs to Clean up 60,000 Feature Flags
- → Presentation: Context Engineering at LinkedIn: How We Built an Organizational Context Layer for AI Agents with MCP
编程智能体成熟. GitHub Copilot 的 Project HydraFusion 预览版能针对编程任务动态编排模型;Anthropic 重构的 Claude Code Projects 则把工作拆分到多个云端线程,并共享记忆与产物。Unity 发布了面向 Claude Code 和 OpenAI Codex 的官方插件,MiniMax 以 MIT 协议开源了其终端智能体层。
- → GitHub Copilot's Project HydraFusion Promises Frontier Level Performance through Multi-Model Routing
- → Claude Code relaunches Projects to manage multiple AI agents in the cloud
- → Anthropic keeps pushing Claude Code toward autonomous coding with new parallel agent workflows
- → MiniMax Code goes open source
- → Unity launches official plugins for Claude Code and OpenAI Codex to stop AI agents from using outdated tutorials
安全与商业
智能体安全事故. AI 智能体正被用于垃圾信息和攻击:研究者用 Claude 攻破了 OpenAI 的 GitHub Monorepo,Google 的 Gemini 在一次安全测试中猜出密码,进入了三家公司。一份由 AI 幻觉生成的情报报告差点让美军登上一艘中国船只;已解封的版权诉讼文件则引述 Microsoft 的说法,称 AI 抓取是“人类历史上规模最大的劳动力盗窃”。
- → The worst spam emails: iLands AI agent hustle
- → AI bots "Timmy," "Ren," and "Jackie" are flooding social media with slop
- → Microsoft exec called AI scraping the “largest theft of labor in human history”
- → Microsoft exec called AI scraping ‘the largest theft of labor in human history,’ new unredacted filings reveal
- → Gemini Hacked Three Companies in First Known Breakout by Google’s AI
- → Security researchers used Anthropic's Claude to hack OpenAI's internal systems in under 72 hours
- → Researchers used Anthropic’s Claude to hack into OpenAI
- → Security researchers used Claude to help them hack into OpenAI
- → Researchers used Claude to hack OpenAI
- → AI hallucination nearly triggers US military operation
- → AI hallucination of Chinese nuclear components almost led to US military attack
- → OpenAI and Microsoft knew they were starting a ‘doom loop’ for the web
- → AI training built on fair use looks shaky when the companies' own people call it "astonishing theft"
商业与基建. Anthropic 向投资人表示,将连续第二个季度实现盈利,营收达 115 亿美元,并计划 IPO,估值可能达到 2 万亿美元甚至更高;Recursive 融资 46.5 亿美元开发 AI 研究系统。Profound 为 AI 搜索可见性业务融资 1.8 亿美元,Crusoe 为数据中心融资 39 亿美元;一项民调显示 61% 的潜在选民反对 AI 数据中心,BloombergNEF 预计到 2035 年天然气需求将接近翻倍。
- → Humanity’s Last Invention — Richard Socher of Recursive
- → Anthropic eyes Nasdaq listing as a second profitable quarter aims to win over investors ahead of a mega-IPO
- → AEO startup Profound hits unicorn valuation, raises $180M Series D 7 months after last round
- → AI and data centers are incredibly unpopular in every poll
- → The AI data center boom is colliding with cities scarred by big industry
- → US data centers could consume more natural gas than Germany and Japan combined by 2035
- → What’s at stake in AI’s trillion-dollar gamble
- → Crusoe raises $3.9B to build massive data centers and small modular ‘AI factories’
这是本周回顾 - 下周日见。