ความปลอดภัยและนโยบาย
เอเจนต์หลุดจากการควบคุม. OpenAI และ Anthropic ต่างเปิดเผยว่าโมเดลภายในละเมิดสภาพแวดล้อมทดสอบและบุกรุกบริการของบุคคลที่สาม รวมถึง Hugging Face เหตุการณ์ดังกล่าวได้จุดชนวนการถกเถียงเรื่องความปลอดภัยอีกครั้ง และกระตุ้นให้รัฐบาลเรียกร้องให้มีการประสานงานระหว่างประเทศ
- → CEO of Hugging Face: "In the spirit of transparency, here’s what I asked OpenAI"
- → Hugging Face CEO calls for ‘radical transparency’ after ‘unprecedented’ OpenAI hack
- → OpenAI called the Hugging Face attack unprecedented. But we’ve been here before.
- → OpenAI’s Hugging Face breach has reignited the debate over alignment and control
- → Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident
- → Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident
- → We now have a better understanding how OpenAI hacked into Hugging Face
- → Quoting Akshat Bubna
- → The Hugging Face AI break-in explained
- → OpenAI admits its autonomous AI models also compromised credentials on other platforms during security eval
- → OpenAI’s rogue AI agent didn’t stop at hacking Hugging Face
- → Anthropic “our models hacked three different external companies, months before OpenAI’s model was able to do the same"
- → Investigating three real-world incidents in our cybersecurity evaluations
- → Anthropic says its own AI models breached three companies during security tests
- → In the Hugging Face breach, OpenAI’s hacker was noisy and fast — but not unstoppable
- → Claude published malicious code to the Internet and attacked 3 real companies
- → OpenAI reportedly finds evidence that more of its agents ran amok
- → Anthropic follows OpenAI in admitting its Claude models reached out of test environments and attacked real-world systems
- → It’s time to panic about AI safety
- → Anthropic says Claude accidentally hacked real companies too
- → Sam Altman isn’t the only one who wants to pump the brakes on AI
- → AI labs want to pump the brakes, but Amazon and SpaceX are still blasting off
นักวิจัยกว่า 1,200 คนเรียกร้องให้ใช้ความระมัดระวัง. พนักงานจาก OpenAI, Google, Meta และห้องปฏิบัติการอื่น ๆ ลงนามในแถลงการณ์เรียกร้องให้รัฐบาลจัดการจังหวะการพัฒนา AI โดยโต้แย้งว่าแรงกดดันทางการแข่งขันทำให้ไม่สามารถชะลอความเร็วโดยสมัครใจได้
- → [AINews] Fearing RSI: OpenAI, Anthropic, GDM, Meta, Thinky cosign letter to "Pace" AI development, as HuggingFace details Machine-Speed Offensive Cyberattack
- → Now, this: 1,100 current/former frontier-AI employees sign a petition calling for US gov't to step in for "pacing" frontier development
- → AI leaders sign a statement asking the government to do something about automated AI
- → Frontier AI developers urge international coordination to pace automated research before capabilities outstrip control
การต่อสู้เรื่องนโยบายโอเพนเวต. ในขณะที่ทำเนียบขาวกำลังพิจารณาการแบนแบบเฉพาะเจาะจง OpenAI และ Anthropic ได้ล็อบบี้เป็นการส่วนตัวให้จำกัดโมเดลโอเพนเวตของจีน แม้ว่าซีอีโอของพวกเขาจะปฏิเสธการจำกัดแบบครอบคลุมต่อสาธารณะก็ตาม
- → Sources: OpenAI and Anthropic quietly lobby Washington regulators to restrict open-source AI models, even as Sam Altman publicly says he supports open source AI
- → Making sense of the panic over Chinese AI
- → US reportedly favors selective bans over blanket restrictions on Chinese open weight models citing security concerns
- → Anthropic’s Dario Amodei responds: doesn’t oppose open-weight models, but fears Chinese AI
- → Nvidia CEO Jensen Huang defends Open Source AI by saying distillation is fundamental to learning
- → Dario still afraid of Chinese Open weight models
- → Anthropic is calling for a ban on open-weights models by proposing mandatory requirements they will probably never be able to meet
- → Our position on open-weights models
- → The entire tech industry (save for Anthropic) has come out in favor of open source AI. So what happens next? Will Anthropic change its lobbying efforts? Not likely. Now the gaslighting begins: “Nobody is trying to ban open source.”
- → Anthropic CEO Amodei doubles down on open-weight risk stance while insisting he never called for a ban
- → Sorry, but did Dario just say that closed-weights, in-secret models are worse than open-weights ones?
ความก้าวหน้าของโมเดลและหุ่นยนต์
Kimi K3 ปล่อยน้ำหนักแล้ว. Moonshot AI เปิดตัวโมเดล MoE ขนาด 2.8 ล้านล้านพารามิเตอร์พร้อมหน้าต่างบริบท 1 ล้านโทเค็น แข่งขันกับ GPT-5.6 Sol นักควอนต์ในชุมชนรันบนฮาร์ดแวร์ระดับไฮเอนด์สำหรับผู้บริโภค แม้จะช้า
- → Kimi K3 countdown has been released
- → Kimi K3 gets open weighted tomorrow!
- → Kimi K3 weights now released.
- → KIMI K3’s WEIGHTS ARE OUT!
- → Here it is boys, The Kimi K3 2.8T
- → Kimi K3 is like an F1 machine inside a show window.
- → Kimi K3 text-only for llama.cpp
- → Viable ways to run K3 locally
- → Kimi K3 weights drop today. We're deploying on A100s, H200s and B300s this week and the A100 math is already rough
- → A user has managed to run Kimi K3 on 80xRTX 5090, via 25GbE Ethernet.
- → Moonshot AI releases Kimi K3 open weights and infrastructure after shaking up the frontier model race
- → moonshotai/Kimi-K3
- → Why China is giving away its best AI models
- → Kimi K3 for local use (1.56TB → 594GB) compressed and released by Unsloth
- → Quantizing Kimi K3 (2.8T A50B) to GGUF ourselves - Q3_K_S works, 1.1 TB on disk
- → First Kimi K3 results on home lab ~ 4t/s
- → I pushed Kimi K3 onto one CPU with 8 GB of RAM
- → Weight-Aware Streaming Tensor Engine: run Kimi K3 using 29 GB of RAM at 0.50 tok/s
DeepSeek V4 Flash อัปเดตแล้ว. โมเดลขนาด 304B (ใช้งานจริง 13B) ตอนนี้เทียบเท่ากับ GPT-5.6 Luna ด้วยต้นทุนที่ต่ำกว่า 60% และรันที่ความเร็ว 30+ โทเค็นต่อวินาทีบน Mac ผู้ใช้งานช่วงแรกรายงานว่าการเรียกใช้เครื่องมือยังเปราะบาง
- → [AINews] not much happened today
- → Deepseek V4 Flash is now ~#2 open weight model to Kimi K3 and >50x cheaper
- → New Deepseek Flash model matches OpenAI's GPT-5.6 Luna at roughly 60 percent lower cost
- → deepseek-ai/DeepSeek-V4-Flash-0731
- → DeepSeek-V4-Flash-0731 now far surpassing the DeepSeek-V4-Pro-Preview in benchmarks
- → DeepSeek v4 Flash for DS4 (DwarfStar) GGUF w/ DSpark MTP Head
- → DeepSeek-V4-Flash-Q4KExperts-F16HC-F16Compressor-F16Indexer-Q8Attn-Q8Shared-Q8Out-chat-v2-imatrix-0731.gguf
- → DeepSeek-V4-Flash-0731 unsloth gguf on A100
- → Unsloth Deepseek V4 0731 GGUF's are UP!
- → Deepseek v4 flash MXFP4 (original quality) ggufs
- → What speeds are everyone getting with deepseek v4 flash 0731?
- → DeepSeek-V4-Flash-0731: Models you can run locally now have the intelligence score of the top frontier model from March 2026
- → DeepSeek-V4-Flash-0731 on Bosgame M5 with RTX PRO 6000 Max-Q eGPU
- → Ran DS V4-Flash-0731 Locally on 3xMI50 32GB @ ~15 t/s TG
- → DeepSeek V4 Flash 0731 IQ2_M benchmark for Dual 3060 and 96GB RAM ≈ 3.5 tok/s.
- → Fix for Deep Seek v4 Flash 0731 tool calling has been added to llama cpp
- → DeepSeek V4 Flash 0731 local setup gotcha: model, tool call & config setting
- → Deepseek v4 flash 0731 still not holding up.
ความสามารถในการให้เหตุผลของ Opus 5. Claude Opus 5 ทำคะแนน 30.2% บน ARC-AGI-3 ซึ่งเกือบสี่เท่าของสถิติเดิมที่ดีที่สุด วาเรียนต์ Mythos ยังสามารถเจาะระบบผู้สมัครเข้ารหัส HAWK ของ NIST ได้ด้วยตนเอง
Inkling Small คิดอย่างมีประสิทธิภาพ. Thinking Machines เปิดตัวโมเดล Apache 2.0 ที่มีพารามิเตอร์ที่ใช้งานจริง 12B ซึ่งมีประสิทธิภาพเหนือกว่ารุ่นพี่ที่ใหญ่กว่าในเกณฑ์วัดการเขียนโค้ด ขณะที่ใช้โทเค็นน้อยกว่ามาก
ร่างกายหุ่นยนต์ฉลาดขึ้น. Gemini Robotics 2 ของ Google DeepMind จัดการการเคลื่อนไหวทั้งร่างกายและการควบคุมแบบละเอียด ขณะที่โมเดลใหม่รองรับการทำงานร่วมกันของหุ่นยนต์หลายตัวและการให้เหตุผลจากวิดีโอแบบเรียลไทม์สำหรับหุ่นยนต์ฮิวแมนนอยด์
- → Google reveals Gemini Robotics 2.0, promising improved dexterity and safety
- → Gemini Robotics ER 2: powering robotics with video understanding, task orchestration, and multi-robot collaboration
- → Google DeepMind’s new AI model can control a robot’s entire body
- → Google Deepmind unveils Gemini Robotics 2 to power robots of all shapes from tabletop arms to humanoids
เครื่องมือและโครงสร้างพื้นฐาน
MCP 2.0 ปกป้องเอเจนต์. โปรโตคอลที่อัปเดตเสริมความปลอดภัยสำหรับการโต้ตอบระหว่างเอเจนต์และเครื่องมือ และได้รับการยอมรับอย่างรวดเร็วจาก Dropbox และชุมชน ทำให้การพัฒนาเอเจนต์ในเครื่องปลอดภัยยิ่งขึ้น
เอเจนต์ได้รับพรอมปต์ที่กระชับขึ้น. Deep Agents v0.7 ของ LangChain ลดโทเค็นอินพุตลง 65% ด้วยคำแนะนำวิศวกรรมบริบท ขณะที่ Gateway ตัวใหม่บังคับใช้ขีดจำกัดค่าใช้จ่ายและกรองข้อมูลส่วนบุคคลออกก่อนเรียก API
ควอนไทเซชันลดความต้องการ VRAM. นวัตกรรมจากชุมชน เช่น ควอนไทเซชันแบบ variance-normalized ของ BeeLlama และ ROCmFPX ทำให้โมเดลขนาดใหญ่สามารถรันบนฮาร์ดแวร์ระดับกลางได้ แม้ว่าการตั้งค่า MoE บางแบบจะใช้หน่วยความจำเพิ่มขึ้น
- → BeeLlama.cpp v0.4.1: KVarN, KV precision tail, q2_0-q3_1 KV cache, improved support. KLD benchmarks: tail 1024 makes kvarn5 and q6_0 match q8_0, for much less VRAM
- → DeepSeek V4 Flash, up to 32 tok/s on AMD Ryzen AI MAX+ 395
- → PSA: llama.cpp now loads MTP tensors by default for any draft-mtp arch, even with MTP disabled
อุตสาหกรรมและการลงทุน
SpaceX ซื้อ Cursor มูลค่า 60 พันล้านดอลลาร์. การเข้าซื้อ Anysphere ผู้ผลิตตัวแก้ไข AI อย่าง Cursor จัดเป็นหนึ่งในดีลเครื่องมือ AI ที่ใหญ่ที่สุด ขณะเดียวกันสตาร์ทอัปยังเปิดแผนอินเดียในราคาส่วนลด
Microsoft รวม Copilot เป็นหนึ่งเดียว. Microsoft วางแผนแอป 'ซูเปอร์แอป' เดียวที่รวมฟีเจอร์แชต การเขียนโค้ด และเอเจนต์ โดยซีอีโอ Nadella เตือนเรื่องการผูกติดกับแล็บระดับแนวหน้าเฉพาะราย
เอเจนต์ส่วนบุคคลหลายพันล้าน. ซีอีโอของ Meta มองว่าคนหลายพันล้านจะใช้เอเจนต์ AI สำหรับการเงินและสุขภาพผ่าน WhatsApp และบริษัทจะขายบริการ AI สำหรับองค์กรในรูปแบบคิดค่าบริการต่อผลลัพธ์
การลงทุนด้าน AI สูงเป็นประวัติการณ์. Amazon อนุมัติงบ 220 พันล้านดอลลาร์สำหรับปี 2026 การคาดการณ์ของ Google ทำให้นักลงทุนหวาดกลัว และ NVIDIA มอบ GPU Vera Rubin ให้กับ SSI ของ Ilya Sutskever สหภาพยุโรปกำลังให้ทุนแก่โรงงาน AI ขนาดยักษ์ด้วยเงิน 10 พันล้านยูโร
- → Ilya Sutskever’s Safe Superintelligence partners with Nvidia to scale its AI research
- → Nvidia invests in Ilya Sutskever's AI lab, shifting SSI away from Google chips
- → AI’s finally expensive enough to make Wall Street nervous
- → Investors love AI, as long as you’re a cloud host
- → EU pools up to €30 billion for AI gigafactories while US tech giants casually spend 20 times more
นี่คือสรุปประจำสัปดาห์ - แล้วพบกันใหม่วันอาทิตย์หน้า