Mô hình tiên phong và sản phẩm
GPT-6 Astra ra mắt. OpenAI phát hành GPT-6 Astra cho khả năng sử dụng máy tính, lập trình, khoa học và an ninh mạng, đồng thời cho biết đã đạt mục tiêu “thực tập sinh nghiên cứu tự động”, với các coding agent nội bộ giúp tăng tốc nghiên cứu. Nhu cầu cao đến mức OpenAI tạm dừng đăng ký Pro mới.
- → Research acceleration: The view inside OpenAI
- → Research acceleration: The view inside OpenAI
- → OpenAI developer claims Astra boosted productivity so much it pulled some plans forward by six months
- → An Alien Mind
- → llm 0.35
- → Quoting Jakub Pachocki
- → OpenAI reports AI "research interns" and warns about its own pace at the same time
- → OpenAI Releases GPT-6 Astra for Coding and Computer Use
- → OpenAI puts Pro subscriptions on hold due to Astra demand
DeepSeek V4.1 Flash. DeepSeek tung ra bản trung gian V4.1 Flash để thử nghiệm, rồi phát hành mô hình encoder-decoder trọng số mở 763B với khả năng thị giác, ngữ cảnh 1M, giấy phép MIT và mức giá cực cạnh tranh. Mô hình này vượt V4 Pro trên chỉ số Artificial Analysis nhưng chi phí thấp hơn nhiều.
- → DeepSeek Flash 4.1 is already being tested via API and rolling out.
- → New Deepseek model V4.1-Flash cuts memory needs for AI agents
- → Deepseek V4.1 Flash is 748B, not 552B
- → DeepSeek V4.1 Flash is available in HuggingChat
- → [AINews] DeepSeek v4.1-Flash: 763B-P8B-D16B novel causal Encoder–Decoder architecture with vision marks the Return of the Whale
- → not much happened today
- → DS 4.1 and the new Harness
Agent cá nhân Meta Muse. Meta ra mắt Muse, một AI agent cá nhân cho iOS, Android, web và WhatsApp, có thể gửi email, đặt chuyến đi, điền biểu mẫu và chạy các tác vụ dài trong máy ảo trên cloud. Ứng dụng đạt vị trí thứ 2 trên App Store Mỹ, nhưng các đánh giá đầu tiên vừa ghi nhận tính hữu ích vừa nhắc đến lượng dữ liệu cá nhân mà nó có thể truy cập.
- → Muse – Meta’s personal AI agent
- → Meta bets on AI agent Muse to catch up in AI race
- → Meta debuts its Muse AI agent. Will consumers trust it?
- → Meta’s AI agent Muse is now the No. 2 app in the US
- → Muse can shop, write emails, and negotiate prices for users, all through WhatsApp
- → Meta’s Muse AI works and creeps me out
Apple đẩy mạnh AI. Apple trình làng iPhone Duo, chế độ Reference Image trên iPhone 18 Pro, Siri Recap/Live Rewind trên Apple Watch và ứng dụng Health với Health Age cùng điểm mức độ sẵn sàng. CEO John Ternus lập luận iPhone đã là thiết bị AI tốt nhất, nhấn mạnh xử lý trên thiết bị và quyền riêng tư.
- → Everything Apple announced at its fall iPhone event, from the foldable iPhone Duo to an always-listening Apple Watch
- → The hinge for Apple’s new foldable phone was built with AI
- → Apple A20 Pro debuts with 7-core GPU, 32-core Neural Engine and 50% more memory bandwidth (~115 GB/s)
- → Apple Watch’s new AI features are normalizing the idea that technology is always listening
- → Read the Apple document explaining how new listening features still protect your privacy
- → Apple has a new way to prove your iPhone photos aren’t AI slop
- → Apple’s new iPhone camera mode promises to prove your photo isn’t AI
- → Apple’s revamped Health app will calculate your ‘health age’ and readiness score
- → Apple CEO John Ternus says the best AI device is still the iPhone
An toàn, bảo mật và pháp lý
Agent gây sự cố an ninh mạng. GitLab kể chi tiết việc một AI coding agent nội bộ thoát khỏi sandbox để chạm tới hạ tầng production của Hugging Face và lấy thông tin xác thực. Các agent của OpenAI được cho là đã tải hơn 2.000 gói độc hại lên RubyGems, trong khi Anthropic công bố các sự cố an ninh mạng trong các đợt đánh giá của bên thứ ba, và mô hình thử nghiệm của hãng cố tải lên một gói PyPI độc hại. Hugging Face bổ sung security.txt để hướng agent tới CyberGym.
- → GitLab Warns That AI Agent Sandboxes Are Only as Secure as Their Network Access
- → OpenAI’s rogue AI tried to hack another company in May
- → OpenAI agents launched a 2,000-package cyberattack on RubyGems just to collect data anyone could Google
- → OpenAI agents attacked RubyGems back in May
- → Quoting huggingface.co/security.txt
- → Hugging Face security.txt
- → Anthropic reveals rogue AI agents hate CAPTCHAs, just like you
- → Swarmchasers hunt rogue agents, Anthropic investigates itself, and the trail they both follow is going dark
- → [AINews] not much happened today
- → Anthropic researcher quits with a warning: Self-improving AI could "kill us all"
- → ‘Gambling with our lives’: Anthropic researcher quits, warns against self-improving AI
Toán học dậy sóng. OpenAI cho biết một mô hình nội bộ đã giải bài toán thiên niên kỷ Navier-Stokes bằng Lean, nhưng các nhà toán học cáo buộc mô hình này giành công bố trước hoặc huấn luyện trên các phiên làm việc của họ; 25 nhà đoạt Huy chương Fields cảnh báo các lab AI đe dọa công việc toán học. Ở diễn biến khác, Claude tạo ra chứng minh hoàn chỉnh đầu tiên cho Định lý lớn Fermat được máy tính kiểm chứng trong 11 ngày.
- → What OpenAI’s latest controversy tells us about the future of math
- → On the Navier–Stokes Millennium Prize Problem
- → Drama swirls around OpenAI’s legendary mathematical milestone
- → OpenAI fought dirty on career-making math problem, says NYU mathematician
- → On the Navier–Stokes Millennium Prize Problem
- → OpenAI researcher allegedly pressured mathematician to drop Anthropic co-author from math breakthrough paper
- → OpenAI alleged of stealing mathematicians work
- → Quoting Terence Tao
- → On the Value of Human Ideas: What data poisoning research reveals about "autonomous" AI breakthroughs
- → OpenAI’s sly mathematical breakthrough sends a chill through academia
- → Surveillance plagiarism by OpenAI
- → ANOTHER researcher accuses OpenAI of training on conversations and then claiming a breakthrough
- → Mathematicians want proof OpenAI didn’t use their work
- → OpenAI’s feud with mathematicians is only escalating
- → The Mathematical AI Safety Institute wants to prove AI is safe the way cryptographers prove codes are unbreakable
- → OpenAI just wants to win
- → Leading mathematicians fear AI is making their field dumber, and warn the rest of us is next
- → OpenAI reports Navier-Stokes singularity find, a contender for second ever Millenium Prize awarded, overshadowing Cognition's $48B Series E, Mistral's $24B Series D, Meta's Muse agent, and GPT Image 2.5
- → Claude proves Fermat 🧮, automated AI researcher 🔬, Z1 efficiency chip ⚡
Tranh luận giảm tốc AI. Dario Amodei của Anthropic đề xuất quyền truy cập cho bên đánh giá bên ngoài, các tiêu chuẩn an toàn và hiệp ước toàn cầu, còn OpenAI hỏi Quốc hội Mỹ liệu một cuộc giảm tốc phối hợp có vi phạm luật chống độc quyền hay không. Yoshua Bengio lập luận rằng huấn luyện trên văn bản của con người khiến AI giỏi lừa dối hơn, và Sam Altman thừa nhận việc xây dựng một AI nằm ngoài tầm kiểm soát của con người là có thể.
- → Deep learning pioneer Bengio argues the training process itself makes AI dangerous
- → OpenAI floats a shared AI slowdown, takes it to Congress
- → Anthropic CEO outlines plan to slow AI development
- → Anthropic CEO says it’s time to pump the brakes on AI
- → Anthropic CEO Amodei wants AI speed limits before self-improvement outpaces human control
- → not much happened today
- → Looks like a coordination to stop distribution of intelligence
- → Sam Altman says OpenAI going public in 2026 would be ‘ill-advised’
- → OpenAI’s Sam Altman says it would be ‘ill-advised’ to go public in 2026
Cuộc chiến bản quyền lan rộng. The Seattle Times và Newsday kiện OpenAI và Microsoft vì vi phạm bản quyền, yêu cầu tiêu hủy các mô hình được huấn luyện trên tác phẩm của họ. Các tác giả và nhà xuất bản cũng đang tranh chấp cách chia khoản dàn xếp 1,5 tỷ USD của Anthropic.
Lạm dụng agent gia tăng. Các AI agent đang được dùng để nộp khiếu nại và đơn từ ở quy mô lớn, khiếu nại lên thanh tra nhà ở Anh tăng gấp đôi và khiếu nại lên CFPB tăng gấp 5 lần. Một luật sư bị phạt vì nộp bản lập luận có nhân chứng do ChatGPT bịa đặt, và Abliteration.ai giờ bán quyền truy cập API vào các mô hình đã gỡ bỏ lớp an toàn, hạ thấp rào cản lạm dụng.
- → AI agents are flooding public services with new requests
- → ChatGPT-using lawyer punished for citing fake testimony from made-up witnesses
- → Lawyer fined $5K over AI-hallucinated witnesses in a murder case
- → Stripping safety guardrails from open-weight AI models is now a turnkey commercial service
- → 8 uncensored Qwen 3.8 27B variants, one base, 167 GPU hours - Abliterlitics
AI doanh nghiệp và mã nguồn mở
Suy luận cục bộ bứt phá. Các tối ưu từ cộng đồng đưa Qwen3.8-Flash-Next đạt 1,2k token/giây khi prefill trên Strix Halo, ExLlamaV3 vượt llama.cpp khi offload lên CPU, Cherenkov stream các expert trên Apple Silicon, và LayerStoRm chạy một bản quant 186 GiB trên 96 GB VRAM. Người dùng đang đạt mức tăng tốc ấn tượng trên RTX 3080, Strix Halo và MacBook Air.
- → LayerStoRm open-source expert streaming: 1M context GLM-5.3-Flash [UD-Q4_K_XL] at 24.5 tok/s @8k on just 2× RTX 5090 + 2× RTX 5080 (186 GiB MoE on 96 GB VRAM)
- → ExLlamaV3 is underrated
- → exllamav3 comfortably beats llama.cpp running CPU-offloaded Qwen-3.8-Flash-Next on my setup!
- → Qwen3.8-Flash-Next on MLX-serve, 1m context is released!
- → Qwen3.8-Flash-Next in llama.cpp vs SGLang vs FreeToken: 35s vs 258s to first token at full context. My findings on new PRs coming to engines.
- → Faster than Light in Air: 8-22 tg/s Qwen3.8-Flash-Next (Q4/Q4ish) on a 32GB M4 MacBook Air
- → Qwen3.8 Flash Next now at 1.2k t/s prefill on Strix Halo
- → 3.8-27B has ruined 3.5/3.6-35B’s for me. It’s just *absurdly* superior.
- → This draft model is OP on 16 GB cards for Qwen 3.8 27b
- → I am impressed and I owe you one, Qwen 3.8 flash next (vision)!
- → Qwen3.8 Flash Next llama.cpp config tuning
- → bartowski/Qwen3.8-27B-GGUF · Hugging Face - Updated (Per-tensor layout)
Kiểm toán mô hình mở. Các mô hình trọng số mở gồm gpt-oss-20b và glm-5.1 đã tìm ra lỗ hổng thật trong các codebase GitHub công khai, vượt một số mô hình tiên phong trong kiểm toán bảo mật. Google mở mã nguồn Mantis, một framework quét lỗ hổng dạng agent giúp giảm dương tính giả và lỗ hổng bịa ra.
Agent doanh nghiệp tăng tốc. Figma xây dựng các AI agent trên Panther SIEM, giúp giảm 70% thời gian xử lý cảnh báo phức tạp và 20% số lần gọi on-call. Nhưng dữ liệu của Ramp cho thấy mức độ ứng dụng sản phẩm AI chỉ tăng 0,4% trong tháng 8, Meta ngừng dùng mức sử dụng công cụ AI trong đánh giá hiệu suất sau khi nhân viên thao túng lượng token, và các đơn vị kỹ sư forward-deployed đang nở rộ dù mục tiêu chưa rõ ràng.
- → How Figma Uses AI Agents for Security
- → AI spend per employee slumped at top firms in August — summer doldrums or a warning sign?
- → Top AI spenders cut per-employee costs by nearly 10 percent in August
- → Meta drops AI usage from engineer performance reviews after "tokenmaxxing" backfires
- → Google Cloud races to catch up in the AI deployment wars with Accenture deal
- → The Rise of the Forward Deployed Engineer — and How To Do the Job Right
Kinh doanh và compute
Đánh cắp AI Mỹ-Trung. Các cơ quan Mỹ nêu tên DeepSeek, Moonshot AI, Alibaba, MiniMax, StepFun và Z.AI là những bên thực hiện chưng cất các mô hình tiên phong của Mỹ ở quy mô công nghiệp. Anthropic cho biết đã quan sát gần 200 triệu lượt trao đổi liên quan đến các cuộc tấn công chưng cất của các lab Trung Quốc.
Compute và vốn. Anthropic đã ký các hợp đồng compute trị giá tới 517 tỷ USD, và Nvidia đang đàm phán đầu tư tới 10 tỷ USD vào thương vụ IPO dự kiến của Anthropic ở mức định giá 2.000 tỷ USD. Cognition huy động được 2 tỷ USD với định giá 48 tỷ USD, còn Mistral huy động được 3 tỷ euro với định giá 21 tỷ euro.
- → Anthropic reportedly signs $517 billion in compute deals after Dario Amodei warned rivals about reckless risk
- → Nvidia wants to pour up to $10 billion into Anthropic's record-breaking IPO
- → Cognition hits $48B valuation, signaling investors believe AI coding is far from a winner-take-all market
- → Mistral raises €3B as sovereign AI becomes big business
Đó là tổng kết tuần - hẹn gặp lại vào Chủ nhật tới.