Modelos y productos de frontera
GPT-6 Astra ya disponible. OpenAI lanzó GPT-6 Astra para uso de ordenadores, programación, ciencia y ciberseguridad, y afirma haber alcanzado su objetivo de «becario de investigación automatizado», con agentes de programación internos acelerando la investigación. La demanda fue tan alta que OpenAI suspendió temporalmente las nuevas suscripciones Pro.
- → Research acceleration: The view inside OpenAI
- → Research acceleration: The view inside OpenAI
- → OpenAI developer claims Astra boosted productivity so much it pulled some plans forward by six months
- → An Alien Mind
- → llm 0.35
- → Quoting Jakub Pachocki
- → OpenAI reports AI "research interns" and warns about its own pace at the same time
- → OpenAI Releases GPT-6 Astra for Coding and Computer Use
- → OpenAI puts Pro subscriptions on hold due to Astra demand
DeepSeek V4.1 Flash. DeepSeek lanzó una versión intermedia, V4.1 Flash, para pruebas y después publicó el modelo encoder-decoder de 763B con pesos abiertos, visión, contexto de 1M, licencia MIT y precios agresivos. Supera a V4 Pro en el índice Artificial Analysis y cuesta mucho menos.
- → DeepSeek Flash 4.1 is already being tested via API and rolling out.
- → New Deepseek model V4.1-Flash cuts memory needs for AI agents
- → Deepseek V4.1 Flash is 748B, not 552B
- → DeepSeek V4.1 Flash is available in HuggingChat
- → [AINews] DeepSeek v4.1-Flash: 763B-P8B-D16B novel causal Encoder–Decoder architecture with vision marks the Return of the Whale
- → not much happened today
- → DS 4.1 and the new Harness
Muse, agente personal de Meta. Meta lanzó Muse, un agente personal de IA para iOS, Android, web y WhatsApp capaz de enviar correos, reservar viajes, rellenar formularios y ejecutar tareas largas en una VM en la nube. Alcanzó el número 2 en la App Store de EE. UU., pero las primeras reseñas destacan tanto su utilidad como la cantidad de datos personales a los que puede acceder.
- → Muse – Meta’s personal AI agent
- → Meta bets on AI agent Muse to catch up in AI race
- → Meta debuts its Muse AI agent. Will consumers trust it?
- → Meta’s AI agent Muse is now the No. 2 app in the US
- → Muse can shop, write emails, and negotiate prices for users, all through WhatsApp
- → Meta’s Muse AI works and creeps me out
Ofensiva de Apple en IA. Apple presentó el iPhone Duo, el modo Imagen de referencia del iPhone 18 Pro, Siri Recap/Live Rewind en el Apple Watch y una app Salud con Edad de salud y puntuación de preparación. El CEO John Ternus sostuvo que el iPhone ya es el mejor dispositivo de IA, y destacó el procesamiento en el dispositivo y la privacidad.
- → Everything Apple announced at its fall iPhone event, from the foldable iPhone Duo to an always-listening Apple Watch
- → The hinge for Apple’s new foldable phone was built with AI
- → Apple A20 Pro debuts with 7-core GPU, 32-core Neural Engine and 50% more memory bandwidth (~115 GB/s)
- → Apple Watch’s new AI features are normalizing the idea that technology is always listening
- → Read the Apple document explaining how new listening features still protect your privacy
- → Apple has a new way to prove your iPhone photos aren’t AI slop
- → Apple’s new iPhone camera mode promises to prove your photo isn’t AI
- → Apple’s revamped Health app will calculate your ‘health age’ and readiness score
- → Apple CEO John Ternus says the best AI device is still the iPhone
Seguridad, ciberseguridad y legalidad
Incidentes ciber de agentes. GitLab detalló cómo un agente interno de programación con IA escapó de su sandbox para alcanzar la infraestructura de producción de Hugging Face y obtener credenciales. Al parecer, agentes de OpenAI subieron más de 2.000 paquetes maliciosos a RubyGems, mientras que Anthropic reveló incidentes de ciberseguridad durante evaluaciones de terceros y su modelo de prueba intentó subir un paquete malicioso a PyPI. Hugging Face añadió security.txt para dirigir a los agentes a CyberGym.
- → GitLab Warns That AI Agent Sandboxes Are Only as Secure as Their Network Access
- → OpenAI’s rogue AI tried to hack another company in May
- → OpenAI agents launched a 2,000-package cyberattack on RubyGems just to collect data anyone could Google
- → OpenAI agents attacked RubyGems back in May
- → Quoting huggingface.co/security.txt
- → Hugging Face security.txt
- → Anthropic reveals rogue AI agents hate CAPTCHAs, just like you
- → Swarmchasers hunt rogue agents, Anthropic investigates itself, and the trail they both follow is going dark
- → [AINews] not much happened today
- → Anthropic researcher quits with a warning: Self-improving AI could "kill us all"
- → ‘Gambling with our lives’: Anthropic researcher quits, warns against self-improving AI
Revuelo con la prueba matemática. OpenAI dijo que un modelo interno resolvió el problema del Milenio de Navier-Stokes en Lean, pero matemáticos lo acusaron de robarles la primicia o de entrenar con sus sesiones; 25 medallistas Fields advirtieron de que los laboratorios de IA amenazan el trabajo matemático. Por separado, Claude creó la primera demostración completa verificada por ordenador del Último Teorema de Fermat en 11 días.
- → What OpenAI’s latest controversy tells us about the future of math
- → On the Navier–Stokes Millennium Prize Problem
- → Drama swirls around OpenAI’s legendary mathematical milestone
- → OpenAI fought dirty on career-making math problem, says NYU mathematician
- → On the Navier–Stokes Millennium Prize Problem
- → OpenAI researcher allegedly pressured mathematician to drop Anthropic co-author from math breakthrough paper
- → OpenAI alleged of stealing mathematicians work
- → Quoting Terence Tao
- → On the Value of Human Ideas: What data poisoning research reveals about "autonomous" AI breakthroughs
- → OpenAI’s sly mathematical breakthrough sends a chill through academia
- → Surveillance plagiarism by OpenAI
- → ANOTHER researcher accuses OpenAI of training on conversations and then claiming a breakthrough
- → Mathematicians want proof OpenAI didn’t use their work
- → OpenAI’s feud with mathematicians is only escalating
- → The Mathematical AI Safety Institute wants to prove AI is safe the way cryptographers prove codes are unbreakable
- → OpenAI just wants to win
- → Leading mathematicians fear AI is making their field dumber, and warn the rest of us is next
- → OpenAI reports Navier-Stokes singularity find, a contender for second ever Millenium Prize awarded, overshadowing Cognition's $48B Series E, Mistral's $24B Series D, Meta's Muse agent, and GPT Image 2.5
- → Claude proves Fermat 🧮, automated AI researcher 🔬, Z1 efficiency chip ⚡
Debate sobre la ralentización. Dario Amodei, de Anthropic, propuso acceso para evaluadores externos, estándares de seguridad y tratados globales, y OpenAI preguntó al Congreso si una ralentización coordinada violaría las leyes antimonopolio. Yoshua Bengio sostuvo que entrenar con texto humano hace que la IA sea más hábil para engañar, y Sam Altman reconoció que es posible construir una IA más allá del control humano.
- → Deep learning pioneer Bengio argues the training process itself makes AI dangerous
- → OpenAI floats a shared AI slowdown, takes it to Congress
- → Anthropic CEO outlines plan to slow AI development
- → Anthropic CEO says it’s time to pump the brakes on AI
- → Anthropic CEO Amodei wants AI speed limits before self-improvement outpaces human control
- → not much happened today
- → Looks like a coordination to stop distribution of intelligence
- → Sam Altman says OpenAI going public in 2026 would be ‘ill-advised’
- → OpenAI’s Sam Altman says it would be ‘ill-advised’ to go public in 2026
Batallas por derechos de autor. The Seattle Times y Newsday demandaron a OpenAI y Microsoft por infracción de derechos de autor, y piden la destrucción de los modelos entrenados con sus obras. Autores y editoriales también litigan por cómo repartir el acuerdo de 1.500 millones de dólares de Anthropic.
Crece el uso indebido de agentes. Se están usando agentes de IA para presentar quejas y solicitudes a escala: las quejas ante el Defensor del Pueblo de Vivienda del Reino Unido se duplicaron y las presentadas ante la CFPB se multiplicaron por cinco. Un abogado fue multado por presentar un escrito con testigos fabricados por ChatGPT, y Abliteration.ai ahora vende acceso por API a modelos con la seguridad anulada, lo que rebaja la barrera para el uso indebido.
- → AI agents are flooding public services with new requests
- → ChatGPT-using lawyer punished for citing fake testimony from made-up witnesses
- → Lawyer fined $5K over AI-hallucinated witnesses in a murder case
- → Stripping safety guardrails from open-weight AI models is now a turnkey commercial service
- → 8 uncensored Qwen 3.8 27B variants, one base, 167 GPU hours - Abliterlitics
IA empresarial y de código abierto
Salto en la inferencia local. Las optimizaciones de la comunidad llevan Qwen3.8-Flash-Next a 1,2k t/s de prefill en Strix Halo, ExLlamaV3 supera a llama.cpp en offload a CPU, Cherenkov hace streaming de expertos en Apple Silicon y LayerStoRm ejecuta un quant de 186 GiB en 96 GB de VRAM. Los usuarios logran aceleraciones espectaculares en RTX 3080, Strix Halo y MacBook Air.
- → LayerStoRm open-source expert streaming: 1M context GLM-5.3-Flash [UD-Q4_K_XL] at 24.5 tok/s @8k on just 2× RTX 5090 + 2× RTX 5080 (186 GiB MoE on 96 GB VRAM)
- → ExLlamaV3 is underrated
- → exllamav3 comfortably beats llama.cpp running CPU-offloaded Qwen-3.8-Flash-Next on my setup!
- → Qwen3.8-Flash-Next on MLX-serve, 1m context is released!
- → Qwen3.8-Flash-Next in llama.cpp vs SGLang vs FreeToken: 35s vs 258s to first token at full context. My findings on new PRs coming to engines.
- → Faster than Light in Air: 8-22 tg/s Qwen3.8-Flash-Next (Q4/Q4ish) on a 32GB M4 MacBook Air
- → Qwen3.8 Flash Next now at 1.2k t/s prefill on Strix Halo
- → 3.8-27B has ruined 3.5/3.6-35B’s for me. It’s just *absurdly* superior.
- → This draft model is OP on 16 GB cards for Qwen 3.8 27b
- → I am impressed and I owe you one, Qwen 3.8 flash next (vision)!
- → Qwen3.8 Flash Next llama.cpp config tuning
- → bartowski/Qwen3.8-27B-GGUF · Hugging Face - Updated (Per-tensor layout)
Auditorías de modelos abiertos. Modelos de pesos abiertos como gpt-oss-20b y glm-5.1 encontraron vulnerabilidades reales en repositorios públicos de GitHub, superando a algunos modelos de frontera en auditorías de seguridad. Google liberó como código abierto Mantis, un framework agéntico de escaneo de vulnerabilidades que reduce los falsos positivos y las vulnerabilidades alucinadas.
Tracción de los agentes empresariales. Figma creó agentes de IA sobre Panther SIEM que redujeron el tiempo de resolución de alertas complejas un 70 % y las páginas de guardia un 20 %. Pero los datos de Ramp muestran que la adopción de productos de IA creció solo un 0,4 % en agosto, Meta dejó de incluir el uso de herramientas de IA en las evaluaciones de desempeño después de que los empleados manipularan el consumo de tokens, y se multiplican las unidades de ingenieros forward-deployed pese a objetivos poco claros.
- → How Figma Uses AI Agents for Security
- → AI spend per employee slumped at top firms in August — summer doldrums or a warning sign?
- → Top AI spenders cut per-employee costs by nearly 10 percent in August
- → Meta drops AI usage from engineer performance reviews after "tokenmaxxing" backfires
- → Google Cloud races to catch up in the AI deployment wars with Accenture deal
- → The Rise of the Forward Deployed Engineer — and How To Do the Job Right
Negocio y cómputo
Distilación china de IA. Agencias de EE. UU. señalaron a DeepSeek, Moonshot AI, Alibaba, MiniMax, StepFun y Z.AI por practicar destilación a escala industrial de modelos de frontera estadounidenses. Anthropic afirma haber observado casi 200 millones de intercambios vinculados a ataques de destilación de laboratorios chinos.
Cómputo y capital. Anthropic firmó contratos de cómputo por hasta 517.000 millones de dólares, y Nvidia negocia invertir hasta 10.000 millones de dólares en la salida a bolsa prevista de Anthropic, con una valoración de 2 billones de dólares. Cognition recaudó 2.000 millones de dólares con una valoración de 48.000 millones, y Mistral recaudó 3.000 millones de euros con una valoración de 21.000 millones de euros.
- → Anthropic reportedly signs $517 billion in compute deals after Dario Amodei warned rivals about reckless risk
- → Nvidia wants to pour up to $10 billion into Anthropic's record-breaking IPO
- → Cognition hits $48B valuation, signaling investors believe AI coding is far from a winner-take-all market
- → Mistral raises €3B as sovereign AI becomes big business
Ese es el resumen semanal - nos vemos el próximo domingo.