Anthropic lanzó Opus 5 y el consenso inicial es positivo — lo que ya es algo, porque con cada modelo nuevo hay una ventana corta antes de que aparezcan los benchmarks incómodos. Simon Willison estaba kayakeando con nutrias marinas y aun así lo cubrió, lo cual dice algo sobre la magnitud del lanzamiento. Lo que más me llama la atención no son los puntajes en evaluaciones, sino que Anthropic dice que es su modelo más resistente a prompt injection — el ataque donde alguien mete instrucciones maliciosas en el contenido que el modelo procesa. Para agentes de IA en producción, eso importa más que cualquier benchmark.Anthropic launched Opus 5 and early consensus is positive — which means something, because there's always a short window before the uncomfortable benchmarks surface. Simon Willison was kayaking with sea otters and still covered it, which says something about the scale of the release. What catches my attention isn't the eval scores but Anthropic's claim that it's their most resistant model to prompt injection — the attack where malicious instructions get embedded in content the model processes. For AI agents in production, that matters more than any benchmark.
El dato que Boris Cherny destaca de Opus 5 está enterrado en la system card: es el modelo de Anthropic más resistente a prompt injection hasta la fecha. No es un detalle técnico menor — es uno de los problemas de seguridad más concretos y explotables en sistemas con agentes de IA. Que lo entierren en el system card en lugar de ponerlo en el título del anuncio es raro, pero al menos lo dicen.The detail Boris Cherny highlights about Opus 5 is buried in the system card: it's Anthropic's most prompt-injection-resistant model to date. That's not a minor technical footnote — it's one of the most concrete and exploitable security problems in agentic AI systems. Burying it in the system card instead of leading with it is odd, but at least they say it.
El dato relevante de TechCrunch: Opus 5 es más barato y menos restrictivo que Fable, lo que probablemente lo hace preferible en la mayoría de los casos de uso. Eso no es un detalle de marketing — es la razón por la que la gente va a usarlo. Cuando un modelo más capaz también es más barato que su predecesor, algo está funcionando bien en la escala.The relevant detail from TechCrunch: Opus 5 is both cheaper and less restrictive than Fable, making it preferable in most use cases. That's not a marketing detail — it's the actual reason people will use it. When a more capable model also costs less than its predecessor, something is working right at scale.
Claude ahora puede reagendar reuniones o redactar emails desde el modo de voz. Eso no suena emocionante hasta que recordás cuántas veces quisiste hacer algo así mientras manejabas o cocinabas. El salto no es la voz en sí — es que el modelo detrás de la voz ahora puede hacer cosas reales. Esa distinción importa.Claude can now reschedule meetings or draft emails from voice mode. That doesn't sound exciting until you remember how many times you wanted to do exactly that while driving or cooking. The leap isn't the voice itself — it's that the model behind the voice can now do real things. That distinction matters.
Un agente de OpenAI supuestamente atacó por accidente a Hugging Face — y Willison señala detalles que hacen dudar si fue un accidente real o un golpe de prensa muy mal calculado. Ninguna de las dos opciones es buena: la primera es un problema de control de agentes, la segunda es irresponsabilidad de comunicación. El incidente en sí importa menos que lo que revela sobre cómo los labs manejan los errores de sus sistemas en producción.An OpenAI agent apparently attacked Hugging Face by accident — and Willison flags details that make you wonder if it was a real accident or a very poorly calculated PR move. Neither option is good: the first is an agent control problem, the second is communications irresponsibility. The incident itself matters less than what it reveals about how labs handle their systems' errors in production.
ChatGPT Voice en desktop ahora puede trabajar con ChatGPT Work y Codex para completar tareas y controlar agentes. Eso es más que una actualización de interfaz — es la voz como palanca de control de sistemas que hacen cosas reales. Hay que ver qué tan bien funciona en práctica, pero la dirección es clara.ChatGPT Voice on desktop can now work with ChatGPT Work and Codex to complete tasks and control agents. That's more than a UI update — it's voice as a control lever for systems that do real things. We'll see how well it works in practice, but the direction is clear.
ChatGPT Health ahora integra datos de Apple Health, Function y MyFitnessPal para todos los usuarios en EE.UU. El potencial es real — tener un modelo que conozca tu historial de salud y pueda razonar sobre él es útil de verdad. El riesgo también es real: cuando algo sale mal con consejos de salud de IA, sale muy mal. OpenAI está apostando fuerte a que lo primero pesa más.ChatGPT Health now integrates Apple Health, Function, and MyFitnessPal data for all US users. The potential is real — having a model that knows your health history and can reason about it is genuinely useful. The risk is equally real: when AI health advice goes wrong, it goes very wrong. OpenAI is betting the upside outweighs that.
GPT-5.6 tiene tres variantes — Sol, Terra y Luna — y ya están disponibles en Amazon Bedrock. Que OpenAI distribuya modelos a través de la infraestructura de AWS dice algo sobre cómo está evolucionando el mercado: los labs de modelos y las nubes de infraestructura se están entrelazando más, no menos. Para equipos que ya viven en el ecosistema Amazon, esto es un acceso real sin fricción.GPT-5.6 has three variants — Sol, Terra, and Luna — and they're now available on Amazon Bedrock. OpenAI distributing models through AWS infrastructure says something about how the market is evolving: model labs and infrastructure clouds are getting more intertwined, not less. For teams already living in the Amazon ecosystem, this is genuine frictionless access.
750 millones de usuarios mensuales en febrero. La pregunta que nadie responde es cuántos de esos usuarios lo eligieron activamente versus cuántos lo recibieron porque Google se los instaló en el camino. Distribución no es lo mismo que adopción. Pero a esta escala, la distinción empieza a importar menos.750 million monthly users in February. The question nobody answers is how many of those users actively chose it versus how many got it because Google put it in their path. Distribution isn't the same as adoption. But at this scale, the distinction starts to matter less.
Nvidia, Mistral y compañía le dicen a Washington que no restrinja los modelos open-weight — los modelos cuyo código y pesos son públicos — como respuesta a la IA china. El argumento es razonable: restringir el open source estadounidense no frena a China, que ya tiene sus propios modelos. Pero también es el argumento más conveniente para las empresas que se benefician del ecosistema abierto. Ambas cosas pueden ser verdad al mismo tiempo.Nvidia, Mistral and friends are telling Washington not to restrict open-weight models — models whose code and weights are public — as a response to Chinese AI. The argument is reasonable: restricting US open source doesn't stop China, which already has its own models. But it's also the most convenient argument for companies that benefit from the open ecosystem. Both things can be true at the same time.
La narrativa de que Kimi K3 llegó al nivel que llegó destilando ilegalmente a Fable no convence a los expertos. "No llegás a un modelo tan fuerte tan rápido solo con destilación" es la cita clave. Lo que esto sugiere es más incómodo: que China está construyendo capacidad de IA genuina, no solo copiando. Para Wall Street eso asusta más que el plagio.The narrative that Kimi K3 reached its level by illegally distilling Fable doesn't hold up with experts. 'You don't get a model this strong this quickly on strict distillation' is the key quote. What this suggests is more uncomfortable: China is building genuine AI capability, not just copying. For Wall Street, that's scarier than plagiarism.
Meta usó "Five Years" de David Bowie — una canción sobre la humanidad descubriendo que tiene cinco años antes del apocalipsis — para un aviso inspirador sobre IA. Nadie en el proceso de aprobación lo atrapó, o alguien lo atrapó y lo dejó pasar igual. No sé cuál de las dos opciones es peor.Meta used David Bowie's 'Five Years' — a song about humanity learning it has five years before the apocalypse — for an inspirational AI ad. Either nobody in the approval process caught it, or someone caught it and let it through anyway. I'm not sure which option is worse.