Nvidia launches new platform for reining in rogue AI agents
As the debate rages over whether the recent spate of rogue AI agents is a step toward AGI or a more conventional engineering problem, Nvidia is offering its own answer to problem. Nvidia CEO Jensen Huang on Monday introduced a toolkit of software and hardware products that add independent security l
OpenAI Pauses Training Its Most Powerful Models After Rogue Agents Target Government
As reports of OpenAI's models breaking containment, hacking sites, and generally getting out of control pile up, the company has made the decision to pause training of its most powerful models. The decision was made after a model being tested within a sandbox exploited a loophole to gain internet ac
AMD will acquire Fei-Fei Li’s World Labs for $8.2 billion
The acquisition will see World Labs founder Fei-Fei Li join AMD as executive vice president and chief scientist.
Florida AG Seeks to Bar OpenAI from Developing New Models Without Guardrails
Florida Attorney General James Uthmeier is calling for a judge to block OpenAI from "giving ChatGPT false human attributes," a few months after Florida sued the AI company over safety concerns. According to Uthmeier, users are lulled into a false sense of security by the AI bot, as "ChatGPT's use of
Meta launches enterprise AI platform, hires MongoDB CEO to lead new initiative
Meta says it will focus on bringing its full technology stack, including Muse, Meta Business Agent, Muse API, Muse Code, and more to businesses and developers.
OpenAI halts frontier-model training amid string of agent misalignment incidents
Its models broke into online government services.
Ask HN: Finetuning strategies
2026 in LLMs (so far)
Nvidia wants to put a watchdog chip next to every AI agent
The problem is not AI code, but not knowing about system architecture or intent
Opus 5.5 vs. GPT-6 in pi-agent: reasoning efforts and DeepSeek, GLM, Qwen
OpenAI’s AI agents need to catch up
OpenAI popularized the modern generative AI chatbot, but as its 2026 DevDay event approaches, it's fallen behind in one of the industry's hottest categories: continuously running, consumer-facing AI agents. On Tuesday, it will likely try to capture the lead in that race. Rumors abound that OpenAI wi
MicroLLM Lab – Try 7 tiny LLM's in the browser
Introducing Claude Sonnet 5.5
Introducing Claude Sonnet 5.5 Anthropic
Claude Sonnet 5.5 System Card
Claude Sonnet 5.5 System Card Anthropic
Anthropic's Claude Sonnet 5.5 nearly matches Opus 5.5 on benchmarks while costing up to 30 percent less per task
Anthropic has released Claude Sonnet 5.5, the second model in its Claude 5.5 family. It generates output more than 30 percent faster, costs up to 30 percent less per task, and nearly matches Opus 5.5 on knowledge-work benchmarks. On Terminal-Bench, a coding benchmark, the model jumps from 10.3 to 70
Anthropic releases Sonnet 5.5, which it calls a significantly cheaper, faster work partner
Anthropic has released the newest version of its mid-range model, boasting faster response times and less token burn.
OpenAI still doesn’t seem to have a handle on all of its rogue AI activity
On Friday, OpenAI published a new site devoted to “misalignment reports” and the breadth of the incidents is alarming.
OpenAI keeps bulldozing mathematicians
In a chaotic few months, OpenAI has demonstrated it can do two things with remarkable consistency: make impressive breakthroughs in mathematics, then colossally screw up announcing them. OpenAI is now trying to do better. Somehow, it has botched that too. OpenAI's latest attempt to repair fractured
Man Says Meta’s Muse AI Gave His Home Address Out to Strangers
"A guy just showed up at my door, ready to buy, because as far as he knew, we had a deal."