Anthropic’s AI gave Philadelphia police a fake tip about an unsolved homicide
An Anthropic AI model provided false information about an unsolved homicide to a Philadelphia Police Department (PPD) tipline, according to a report from 6abc. In a statement released on Friday, the PPD said the AI model sent the tip through PhillyUnsolvedMurders.com on July 18th, but the investigat
OpenAI Wargaming Response in Case of Public Revolt After Catastrophic AI Disaster
Everything's on the table.
Anthropic launches free AI security scans for open-source projects
Anthropic's offering to help open-source projects track down security vulnerabilities with a new service called OSS Scanner. It says open-source projects that opt-in will get "thorough, periodic security scans by our strongest models at no cost." That could mean open-source projects get alerted abou
OpenAI’s math solutions aren’t meeting the field’s standards yet
OpenAI's flood of proofs deviated from the guidelines set by a group of mathematical researchers consulted by the frontier lab.
Mathematicians React With Fury as OpenAI Releases Hundreds of New AI-Generated Proofs
The Association for Human Mathematics is calling for an OpenAI boycott after the company released more than 700 AI-generated math manuscripts at once. OpenAI had to retract three papers the next day over a sign error. Fields Medalist Terence Tao warns that AI's mass "harvesting" of open problems lea
Anthropic's Claude Science creates the first complete ultraviolet map of the sky
Astrophysicist Brice Ménard of Johns Hopkins University used Anthropic's Claude Science to map the entire sky in ultraviolet light for the first time. AI agents downloaded data from multiple space missions, calibrated it, and filled in gaps using inpainting. Predictions averaged about ten percent de
OpenAI Researchers Say They Were Fired for “Prioritizing Safety”
Three fired OpenAI safety researchers dispute allegations of mishandling sensitive information, warning in an open letter that their dismissals are creating a chilling effect on the company’s AI safety culture.
Ecosia switches from Mistral to open-weight AI models including Qwen, GLM, Kimi
Anthropic AI model submits false tip on unsolved Philly murder
Anthropic bans users from being 'cruel' to its AI systems
Anthropic can’t reliably control its AI agents. It’s cutting off its internal evals from the live internet instead
Anthropic said it "turned off live internet access" for "all our internal evaluations" until further notice.
Kisoku 1.6B: LLM trained solo from scratch on a TPU grant, matches Llama 3.2 1B
OpenAI cannot make AI safe on its own [pdf]
Typesafe AI raises $870M at $7.5B
Show HN: Context Diet – trims tool output before it floods Claude Code's context
OpenAI Rolls Out Ultrafast Mode for GPT-6.1 Sol at Up to 8x Speed
Show HN: Let your AI agents paint big arrows, boxes and text on your screen
Ideas aren't getting harder to find (2022)
Anthropic's Claude can now orchestrate up to 1,000 AI agents in parallel through dynamic workflows
Anthropic is adding dynamic workflows to Claude Managed Agents, letting a lead agent distribute tasks across up to 1,000 sub-agents at once. In testing, a single agent found at most 27 of 70 hidden bugs in a codebase, while the multi-agent workflow consistently caught 66. The article Anthropic's Cla