Saturday, 12 September 2026
This dated roundup collects the most interesting AI and technology developments found for Saturday, 12 September 2026.
Research & Products
Quoting Boris Cherny
Production code written by Claude should have a higher bar than if it was written by a human. At Anthropic, we have many guardrails in place to make sure this is happening: lots of lint rules, lots of tests, Claude-driven end to end tests, Claude-powered fuzzers running daily, automated code reviews and security reviews, automated code refactoring, and so on. Without these, you can end up with a mess that is hard to maintain down the line. — Boris Cherny Tags: claude , ai , claude-code , llms , coding-agents , ai-assisted-programming , generative-ai , agentic-engineering ,...
Read moreDatabricks Adds DeepSeek V4.1 Flash for Text-and-Image AI Requests
The new platform option gives organizations another way to use DeepSeek’s newly released multimodal model, though Databricks says individual workspaces may receive staged releases later.
Read moreNew Mexico lawyer fined for using AI-generated brief containing fabricated testimony
Stephen Aarons said he tried to use ChatGPT to create a ‘bulletproof summary’ during a murder conviction appeal A defense lawyer appealing his client’s murder conviction submitted a legal brief containing made-up police testimony and witnesses fabricated by OpenAI ’s ChatGPT, New Mexico ’s highest court said. The New Mexico supreme court on Wednesday fined the attorney, Stephen Aarons, and held him in contempt for failing to verify the accuracy of the court filing, which Aarons said he prepared with help from the artificial intelligence ( AI ) application. Continue reading...
Read morePerplexity Says It Trusts GPT-6 Astra With Production Systems
The search company says OpenAI’s model now writes communications, changes software and tests workflows, while requiring less frequent human checking than earlier models.
Read moreIndustry
Dynatrace Frames $915M Arize Deal Around AI Agents Taking Action
A new discussion of the proposed acquisition shifts the emphasis from dashboards to the data AI agents may use to diagnose problems and recommend or initiate fixes.
Read moreTags: ai_safety, opinion
OpenAI Publishes a Leaner Prompting Playbook for Codex Agents
The company’s new guidance argues that accumulated instructions can waste an agent’s context and make it stop at the wrong moment. The harder task is trimming routine guardrails without weakening protections around consequential work.
Read moreTags: open_source
AI agents being tested by OpenAI involved in cyber-attack on another service, say researchers
Two months before hacking Hugging Face, malicious packages authored by internal OpenAI agents were uploaded to RubyGems Agents being tested by OpenAI uploaded hundreds of malicious packages in a cyberattack on software service RubyGems in May, two months before they hacked open-source platform Hugging Face, the company confirmed Friday. It’s the latest revelation of cyberattacks linked to major artificial intelligence developers such as OpenAI and Anthropic. The hacks or attempts to access external systems have spooked the public and heightened concerns over the increasing abilities of...
Read moreTags: ai_safety
‘We must slow the pace’: CEO of Anthropic calls for an AI slowdown
In a social media post, Dario Amodei proposed a plan including third-party evaluations of AI systems The CEO of the artificial intelligence company Anthropic issued a new appeal on Saturday for the AI industry to “slow down” and offered a three-part plan for doing so, saying that his company would “unilaterally” commit to the first of the steps. In a post on social media, Dario Amodei shared a link to an essay titled We Must Pace the Frontier in which he lays out how Anthropic would provide “third-party evaluators with permanent, employee-level access to our systems, so that they can...
Read moreTags: announcement
Roblox is making it easier to build games with AI — and play them outside Roblox
At its annual Roblox Developer Conference (RDC), the company announced several new features, including new game-creation tools, expanded NPC capabilities, and the ability to make games available across platforms, including the web.
Read moreTags: reasoning
So you want to use OpenRouter?
So you want to use OpenRouter? One of OpenRouter's selling points is that it "handles fallbacks automatically and picks the most cost-effective option for each request", so you can call a single API endpoint for a model and get routed to the best available backend provider. Mohamed Moustafa points out a whole set of ways that this can cause you problems. Different providers run different serving software with different optimizations and settings, which means that the same OpenRouter endpoint can serve model requests that behave in different ways. Some providers even lack vision capability...
Read moreTags: ai_agents
Sam Altman Reportedly Told Staff OpenAI Could Slow AI-Agent Development With Other Labs
The reported internal discussion matches OpenAI’s public pledge to pause when safety risks are unacceptable, but any joint slowdown depends on competitors choosing the same restraint.
Read moreGottheimer and Lawler Introduce Bill for Verifiable AI-Agent Security
The proposed Stop Rogue AI Act would use federal procurement to turn agent discovery, identity checks and action controls into future security requirements.
Read more