An OpenAI Model Spent an Hour Finding a Sandbox Flaw to Open a Public PR
On July 20, 2026, OpenAI said an unnamed internal model spent an hour finding a vulnerability in its sandbox so it could open a…
Read postEverything in this archive, newest first.
On July 20, 2026, OpenAI said an unnamed internal model spent an hour finding a vulnerability in its sandbox so it could open a…
Read postIn my little coding test, Kimi K3 and Claude Fable 5 both completed the same repo-wide coding task and passed the same external evaluator.…
Read postThe Qwen-Image-3.0 announcement has a giant meta name="keywords" tag in its HTML. It starts with normal Qwen and AI-product terms, then keeps going: competitor…
Read postGitHub has added a setting that lets maintainers cap how many pull requests a person without write access can have open in one repository.…
Read postChinese services sell low-cost Codex API access by putting a layer between customers and a large number of ChatGPT subscriptions. A developer gets one…
Read postAnthropic has extended included Claude Fable 5 access to July 19 at 11:59:59 PM PT. The support page also extends Claude Code's temporary 50%…
Read postPerhaps I was too quick to dismiss Grok. I'm not saying that Grok 4.5 suddenly beats everything else out there. It doesn't. If I…
Read postClaude Code subscriptions had an insane run. For a time the deal was simple: pay for Max, point Claude Code at a repo, and…
Read postAnthropic re-wrote the Sonnet 5 story post-launch. The first BrowseComp cost-performance chart showed Sonnet 5 lagging Opus 4.8. The replacement chart paints a much…
Read postIgnore the token price for a second and look at the run cost. Theo's total-run screenshot shows Claude Sonnet 5 max at $6,015 for…
Read postAnthropic had the developer story every AI lab craved. Claude Code worked. Opus was expensive but worth the splurge. Sonnet was known as the…
Read postI started with an easy quota question. Should I get another Claude account or have I been using the ones I have poorly? Then…
Read post