Anthropic finds GLM-5.3 can build browser exploits

Anthropic reported on September 29 that Z.ai’s open-weight GLM-5.3 model could build working browser exploits and that researchers could substantially weaken its refusal safeguards. The finding concerns a model that anyone can download, modify and run. Anthropic’s Frontier Red Team tested whether GLM-5.3 could turn known software defects into working attacks and whether it would follow explicitly harmful instructions after common safeguard-bypass techniques. Anthropic researchers Andrew Fasano, Marius Fleischer, Cole McFaul, Robert Xiao and Tripp Gallagher authored the report. The team ran models in isolated environments and combined automated benchmarks with sessions in which security researchers directed the model while examining unfamiliar software targets. ...

September 30, 2026 · Martin Seckar

OpenAI launches always-on Dots agents

OpenAI launched Dots on September 29, introducing persistent agents that run on cloud computers, use connected applications and continue assigned work when the user is absent. The product changes the unit of interaction from a conversation to an ongoing working relationship. A Dot can keep several projects active, retain context across ChatGPT, Slack and Microsoft Teams, and contact its user with progress reports or decisions that require approval. Why it matters: A worker delegating a continuing responsibility gives the system more time, context and opportunity to act than a single chat permits. That makes permissions, audit records and interruption controls part of the product rather than optional deployment work. ...

September 30, 2026 · Martin Seckar

OpenAI releases GPT-6.1 Sol

OpenAI released GPT-6.1 Sol on September 29, pricing the model at $2 per million input tokens and $10 per million output tokens through its API. The company positions the model between its existing GPT-6 Sol and flagship GPT-6 Astra. OpenAI says the new version approaches Astra on coding, computer-use and professional-work evaluations while charging one-fifth of Astra’s standard input and output prices. Why it matters: Developers running agents repeatedly pay for long prompts, tool results and generated output. A lower price at near-flagship capability can change which model they leave active for routine work and which tasks still justify Astra. ...

September 30, 2026 · Martin Seckar

London rail face-scan trial yields no alert-led arrests

A British Transport Police facial-recognition trial scanned more than half a million faces but produced one incorrect watchlist alert and no arrests caused by an alert, the Guardian reported on September 29. The six-month pilot covered 18 deployments at busy London railway stations between February and July. Records obtained through a freedom-of-information request put equipment and staffing costs at £320,786 and police time at almost 100 hours. Why it matters: Rail passengers had their biometric data processed at scale, while the system generated no correct watchlist match during the reported period. That result gives lawmakers and oversight bodies a concrete deployment record for judging whether the intrusion was proportionate. ...

September 30, 2026 · Martin Seckar

OpenAI adds shared Spaces and team tasks

OpenAI added shared Spaces, collaborative Pages and team tasks to ChatGPT on September 29, extending the product from individual conversations into a workspace for people and agents. The release gives teams a shared location for knowledge and ongoing work. ChatGPT can organize a Space using team instructions, while Pages support joint writing, research, charts and images. Business and Enterprise customers can also assign recurring tasks that use connected tools on a schedule or after events such as a new email. ...

September 30, 2026 · Martin Seckar

AI Daily Digest for 30 September 2026

PostHog released a reasoning-based decision model while OpenAI expanded its developer, pricing and enterprise-distribution offers. Community discussions focused on privacy measurements, data-centre economics and possible restrictions on Chinese open weights. » Why it matters: The day’s lower-ranked releases and discussions show where AI products are being packaged for routine work—and where users still lack independent evidence about performance, privacy and cost. In brief PostHog adds reasoning to a decision model. Jeeves is a 9-billion-parameter, Jev-compatible model that answers yes-or-no, multiple-choice and rating questions after an optional reasoning stage. PostHog released weights, code and data and reports stronger results than Jev on selected public tests; those results remain developer-run. It is worth watching because it combines a constrained decision interface with a slower reasoning path instead of sending every classification task to a general chatbot. Source ...

September 30, 2026 · Martin Seckar

AI Community Digest for 29 September 2026

AI Community Digest for 29 September 2026 Developers debated how to verify AI output, while researchers shared a decision-model project and revisited coding-agent evaluation. The items distinguish opinions and research claims from independent findings. » Why it matters: A launch announcement answers what a supplier offers. These discussions ask how a person checks the result, retains understanding and decides when an apparently successful task is incomplete. Hacker News Two coding essays put understanding beside output speed. Alex Ewerlöf’s September 26 essay argues that generating code does not remove responsibility for maintenance and correctness. A separate September 28 discussion of system architecture asks how developers retain enough understanding to review changes. These are related professional arguments, merged here as one item. Their value is a practical question for teams: can the person accepting a patch explain its effect on the surrounding system? Neither thread establishes a measured productivity loss. Ewerlöf discussion; author’s essay; architecture discussion. ...

September 29, 2026 · Martin Seckar

AMD agrees to buy World Labs for $8.2 billion

AMD agrees to buy World Labs for $8.2 billion The proposed share-based acquisition would bring spatial-model research into AMD. Closing and the promised engineering benefits remain future events. AMD, the semiconductor developer, said on September 28 that it agreed to acquire World Labs, a spatial-intelligence research company, in an all-stock transaction valued at approximately $8.2 billion. The chip company is buying a team that builds AI models of three-dimensional environments. It wants that research to inform its computing products. Customers should distinguish that plan from hardware or software improvements already delivered. ...

September 29, 2026 · Martin Seckar

Anthropic releases Claude Sonnet 5.5

Anthropic releases Claude Sonnet 5.5 Anthropic keeps Sonnet’s token prices while claiming lower task costs. The useful comparison is the cost of work that passes review. Anthropic, an AI developer, released Claude Sonnet 5.5 on September 28, positioning the language model for routine coding and office work with claimed efficiency gains over Sonnet 5. The company has updated its everyday work model. Customers pay the same advertised rates for text processing. Whether they save money depends on how much work the model needs to finish a task. ...

September 29, 2026 · Martin Seckar

CASP researchers examine AI research feedback risks

CASP researchers examine AI research feedback risks The report asks whether automated AI development could sharply accelerate progress. It presents a risk scenario with uncertainty, not an established timetable. Researchers published a report on September 28 through the Cambridge Programme on AI Science and Policy, or CASP, examining whether automation of AI research could trigger unusually rapid capability growth. The report considers AI systems helping to build better AI systems. Its authors argue that this feedback could shorten development cycles. They call for better oversight while acknowledging substantial uncertainty about the outcome. ...

September 29, 2026 · Martin Seckar