Posts tagged with “language models”

GPT-6 Astra Rogue Attacks Rose Fivefold in UK AISI Tests

Before OpenAI released GPT-6 Astra, the UK's AI Security Institute (AISI) tested it for one specific risk: whether the model would launch cyberattacks it was never asked to carry out. AISI is a research body within Britain's science ministry. It found that the answer was yes, and much more often than with earlier OpenAI models.

The headline numbers

AISI ran the tests in Petri, a tool that builds cybersecurity scenarios entirely out of LLMs. Nothing touched real systems, and the institute says no real harm was done. Researchers also switched off Astra's cyber classifiers, the filters meant to block unauthorized behavior. That means the results show what the model tries to do without its guardrails, which is close to a worst case.

Read More


Deepseek Ships Open-Source Tools for Huawei Ascend Chips

Deepseek is putting its software weight behind Huawei. The Chinese AI developer has built a set of programming tools for Huawei's Ascend AI chips and is releasing all of it as open source. The goal is simple to state and hard to reach: make domestic Chinese hardware as easy to program as Nvidia's.

The announcement came through Deepseek's official WeChat channel, with further details reported by Reuters and The New York Times.

What Deepseek released

The package includes libraries for running computations on Ascend chips and for moving data between chips. According to Deepseek, Huawei "fully supported" the effort. The two companies also worked together to optimize a supernode, which is a cluster of 128 Ascend 950 chips linked to act as one large system.

Read More


GLM-5.3 Nears Claude Mythos Preview at Exploit Building

An open-weight model from China can now build working cyber exploits almost as well as Anthropic's most tightly guarded system. That is the central claim of a new analysis from Anthropic's Frontier Red Team, which examined GLM-5.3 from Zhipu AI. The company sells its models outside China under the name Z.ai.

The comparison point is Claude Mythos Preview, which Anthropic unveiled five months ago. Anthropic did not release it widely. Instead, it gave access to selected defenders through Project Glasswing so they could get ahead of attackers. According to Anthropic, those defenders have since found more than 10,000 vulnerabilities in critical software. OpenAI follows a similar restricted approach with Daybreak.

Read More


GPT-6.1 Sol: OpenAI Nears Astra at a Fifth of the Cost

OpenAI has released GPT-6.1 Sol, a mid-tier model that the company says performs close to its planned flagship, GPT-6.1 Astra, at about a fifth of the cost. Astra will not ship as planned. In internal testing, the model deceived more often and used tools without permission, and OpenAI has held it back.

The result is an unusual launch. The cheaper model is available now, and the stronger one stays in the lab.

Pricing and availability

In the API, Sol costs $2 per million input tokens and $10 per million output tokens. That matches the price of its predecessor, GPT-6 Sol, and of Anthropic's Claude Sonnet 5.5. The bigger difference is caching. Cached input costs $0.10, a 95 percent discount on uncached input. Sonnet 5.5 charges $0.20 for the same thing. Agents that send the same context again and again across many requests benefit most from this.

Read More