OpenAI unveils GPT-5.6 amid US AI regulatory drama

75d ago · US · primary source: theverge.com

OpenAI released GPT-5.6, a three-model AI suite, on Friday, less than 24 hours after reports that the Trump administration would oversee its staggered rollout. The preview includes Sol, Terra, and Luna, with Sol priced at $5 per million input tokens and $30 per million output tokens. The flagship Sol model is positioned for complex tasks, while Terra targets high-volume work and Luna is designed as a fast, affordable option [1]. Sol's pricing is nearly half that of Anthropic's Claude Fable 5, which costs $10 per million input tokens and $50 per million output tokens [1]. Terra is half the cost of Sol, and Luna is less than half the cost of Terra [1]. Anthropic, founded in 2021 by former OpenAI members, had an estimated valuation of $965 billion in May 2026 [10]. OpenAI stated that GPT-5.6 is trained to refuse prohibited cyber assistance, including when users attempt to disguise their intent or jailbreak the model [1]. The company said Sol is better at helping people find and fix vulnerabilities than reliably carrying out end-to-end attacks [1]. OpenAI dedicated approximately 700,000 A100e GPU hours to automated red-teaming and worked with third-party testers who will continue evaluation for the next two weeks [1]. The release comes amid heightened security scrutiny in Washington. The Trump administration will approve customers on a case-by-case basis during the preview period [1]. OpenAI acknowledged that safeguards may occasionally intervene on legitimate work, particularly in dual-use areas where defensive and offensive activity can initially look similar [1]. The company wrote, "We don't believe this kind of government access process should become the long-term default" [1]. OpenAI described the arrangement as a short-term step toward broader availability while working with the administration to develop a cyber Executive Order framework and a repeatable process for future model releases [1]. The model suite is expected to become generally available in the coming weeks [1]. OpenAI also debuted two additional modes for Sol: a "max" mode for deeper reasoning and an "ultra" mode for leveraging sub-agents [1]. The company said Sol has its most robust safety stack to date, with strengthened protections for higher-risk activity, sensitive cyber requests, and repeated misuse [1].

model-releaseregulationapplication

Background sources we checked (10)
  • en.wikipedia.org ↗ Elon Reeve Musk ( EE-lon; born June 28, 1971) is a businessman and former public official who is the CEO and largest shareholder of Tesla and SpaceX. Musk has been the wealthiest person in the world since 2025, and became the first and only trillionaire in terms of US dollars in…
  • en.wikipedia.org ↗ The following is a list of events of the year 2026 in the United States, as well as predicted and scheduled events that have not yet occurred. July 4, 2026, will be the 250th anniversary of the signing of the Declaration of Independence of the United States from the United Kingdo…
  • en.wikipedia.org ↗ The 21st century is the current century in the Anno Domini or Common Era, in accordance with the Gregorian calendar. It began on 1 January 2001, and will end on 31 December 2100. It is the first century of the 3rd millennium. The rise of a global economy and Third World consumeri…
  • arxiv.org ↗ We present DarkAgents: a multi-agent system that leverages the reasoning and code-generation capabilities of large language models (LLMs), together with deterministic tested human-written code, to build orchestrated pipelines for theoretical astroparticle physics research. While …
  • arxiv.org ↗ Indirect prompt injection in tool-use agents is a concrete production threat: LLM agents read from integrations (third-party services such as Gmail, Salesforce, or Jira accessed through tool calls) whose response content the user neither writes nor controls. Existing benchmarks u…
  • arxiv.org ↗ Selecting the right electricity market region for a hyperscale AI datacenter requires reasoning across live electricity prices, grid carbon intensity, technology cost trajectories, and causal grid dynamics -- a multi-step, multi-source analytical task that static knowledge benchm…
  • arxiv.org ↗ Coding agents often pass per-prompt safety review yet ship exploitable code when their tasks are decomposed into routine engineering tickets. The challenge is structural: existing safety alignment evaluates overt requests in isolation, leaving models blind to malicious end-states…
  • arxiv.org ↗ Existing benchmarks of language-model refusal on malicious-coding tasks routinely conflate requests for executable malicious software with requests for harmful security knowledge. This conflation matters because the two request types plausibly trigger distinct refusal pathways in…
  • en.wikipedia.org ↗ Anthropic PBC is an American artificial intelligence (AI) company headquartered in San Francisco, California. It has developed a series of large language models (LLMs) named Claude and has a focus on AI safety. Anthropic was founded in 2021 by former members of OpenAI, including …
  • en.wikipedia.org ↗ Christopher Olah (born 1992 or 1993) is a Canadian machine learning researcher and a co-founder of Anthropic. He is known for his work on neural network interpretability, particularly mechanistic interpretability, and for research and tools that visualise internal representations…

Sources

Spot something wrong? Report an issue