STOCK TITAN

Cerebras Powers Ultrafast Mode for OpenAI’s GPT-5.6 Sol

(Moderate)
(Neutral)
Tags
crypto AI

Cerebras (NASDAQ: CBRS) announced that its hardware powers Ultrafast mode, a new OpenAI API tier for GPT-5.6 Sol, delivering up to 750 output tokens per second and up to 14× faster processing than the Standard tier, with the same model intelligence.

According to Cerebras, Ultrafast is initially in limited preview and targets workloads needing very low latency. Reported comparisons indicate Ultrafast runs about 5× faster than Claude Opus 4.8 Fast and 11× faster than Claude Fable 5, and achieves a 5.6× speedup on GDP-Val benchmarks with no measured quality loss.

Loading...
Loading translation...

Positive

  • Powers OpenAI GPT-5.6 Sol Ultrafast tier at up to 14× speed
  • Supports up to 750 output tokens per second for GPT-5.6 Sol
  • Ultrafast about 5× faster than Claude Opus 4.8 Fast mode
  • Ultrafast about 11× faster than Claude Fable 5 on output speed
  • 5.6× end-to-end speedup on GDP-Val with no quality loss reported
  • Wafer-Scale Engine keeps 44 GB of model weights on-chip

Negative

  • Ultrafast service initially limited to a small group of customers

News Explained

The limited-preview service uses Cerebras’s Wafer-Scale Engine to keep model weights in 44 GB of on-chip SRAM per wafer-sized chip instead of moving them between on-chip memory and off-chip storage, which the release identifies as the source of its faster inference.

Market Context

Historical event 1085268 recorded a 0.59% reaction after a CrowdStrike partnership. This platform re...
Analysis

Historical event 1085268 recorded a 0.59% reaction after a CrowdStrike partnership. This platform record places the OpenAI service launch beside mixed partnership outcomes, while high short positioning and recent insider net selling remain risks to monitor.

Key Figures

Ultrafast speed: 750 output tokens per second Speed advantage: 14× faster Claude Opus speed advantage: 5x faster +5 more
8 metrics
Ultrafast speed 750 output tokens per second GPT-5.6 Sol Ultrafast limited preview
Speed advantage 14× faster vs. Standard processing
Claude Opus speed advantage 5x faster vs. Claude Opus 4.8 in Fast mode
Claude Fable speed advantage 11x faster vs. Claude Fable 5
Benchmark size 2,500 questions Humanity’s Last Exam
Exam completion time just over 11 hours GPT-5.6 Sol Ultrafast on Humanity’s Last Exam
Accuracy speed advantage nearly 7x faster vs. Claude Fable 5 at comparable accuracy
End-to-end speedup 5.6x GDP-Val benchmark with no loss in quality

Historical Context

5 past events · Latest: Aug 05 (Positive)
Pattern 5 events
Date Event Sentiment 24h Move Catalyst
Aug 05 AI partnership Positive -5.7% Lovable partnership announced to deploy Cerebras inference for latency-sensitive software workflows.
Jul 29 Colocation agreement Positive -12.1% Cerebras signed a long-term data-center colocation agreement with CleanCore Solutions.
Jul 22 Earnings scheduling Neutral +4.9% Cerebras scheduled release of second-quarter 2026 financial results and conference call.
Jul 22 AI partnership Positive +0.6% CrowdStrike deployed Falcon detection models on Cerebras inference infrastructure.
Jul 09 Manufacturing partnership Positive +9.3% Flex expanded manufacturing capacity for Cerebras CS-3 AI supercomputers.

24h Move is the share-price change in the day after each event; other market factors may also have contributed.

Pattern Detected

Positive partnership and commercial-expansion announcements produced mixed reactions, with two positive outcomes and three divergences in the selected history.

Key Terms

inference, sram, memory-bandwidth bottleneck, on-chip memory
4 terms
inference technical
"Cerebras’ inference technology"
Inference is the process of drawing a conclusion from available evidence or data, like a detective piecing together clues to form a likely story. For investors it matters because these judgments turn raw reports, test results, or market signals into expectations about future performance, risk, or regulatory outcomes—so how someone infers from the same facts can change investment decisions and valuation.
sram technical
"44 GB of SRAM on each wafer-sized chip"
Static random-access memory (SRAM) is a fast type of computer memory that stores data instantly and reliably while power is on, similar to a small, very quick notepad a processor keeps next to it. Investors care because SRAM influences product speed, power use, cost and production capacity for electronics and semiconductor companies, affecting profit margins, competitive positioning and supply-chain risk.
memory-bandwidth bottleneck technical
"This eliminates the memory-bandwidth bottleneck"
A memory-bandwidth bottleneck occurs when a computer processor cannot get data from memory fast enough to run at full speed, so the CPU or accelerator sits idle waiting for information. Think of it like a highway where too few lanes slow traffic to a trickle even though the cars (processors) are ready to go; this limits overall performance and can reduce the value of hardware or software that depends on fast data flow. Investors care because such bottlenecks affect product competitiveness, server utilization, and the real-world performance gains behind revenue and cost forecasts.
on-chip memory technical
"between on-chip memory and off-chip storage"
Memory that is built directly into an integrated circuit or processor die, used to store data and instructions close to where they are processed. Like a notepad kept on a worker’s desk instead of across the room, on-chip memory is much faster and more energy-efficient than external memory but usually smaller; its size and speed affect a chip’s performance, power consumption, and cost, which in turn influence product capabilities and competitiveness in the market.

AI-generated analysis. How Rhea-AI works. Not financial advice.

See more from StockTitan in Google Search and AI answers. Adds StockTitan as a preferred source · opens Google
Add on Google

Cerebras powers OpenAI's most capable model at up to 14x the speed, giving frontier intelligence with no compromise on latency

SUNNYVALE, Calif., Aug. 13, 2026 (GLOBE NEWSWIRE) -- Cerebras (NASDAQ: CBRS) today announced that it is powering Ultrafast mode, a new service tier in the OpenAI API for GPT-5.6 Sol. Available initially in limited preview to OpenAI customers, Ultrafast runs GPT-5.6 Sol at up to 750 output tokens per second and up to 14× faster than Standard processing.

GPT-5.6 Sol Ultrafast powered by Cerebras runs with the same intelligence as GPT-5.6 Sol Standard, enabling frontier intelligence at Cerebras' blistering fast speed.

Every major computing shift has been unlocked by a leap in speed, not just capability: the PC era needed the jump from kilohertz to gigahertz, and the internet needed the move from dial-up to broadband before it could reach everyone. AI is no different. Now that frontier models are demonstrably capable, the next constraint on adoption is how fast they run.

“GPT-5.6 Sol on Ultrafast is proof that speed and intelligence are no longer mutually exclusive,” said Andrew Feldman, CEO and co-founder, Cerebras. “Together with OpenAI, we're putting frontier intelligence in the hands of users at unprecedented speed and changing what's possible with AI.”

“By combining GPT-5.6 Sol with Cerebras’ inference technology, we’re exploring what becomes possible when customers can get the intelligence of our most capable models with significantly lower latency. We’re starting with a small group of customers to learn where that speed creates meaningful value, and we’ll use those learnings to inform how we expand the service over time,” said Sachin Katti, VP Compute Strategy & GPT-Infra at OpenAI.

A New Speed-Intelligence Frontier

Until now, organizations needed to choose between the capabilities of larger models and the faster response times of smaller models. Ultrafast powered by Cerebras offers access to the full intelligence of GPT-5.6 Sol at speeds suited to work that cannot wait. Based on output speeds for Anthropic models reported by Artificial Analysis, Ultrafast is 5x faster than Claude Opus 4.8 in Fast mode, and 11x faster than Claude Fable 5. In doing so, it opens up a new region of speed-intelligence: frontier intelligence combined with unprecedented speed.

Benchmarks focused on economically valuable work, from programming to drafting legal documents, show how faster token generation translates directly into higher productivity for users.

On Humanity’s Last Exam, a 2,500-question benchmark spanning graduate-level chemistry, economics and literature, GPT-5.6 Sol Ultrafast answered the full question set in just over 11 hours. This compares to more than three days of continuous compute for Claude Fable 5, with GPT-5.6 Sol Ultrafast reaching comparable accuracy nearly 7x faster. On GDP-Val, a benchmark of economically valuable knowledge-work tasks such as legal briefs, financial models, and engineering reports, Ultrafast delivered a 5.6x end-to-end speedup with no loss in quality.

Ultrafast’s speed comes from Cerebras' Wafer-Scale Engine architecture, which keeps model weights on-chip — 44 GB of SRAM on each wafer-sized chip — rather than shuttling them between on-chip memory and off-chip storage as GPU-based inference must. This eliminates the memory-bandwidth bottleneck that constrains frontier-model inference speed on conventional hardware.

To get notified when capacity expands to more customers, please visit the Cerebras website.

About Cerebras Systems
Cerebras Systems (NASDAQ: CBRS) builds the world’s fastest AI infrastructure. The Cerebras team of pioneering computer architects, computer scientists, AI researchers, and engineers of all types came together to make AI blisteringly fast through innovation and invention. We believe that when AI is fast, it will change the world. Leading global corporations, research institutes, and governments choose Cerebras to run their AI workloads. Cerebras solutions are available on premises and in the cloud. Visit cerebras.ai for more.

Cerebras Disclosure Information
Cerebras uses its investor relations page (investors.cerebras.ai), its X account (@cerebras), and its LinkedIn page (linkedin.com/company/cerebras-systems/) to disclose material nonpublic information and for complying with its disclosure obligations under Regulation FD. Accordingly, investors should monitor these channels, in addition to following Cerebras' press releases, Securities and Exchange Commission (SEC) filings, public conference calls and public webcasts.

Forward-Looking Statements

This press release contains "forward-looking statements" within the meaning of applicable securities laws. All statements other than statements of historical fact could be deemed to be forward-looking. The words "may," "will," "shall," "should," "expects," "plans," "anticipates," "could," "intends," "target," "projects," "contemplates," "believes," "estimates," "predicts," "potential," "objective," or "continue," or the negative of these words or other similar terms or expressions that concern our expectations, strategy, plans, or intentions are intended to identify forward-looking statements, although not all forward-looking statements contain these identifying words. These forward-looking statements are subject to a number of risks and uncertainties, many of which involve factors or circumstances that are beyond Cerebras' control. These risks and uncertainties include, but are not limited to: Cerebras' ability to sustain and manage its growth, access borrowings and other sources of capital on acceptable terms, and deploy available capital to support growth; its history of net losses and ability to achieve and maintain profitability; its limited operating history at its current scale and ability to accurately forecast revenue and appropriately budget and manage expenses; its dependence on a limited number of significant customers, including OpenAI, Group 42 Holding Ltd, Mohamed bin Zayed University of Artificial Intelligence, and AWS, and the potential impact of any reduction in demand from, material adverse development in its relationships with, or failure to meet its obligations to, such customers; the timing, execution and expected benefits of its strategic customer, partner and financing arrangements; its historical reliance on sales of hardware systems and the early-stage, rapidly evolving market for its cloud-based offerings and AI infrastructure; its ability to secure sufficient data center capacity and capital to support its cloud-based offerings; its ability to launch new offerings and add new product capabilities; and its ability to compete effectively in the rapidly evolving and competitive market for AI computing solutions.

Cerebras' actual results could differ materially from those stated or implied in forward-looking statements due to a number of factors. Accordingly, undue reliance should not be placed on such statements. These forward-looking statements are made as of the date they were first issued and are based on information available to Cerebras together with Cerebras' expectations, estimates, forecasts, projections, beliefs, and assumptions as of such date. These forward-looking statements should not be relied upon as representing Cerebras' views as of any date subsequent to the date of this press release. Past performance is not necessarily indicative of future results. Cerebras undertakes no intention or obligation to update or revise any forward-looking statements, whether as a result of new information, future events, or otherwise, except as required by law.

Further information on potential risks that could affect actual results is included in Cerebras' most recent filings with the SEC, including in Cerebras' most recent Quarterly Report on Form 10-Q, copies of which may be obtained by visiting Cerebras' Investor Relations website at investors.cerebras.ai or the SEC's website at www.sec.gov.

Contacts

Kriselle Laran
Media Relations
pr@cerebras.ai

Sean Dorsey
Investor Relations
investors@cerebras.ai


FAQ

What did Cerebras (NASDAQ: CBRS) announce about powering OpenAI GPT-5.6 Sol Ultrafast?

Cerebras announced it powers OpenAI’s new GPT-5.6 Sol Ultrafast API tier, delivering up to 750 tokens per second and up to 14× faster processing than Standard. According to Cerebras, Ultrafast maintains the same model intelligence while significantly reducing latency for demanding workloads.

How fast is GPT-5.6 Sol Ultrafast powered by Cerebras compared with the Standard tier?

According to Cerebras, GPT-5.6 Sol in Ultrafast mode runs up to 14× faster than Standard processing and can generate up to 750 output tokens per second. This speed targets applications where response time is critical but users still require the model’s full intelligence.

How does Cerebras-powered Ultrafast compare to Anthropic models for speed?

Cerebras reports that GPT-5.6 Sol Ultrafast is about 5× faster than Claude Opus 4.8 in Fast mode and about 11× faster than Claude Fable 5. These comparisons are based on output speed data reported by Artificial Analysis, according to Cerebras.

What benchmark results did Cerebras share for GPT-5.6 Sol Ultrafast on Humanity’s Last Exam?

According to Cerebras, GPT-5.6 Sol Ultrafast completed the 2,500-question Humanity’s Last Exam benchmark in just over 11 hours, reaching comparable accuracy to Claude Fable 5 nearly 7× faster. The benchmark spans graduate-level chemistry, economics, literature and other knowledge-work domains.

What is the GDP-Val benchmark result for Cerebras-powered GPT-5.6 Sol Ultrafast?

On the GDP-Val benchmark of economically valuable tasks, GPT-5.6 Sol Ultrafast delivered a 5.6× end-to-end speedup with no loss in quality, according to Cerebras. GDP-Val includes knowledge-work tasks such as legal briefs, financial models, engineering reports and similar productivity-focused workloads.

How does Cerebras hardware achieve higher GPT-5.6 Sol Ultrafast inference speed?

According to Cerebras, its Wafer-Scale Engine architecture keeps 44 GB of model weights in on-chip SRAM, avoiding transfers to off-chip memory. This design removes the memory-bandwidth bottleneck that often constrains frontier-model inference speed on conventional GPU-based hardware platforms.

Is OpenAI’s GPT-5.6 Sol Ultrafast powered by Cerebras widely available to CBRS investors’ customers?

Cerebras states that GPT-5.6 Sol Ultrafast is initially available in limited preview to a small group of OpenAI customers. Interested users can register on the Cerebras website to be notified when capacity increases and the service expands to more customers.