HomeTechnologyOpenAI sidesteps Nvidia with unusually fast coding model on plate-sized chips

OpenAI sidesteps Nvidia with unusually fast coding model on plate-sized chips

TechnologyFebruary 13, 2026
1 min read
OpenAI sidesteps Nvidia with unusually fast coding model on plate-sized chips
OpenAI's new GPT‑5.3‑Codex‑Spark is 15 times faster at coding than its predecessor.
Reading Settings

On Thursday, OpenAI released its first production AI model to run on non-Nvidia hardware, deploying the new GPT-5.3-Codex-Spark coding model on chips from Cerebras. The model delivers code at more than 1,000 tokens (chunks of data) per second, which is reported to be roughly 15 times faster than its predecessor. To compare, Anthropic's Claude Opus 4.6 in its new premium-priced fast mode reaches about 2.5 times its standard speed of 68.2 tokens per second, although it is a larger and more capable model than Spark.

"Cerebras has been a great engineering partner, and we're excited about adding fast inference as a new platform capability," Sachin Katti, head of compute at OpenAI, said in a statement.

Codex-Spark is a research preview available to ChatGPT Pro subscribers ($200/month) through the Codex app, command-line interface, and VS Code extension. OpenAI is rolling out API access to select design partners. The model ships with a 128,000-token context window and handles text only at launch.

Read full article

Comments

Source: Ars Technica

Share this article

Related Articles

The Real Reason Data Center Gas Power Plants Are So Dirty
Aug 149 hours ago

The Real Reason Data Center Gas Power Plants Are So Dirty

A massive new gas plant in Texas will be built with much less efficient technology than regular gas plants. It’s far from the only data center power project to rely on dirty turbines.

6a7e221c89f89b7e8931115c5 min read
Read More