Tech & AI News
Ars Technica

OpenAI sidesteps Nvidia with unusually fast coding model on plate-sized chips

OpenAI deployed its GPT-5.3-Codex-Spark coding model on Cerebras chips, delivering code at over 1,000 tokens per second. This model is roughly 15 times faster than its predecessor. It features a 128,000-token context window and is available to ChatGPT Pro subscribers.