Same model, same weights, different silicon: the new API tier parks GPT-5.6 Sol in on-chip SRAM and skips the memory bus entirely.
OpenAI has opened a limited preview of OpenAI Ultrafast, an API service tier that serves GPT-5.6 Sol at up to 750 output tokens per second, roughly 14 times what the Standard tier manages. The model itself has not ...