As reported by toolnavs, OpenAI unveiled its ultra-low-latency inference tier, known as Ultrafast, concurrently with the launch of GPT-5.6 Sol in August of this year. Official statistics reveal that this tier can process data at a remarkable rate of up to 750 tokens per second, marking a speed that is 14 times greater than that of the Standard tier, aligning with the fast-paced demands of modern AI applications.
