OpenAI's Planned Expansion of Its Ultrafast, Low-Latency Inference Tier
1 day ago / Read about 0 minute
Author:小编   

As reported by toolnavs, OpenAI unveiled its ultra-low-latency inference tier, known as Ultrafast, concurrently with the launch of GPT-5.6 Sol in August of this year. Official statistics reveal that this tier can process data at a remarkable rate of up to 750 tokens per second, marking a speed that is 14 times greater than that of the Standard tier, aligning with the fast-paced demands of modern AI applications.