Recently, Sugon unveiled its newly enhanced AI inference-native storage solution, FN Neo. According to real-world testing, this upgraded product slashes the first Token response time for models by 73% and boosts the overall Token throughput capacity by 150%. This results in a significant reduction in user interaction wait times and an enhancement in the system’s ability to handle concurrent services. Designed to meet the round-the-clock operational needs of production-level AI services, FN Neo boasts a high reliability rate of 99.999%. It incorporates multiple fault protection mechanisms to ensure uninterrupted and continuous inference services. Moreover, the product is compatible with mainstream inference frameworks and various networking environments, enabling enterprises to deploy it without the need for extensive modifications to their existing systems. By harnessing the existing inference ecosystem, FN Neo elevates both quality and efficiency, providing robust support for the large-scale production and deployment of large models and intelligent agents.
