29 Jul 2026

MangoBoost Redefines the AI Inference Race Beyond NVIDIA

Stand: Booth 816
MangoBoost Redefines the AI Inference Race Beyond NVIDIA
Validates Performance with AMD and Launches Mango Inference to Accelerate Global AI Platform Expansion


- Up to 1.25× higher token throughput than NVIDIA Blackwell B300
- Up to 5.8× higher throughput and over 2× lower latency than NVIDIA H100 clusters

 


AMD CEO, Lisa Su, and South Korea’s Minister of Science and ICT, Bae Kyung-hoon, visit MangoBoost's booth at AMD Advancing AI 2026, highlighting the company's role in national AI infrastructure.

MangoBoost announced the commercial launch of Mango Inference, its serverless AI inference platform, while unveiling breakthrough inference performance on AMD's latest Instinct™ MI355X accelerators at AMD Advancing AI 2026 in San Francisco.


At the event, the company demonstrated inference performance surpassing industry-leading GPU platforms, highlighting a shift in AI competition from model size to efficient inference delivery and reinforcing its position as a global AI inference platform provider.

Figure 1: MangoBoost software improving performance on the AMD Instinct™ MI355X GPUs vs NVIDIA B300.

Using AMD Instinct™ MI355X GPUs combined with MangoBoost's inference optimization software LLMBoost™, the system achieved up to 1.25× higher token throughput than industry-leading GPU architectures under the same latency constraints in GLM-5/5.1 (FP8) benchmarks. End-to-end latency was reduced by up to 5.1×, cutting AI response time from 142 seconds to 28 seconds.

Figure 2: MangoBoost software improving performance on the AMD Instinct™ MI300X GPUs vs NVIDIA H100. The system also demonstrated outstanding performance on DeepSeek-R1, a reasoning-focused AI model. In collaboration with the AMD team, MangoBoost was able to improve performance on AMD Instinct™ MI300X GPUs, and achieved 5.8x higher throughput, 2.05x lower latency, leading to an overall 7.14x lower cost versus the best NVIDIA H100 configurations. These results were further validated through MLPerf, the industry's leading AI performance benchmark. With these performance results, MangoBoost's product capabilities have led to a strategic collaboration with AMD.

 

As AI moves from training to inference, real-world competitiveness depends on how effectively compute, networking, storage, memory, and system software are integrated and optimized. MangoBoost's achievements demonstrate that AI leadership is defined by system-level optimization rather than hardware specifications alone. Through its collaboration with AMD, MangoBoost validated and improved AI inference performance via MLPerf submissions and unofficial InferenceX projects on AMD Instinct™ GPUs.

The image shows a group of people dressed in formal attire, gathered in a room with a large window behind them, likely attending a professional or technical event.

AI-generated content may be incorrect.

Eriko Nurvitadhi, MangoBoost's Chief AI Officer, briefs AMD CEO Lisa Su and Minister Bae Kyung-hoon on MangoBoost’s next-generation Agentic AI Infrastructure and AMD collaboration during the visit.

Additionally, the event highlighted Korea's growing leadership in AI infrastructure. Deputy Prime Minister and Ministry of Science and ICT Bae Kyung-hoon and AMD Chair and CEO Lisa Su visited the MangoBoost booth to review its AI inference technologies and future vision. Their visit underscored MangoBoost's growing global recognition as a key contributor to South Korea's flagship AI initiatives.

Following successful Seed and Series A funding, MangoBoost is preparing for its Series B financing, supported by validated technology, commercial launch of Mango Inference, and growing global customer traction. The investment will accelerate expansion across North America, the Middle East, Asia, and Europe while advancing its AI inference platform and enterprise offerings.


MangoBoost aims to become a global full-stack AI infrastructure leader, optimizing computing, networking, memory, and storage across next-generation AI data centers. Its long-term vision is to establish the global standard for AI infrastructure optimization in the Agentic AI era.

 

Jangwoo Kim, CEO of MangoBoost, said, "Our industry is no longer competing over which GPU is used. The real competition is about how efficiently hardware and software are integrated into a complete AI infrastructure. Together with AMD, we will continue setting new benchmarks for AI inference and expand Mango Inference as a global AI platform."

 

Overall, this marks a turning point demonstrating that the future of AI inference will be defined by system optimization. With world-class performance and a commercially available AI inference platform, MangoBoost is reshaping the competitive landscape of AI infrastructure worldwide.

About MangoBoost

MangoBoost builds full-stack AI infrastructure software and hardware that raises the efficiency of AI workloads. Founded out of Seoul National University and headquartered in Bellevue, WA, the company develops LLMBoost™ inference optimization software, DPU acceleration and Agent OS, and operates Mango Inference, a managed serverless inference platform. More at https://www.mangoboost.io/

Loading