Aolani, a Singapore-founded neocloud powering AI growth, announced a partnership to supply GPU cloud infrastructure to FriendliAI, the San Francisco-headquartered inference cloud for frontier AI. The deal, disclosed on 10 September 2026, is designed to support rapidly growing demand for inference services as AI applications become embedded in everyday business workflows and organizations move from experimentation to deployment at scale.
The global market for AI inferencing is expanding quickly, driven by the need to run open-weight and custom AI models efficiently in production. FriendliAI, founded by researchers who invented continuous batching—a technique now standard across AI inference serving—has built its inference stack end to end, from optimized GPU kernels to global distribution. The company consistently ranks as one of the fastest inference providers on OpenRouter, with enterprise clients including LG, Kilo Code, and Liner running production inference on its platform.
As access to reliable compute infrastructure becomes a strategic differentiator, more AI natives are turning to Asia for high-performance capacity, attracted by expanding digital infrastructure, strategic connectivity, and a growing AI ecosystem. Aolani, one of the leading neoclouds offering purpose-built next-generation AI infrastructure, helps AI natives scale more efficiently. Its capabilities across orchestration, automation, and lifecycle management actively support FriendliAI’s services, equipping the company with compute to serve customer demand globally and increasingly in Asia.
Nicholas Chia, Chief Executive Officer at Aolani, said, “We’re seeing inference needs grow faster than companies can find compute to support and service their customers. To narrow the supply and demand gap, we actively partner with companies like FriendliAI to deliver compute capacity on time, at scale, and to rigorous standards. We look forward to partnering with the FriendliAI team to grow its services to bring fast and reliable inference to developers worldwide.”
Byung-Gon Chun, Founder and CEO of FriendliAI, said, “We are seeing exponential growth in demand for our frontier AI inference services. Businesses need the freedom to choose the AI models that best suit their applications and the ability to run them efficiently in production. Our job is to deliver high-performance, reliable inference so developers can focus on building their AI applications. Aolani stood out as a trusted infrastructure partner that can help us scale at the pace our customers need. We look forward to working with Aolani to support our mission.”
The partnership matters because it addresses a critical bottleneck in AI deployment: compute supply. As enterprises increasingly rely on AI for daily operations, the ability to run models quickly and reliably becomes a competitive necessity. By combining Aolani’s GPU cloud infrastructure with FriendliAI’s inference expertise, the collaboration aims to accelerate time-to-value for developers and enterprises, particularly in Asia’s fast-growing AI market. For more information, visit https://www.aolanicloud.com/ and https://friendli.ai.

