Aolani and FriendliAI Join Forces to Meet Surging AI Inference Demand

Aolani's partnership with FriendliAI provides GPU cloud infrastructure to scale AI inference services, addressing a critical compute shortage and highlighting Asia's growing role in AI infrastructure.

Houston Metrowire Staff
Technology

Aolani, a Singapore-founded neocloud specializing in AI infrastructure, has announced a partnership with FriendliAI, a San Francisco-based inference cloud, to supply GPU cloud infrastructure. The move comes as demand for AI inference—the process of running trained models to generate outputs—explodes, driven by businesses moving from AI experimentation to full-scale deployment. Inference has become the backbone of everyday AI applications, and access to reliable, high-performance compute is now a strategic differentiator.

FriendliAI, founded by researchers who invented continuous batching, a technique now standard in AI inference serving, has built its inference stack end to end, from optimized GPU kernels to global distribution. It consistently ranks as one of the fastest inference providers on OpenRouter and counts enterprises like LG, Kilo Code, and Liner among its clients. As usage grows, FriendliAI needs dependable compute to keep services responsive. By partnering with Aolani, it gains the capacity to serve rapid customer demand globally, with a particular focus on Asia.

Aolani’s infrastructure capabilities span orchestration, automation, and lifecycle management, enabling AI natives to scale efficiently. The company positions itself as a leading neocloud offering purpose-built, next-generation AI infrastructure. Nicholas Chia, CEO of Aolani, noted, “We’re seeing inference needs grow faster than companies can find compute to support and service their customers. To narrow the supply and demand gap, we actively partner with companies like FriendliAI to deliver compute capacity on time, at scale, and to rigorous standards.”

Byung-Gon Chun, Founder and CEO of FriendliAI, added, “We are seeing exponential growth in demand for our frontier AI inference services. Businesses need the freedom to choose the AI models that best suit their applications and the ability to run them efficiently in production. Aolani stood out as a trusted infrastructure partner that can help us scale at the pace our customers need.”

The partnership underscores a broader trend: as AI inference demand surges, companies are turning to Asia for high-performance compute capacity, attracted by expanding digital infrastructure, strategic connectivity, and a growing AI ecosystem. For FriendliAI, the deal ensures it can meet customer demand without compromising on speed or reliability. For Aolani, it reinforces its role as a key enabler of AI growth in the region and beyond.

This collaboration matters because it directly addresses the compute bottleneck that threatens to slow AI adoption. Without sufficient inference capacity, even the most advanced models cannot deliver value at scale. By combining Aolani’s infrastructure with FriendliAI’s optimized inference stack, the partnership aims to keep AI services fast and reliable as usage grows. It also signals Asia’s rising importance in the global AI supply chain, offering an alternative to traditional compute hubs. As more enterprises deploy AI in production, such partnerships will be critical to sustaining innovation and meeting user expectations.

For more information, visit Aolani and FriendliAI.

Blockchain Registration

QR Code for Blockchain Registration