Aolani and FriendliAI Join Forces to Meet Soaring AI Inference Demand

Aolani will supply GPU cloud infrastructure to FriendliAI, addressing the critical compute shortage for production-scale AI inference and highlighting Asia's growing role in the global AI ecosystem.

LA Metrowire Staff
Technology

Aolani, a Singapore-founded neocloud specializing in AI infrastructure, has announced a strategic partnership to provide GPU cloud infrastructure to FriendliAI, the San Francisco-based inference cloud for frontier AI. The collaboration aims to support the rapidly growing demand for AI inference services as organizations shift from experimentation to full-scale deployment of AI applications.

The global market for AI inferencing is expanding quickly as AI becomes embedded in everyday business workflows. FriendliAI, founded by researchers who invented continuous batching—a now-standard technique in AI inference serving—has built a full inference stack from optimized GPU kernels to global distribution. Its platform consistently ranks among the fastest inference providers on OpenRouter, with enterprise clients such as LG, Kilo Code, and Liner relying on it for production workloads.

Efficient time-to-value and dependable compute are increasingly critical as usage grows. Access to reliable compute infrastructure has become a strategic differentiator, prompting more AI-native companies to turn to Asia for high-performance capacity. The region's expanding digital infrastructure, strategic connectivity, and growing AI ecosystem make it an attractive destination for scaling production workloads.

Aolani, one of the leading neoclouds offering purpose-built next-generation AI infrastructure, helps AI natives scale more efficiently. Its capabilities across orchestration, automation, and lifecycle management directly support FriendliAI's services. This partnership equips FriendliAI with the compute needed to meet rapid customer demand globally and increasingly in Asia.

Nicholas Chia, Chief Executive Officer at Aolani, said: "We're seeing inference needs grow faster than companies can find compute to support and service their customers. To narrow the supply and demand gap, we actively partner with companies like FriendliAI to deliver compute capacity on time, at scale, and to rigorous standards. We look forward to partnering with the FriendliAI team to grow its services to bring fast and reliable inference to developers worldwide."

Byung-Gon Chun, Founder and CEO of FriendliAI, added: "We are seeing exponential growth in demand for our frontier AI inference services. Businesses need the freedom to choose the AI models that best suit their applications and the ability to run them efficiently in production. Our job is to deliver high-performance, reliable inference so developers can focus on building their AI applications. Aolani stood out as a trusted infrastructure partner that can help us scale at the pace our customers need. We look forward to working with Aolani to support our mission."

This partnership underscores the critical role of specialized cloud providers in bridging the compute gap for AI inference. As demand for AI services continues to surge, collaborations between infrastructure providers like Aolani and inference platforms like FriendliAI will be essential to ensure developers and enterprises can deploy AI at scale reliably and efficiently. The move also signals Asia's rising importance as a hub for AI compute capacity, potentially shifting the balance of global AI infrastructure. For more information about Aolani, visit https://www.aolanicloud.com/. To learn about FriendliAI, visit https://friendli.ai/.

Blockchain Registration

QR Code for Blockchain Registration