Aolani and FriendliAI Partner to Meet Surging AI Inference Demand

Aolani will supply GPU cloud infrastructure to FriendliAI, addressing the critical compute shortage for production-scale AI inference and signaling Asia's growing role in the global AI ecosystem.

DC Metrowire Staff
Technology

Aolani, a Singapore-founded neocloud powering AI growth, has announced a partnership to supply GPU cloud infrastructure to FriendliAI, the San Francisco-headquartered inference cloud for frontier AI. The collaboration aims to support the rapidly growing demand for inference services as AI applications become integral to everyday business workflows and organizations transition from experimentation to deployment at scale.

The global market for AI inferencing is expanding quickly, driven by the need to run AI models efficiently in production. FriendliAI, founded by researchers who invented continuous batching—a now-standard technique in AI inference serving—has built its inference stack end to end, from optimized GPU kernels to global distribution. This enables production AI workloads to run fast and reliably at scale. The company consistently ranks as one of the fastest inference providers on OpenRouter, with enterprise clients including LG, Kilo Code, and Liner running their production inference on the platform.

Efficient time-to-value and dependable compute are increasingly important to keep services responsive as usage grows. As access to reliable compute infrastructure becomes a strategic differentiator for companies scaling production workloads, more AI natives are turning to Asia for high-performance compute capacity, attracted by the region's expanding digital infrastructure, strategic connectivity, and growing AI ecosystem.

As one of the leading neoclouds offering purpose-built next-generation AI infrastructure, Aolani helps AI natives scale more efficiently. Aolani's infrastructure capabilities across orchestration, automation, and lifecycle management actively support FriendliAI's services. This partnership equips FriendliAI with the compute to serve rapid customer demand, both across the globe and increasingly in Asia.

Nicholas Chia, Chief Executive Officer at Aolani, said: "We're seeing inference needs grow faster than companies can find compute to support and service their customers. To narrow the supply and demand gap, we actively partner with companies like FriendliAI to deliver compute capacity on time, at scale, and to rigorous standards. We look forward to partnering with the FriendliAI team to grow its services to bring fast and reliable inference to developers worldwide."

Byung-Gon Chun, Founder and CEO of FriendliAI, said: "We are seeing exponential growth in demand for our frontier AI inference services. Businesses need the freedom to choose the AI models that best suit their applications and the ability to run them efficiently in production. Our job is to deliver high-performance, reliable inference so developers can focus on building their AI applications. Aolani stood out as a trusted infrastructure partner that can help us scale at the pace our customers need. We look forward to working with Aolani to support our mission."

The partnership underscores the critical role of specialized cloud infrastructure in the AI era. As inference workloads surge, the ability to provide scalable, high-performance compute becomes a competitive advantage. By combining Aolani's next-generation AI infrastructure with FriendliAI's optimized inference stack, the two companies aim to accelerate the deployment of AI applications globally. For more information, visit Aolani's website and FriendliAI's website.

Blockchain Registration

QR Code for Blockchain Registration