Aolani, a Singapore-founded neocloud powering AI growth, announced a partnership to supply GPU cloud infrastructure to FriendliAI, the San Francisco-headquartered inference cloud for frontier AI, to support the rapidly growing demand for inference services.
The global market for AI inferencing is expanding quickly as AI applications become part of everyday business workflows and organisations move from experimentation to deployment at scale. At the forefront of production-scale AI, FriendliAI serves this exact demand to help developers and enterprises deploy open-weight and custom AI models.
FriendliAI was founded by researchers who invented continuous batching, which is now a standard across AI inference serving. The company has built its inference stack end-to-end, from optimised GPU kernels to global distribution, so that production AI workloads run fast and reliably at scale. FriendliAI consistently ranks as one of the fastest inference providers on OpenRouter, with enterprise clients including LG, Kilo Code, and Liner running their production inference on the platform.
Efficient time-to-value and dependable compute are increasingly important to keep services responsive as usage grows. As access to reliable compute infrastructure becomes a strategic differentiator for companies scaling production workloads, more AI natives are turning to Asia for high-performance compute capacity, attracted by the region's expanding digital infrastructure, strategic connectivity, and growing AI ecosystem.
As one of the leading neoclouds offering purpose-built next-generation AI infrastructure, Aolani helps AI natives scale more efficiently. Aolani's infrastructure capabilities across orchestration, automation and lifecycle management actively support FriendliAI's services. This partnership equips FriendliAI with the compute to serve the rapid customer demand, both across the globe and increasingly in Asia.
Efficient time-to-value and dependable compute are increasingly important to keep services responsive as usage grows

