- Round was led by ARCHIV, an AI and Robotics investment firm, with participation from NVIDIA
- Funding will expand GPU capacity across the U.S., Taiwan, and Southeast Asia as contracted ARR tops $600 million
MOUNTAIN VIEW, Calif.--(BUSINESS WIRE)--GMI Cloud, an AI-native cloud delivering high-performance GPU infrastructure and inference services, today announced $668 million in new financing, comprising $223 million in equity for Series B and a $445 million credit facility led by CTBC. The Series B was led by ARCHIV, a new investment firm based in San Francisco that specializes in AI and robotics, with participation from NVIDIA. This round saw strong participation from investors across the Asia-Pacific region, including DSC Investment, Trend Micro, KB Investment, Kyobo Life, KT Corporation and others.




The funding provides growth capital to support GMI Cloud’s capacity expansion in the United States, Taiwan and the rest of APAC. This new capacity builds on the company’s Taiwan AI Factory, announced in 2025 as its inaugural facility in Asia, as well as its Japan sovereign AI initiative announced earlier this year.
Beyond expanding capacity, the Series B will support the continued growth of GMI Cloud’s inference services and strategic hiring as the company scales.
The round follows a period of rapid commercial growth. GMI Cloud's contracted ARR has reached more than 9x its year-end 2025 level. Live ARR in production today has also grown more than 4.5x over the same period. Its inference platform now processes approximately 4 trillion tokens per week. Notable customers include Fireworks, Higgsfield, Nous Research, OpenRouter, Reflection, Cartesia, Trend Micro, and Utopai Studios.
"Our customers are scaling faster than ever, and they need infrastructure that keeps pace," said Alex Yeh, Founder and CEO of GMI Cloud. "AI is driving a new renaissance, and reliable compute is its foundation. Our goal is to build that foundation across continents, with an ecosystem of products on top of it."
Demand for compute is no longer regional. U.S. AI companies and hyperscalers need capacity in both the United States and Asia, and enterprises across Asia-Pacific want production AI built close to home, under local data and compliance requirements. Most AI clouds are built for one side of that equation. GMI Cloud is built for both, operating as a single platform so customers can place workloads where their users, data, and regulators are.
Serving both markets starts with a secure supply chain. GMI Cloud's deep ties to Taiwan, home to most of the world's AI server manufacturing, give it close relationships with manufacturers and a more predictable path from order to deployment. The company brings that precision and operational discipline to customers worldwide.
"In AI infrastructure, a delivery date is a promise. Customers plan launches, hiring, and revenue around it," said Yeh. "Our place in Taiwan's supply chain is how we keep that promise, cluster after cluster globally."
"Capacity that arrives late is capacity we can't use," said Chenyu Zhao, co-founder of Fireworks. "GMI Cloud has been one of our strongest and most reliable providers across NVIDIA GB200 and GB300 NVL72 systems."
For more information, visit gmicloud.ai.
About GMI Cloud
GMI Cloud is an AI-native cloud infrastructure company powering the next generation of AI applications. Built as One Cloud for Compute, Inference, and Agents, GMI Cloud provides GPU clusters, optimized inference, and agent infrastructure on one unified cloud, giving enterprises and developers the capabilities to build, deploy, and scale production AI. Founded in 2021 and headquartered in Mountain View, California, the company operates GPU infrastructure across the United States and Asia-Pacific. For more information, visit gmicloud.ai.
Contacts
Media Contact:
GMI Cloud
louisa.g@gmicloud.ai





