2026年10月06日

On September 30, we announced $668 million in new fundraising: $223 million in equity for our Series B and a $445 million credit facility led by CTBC. The Series B was led by ARCHIV, a new investment firm based in San Francisco that specializes in AI and robotics, with participation from NVIDIA. Investors across the Asia-Pacific region also took part, including DSC Investment, Trend Micro, KB Investment, Kyobo Life, KT Corporation and others.
Reliable compute should be for everyone
We believe reliable compute should be for everyone. Compute should never be the biggest obstacle standing between vision and breakthrough, or between builders and the products they want to ship. The new GDP is GPUs, data, and power, and together they unlock infinite intelligence. Our mission is to make that intelligence borderless and inclusive, not confined to one geography or reserved for a handful of players.
"Our customers are scaling faster than ever, and they need infrastructure that keeps pace," said Alex Yeh, our Founder and CEO. "AI is driving a new renaissance, and reliable compute is its foundation. Our goal is to build that foundation across continents, with an ecosystem of products on top of it."
Global Capacity in U.S. and APAC
Demand for compute is no longer regional. U.S. AI companies and hyperscalers need capacity in both the United States and Asia, and enterprises across Asia-Pacific want production AI built close to home, under local data and compliance requirements. Most AI clouds are built for one side of that equation. We're built for both, operating as a single platform so customers can place workloads where their users, data, and regulators are.
-3.mp4 Comp 1_00000.png)
We deliver on schedule
In AI infrastructure, on-time delivery is everything. That starts with a secure supply chain. Our deep ties to Taiwan, home to most of the world's AI server manufacturing, give us close relationships with manufacturers and a more predictable path from order to deployment. We bring that precision and operational discipline to customers worldwide.
"In AI infrastructure, a delivery date is a promise. Customers plan launches, hiring, and revenue around it," said Alex. "Our place in Taiwan's supply chain is how we keep that promise, cluster after cluster globally."
But on time is only the start. Our customers have their own missions, and our job is to make sure compute is never the reason they fall short. That means infrastructure that is secure and stable, runs the most advanced models, and is built by a team that knows their business.
"Capacity that arrives late is capacity we can't use," said Chenyu Zhao, co-founder of Fireworks. "GMI Cloud has been one of our strongest and most reliable providers across NVIDIA GB200 and GB300 NVL72 systems."
Beyond 9x ARR
Our contracted ARR has reached more than 9x its year-end 2025 level, and live ARR in production has grown more than 4.5x over the same period. Our inference platform now processes about 4 trillion tokens per week, and in September we passed 1 trillion tokens served through OpenRouter. Customers include Fireworks, Higgsfield, Nous Research, OpenRouter, Reflection, Cartesia, Trend Micro, and Utopai Studios.
This year, we became an NVIDIA Exemplar Cloud on NVIDIA GB300 NVL72 systems for training. We launched GMI Agentbox, our agent infrastructure product, and GMI Router, which optimizes performance and token consumption for clients and customers. We launched SCALE for AI startups in March, now in its second cohort, and the Clouder program for developers and creators in May, now 50+ members across four continents. We also built communities in Singapore and Korea.
.jpg)
Building the future
This funding supports our capacity expansion in the United States, Taiwan and the rest of APAC, building on our Taiwan AI Factory, announced in 2025, and our Japan sovereign AI initiative, announced earlier this year. It will also support the continued growth of our inference services and strategic hiring as we scale.
Thank you to our investors for backing this vision, to our customers for trusting us with their most important work, and to our team for making every delivery date count. We're just getting started. If you want to train on GPU clusters, run inference at scale, or deploy agents in production, all on one reliable cloud, we'd love to work with you.
One Cloud for Compute, Inference, and Agents.
Learn more at gmicloud.ai.
GMI Cloud helps you architect, deploy, optimize, and scale your AI strategies
