Telecom & Connectivity | 3 min read

SoftBank Launches AI GPU Cloud in Japan With NVIDIA GB200, Targeting Sovereign AI Market

SoftBank officially launched its AI Data Center GPU Cloud this month, deploying NVIDIA GB200 NVL72 systems at Japan-based data centers. The neocloud offering positions SoftBank as Japan's primary sovereign AI compute alternative to U.S. hyperscalers.

Hector Herrera
Hector Herrera
A Data Center featuring Data Center, data centers, related to AI GPU Cloud in Japan With a chip manufacturer GB200, Target from an unusual angle or perspective
Why this matters SoftBank officially launched its AI Data Center GPU Cloud this month, deploying NVIDIA GB200 NVL72 systems at Japan-based data centers. The neocloud offering positions SoftBank as Japan's primary sovereign AI compute alternative to U.S. hyperscalers.

SoftBank Launches AI GPU Cloud in Japan With NVIDIA GB200, Targeting Sovereign AI Market

SoftBank officially launched its AI Data Center GPU Cloud this month, deploying NVIDIA GB200 NVL72 systems at Japan-based data centers under what the company calls a neocloud strategy. The launch positions SoftBank as Japan's primary sovereign AI compute alternative to U.S. hyperscalers — and signals that the race for AI infrastructure dominance has opened a significant front in Asia.

Japan has few good options for enterprises that need high-performance AI compute and want their data to stay inside the country. AWS, Google Cloud, and Azure all operate Japan regions, but they are U.S.-headquartered companies subject to American legal jurisdiction. SoftBank's offering changes that calculation.

What SoftBank Is Deploying

The AI Data Center GPU Cloud is built on current-generation hardware and managed infrastructure:

  • NVIDIA GB200 NVL72 rack systems — the highest-density GPU configuration currently available for AI inference and large model training
  • Infrinia AI Cloud OS — the operating system layer managing GPU resource allocation and workload scheduling
  • Kubernetes-as-a-Service — container orchestration for enterprise AI deployments without the operational overhead of managing the underlying cluster
  • LLM Inference-as-a-Service — hosted inference for large language models, available to customers who need AI output without running their own model infrastructure

All data remains within Japan-based data centers. That data residency guarantee is the product's core competitive differentiator for regulated industries — financial services, healthcare, and government — that cannot or will not route sensitive AI workloads through U.S.-based infrastructure.

The Sovereign AI Positioning

"Sovereign AI" refers to a nation's or organization's ability to develop and operate AI systems on domestically controlled infrastructure, under domestic legal jurisdiction, using domestic data. Japan's government has been explicit about prioritizing this capability since 2023.

SoftBank, as both Japan's largest wireless carrier and one of the world's most active technology investors through its Vision Fund portfolio, occupies a unique position to build the physical infrastructure that sovereign AI requires. The company has the real estate relationships, the power access, and the data center operational expertise that most enterprises lack — and that cloud-native AI startups typically cannot provide.

Getting allocation of GB200 NVL72 systems from NVIDIA requires either hyperscaler-scale purchasing commitments or the kind of strategic partnership relationship that SoftBank has built with NVIDIA over multiple years. Most regional cloud providers are still deploying previous-generation H100 systems. SoftBank's ability to launch on current-generation hardware is a meaningful head start in the Japanese enterprise market.

Market Context

Japan's enterprise AI adoption is accelerating but compute-constrained. Major Japanese corporations — in automotive, financial services, heavy industry, and consumer electronics — are building internal AI capabilities but have been reliant on either U.S. hyperscalers or on-premise hardware that is expensive to scale.

The compliance picture drives the decision for many of them. Japan's Act on the Protection of Personal Information (APPI), sector-specific financial regulations, and emerging AI governance guidelines from Japan's Ministry of Economy, Trade and Industry (METI) create strong incentives for enterprises to keep data within Japan's legal jurisdiction. SoftBank's service addresses that directly in a way that non-Japanese hyperscalers structurally cannot.

The competitive risk is real: AWS, Google, and Azure all offer Japan-based regions and have existing enterprise relationships. But SoftBank's Japanese corporate identity — domestic headquarters, domestic governance, Japanese-language support at the enterprise level, and a carrier relationship with most large Japanese enterprises already — gives it structural advantages that pricing alone cannot overcome.

What to Watch

SoftBank has not publicly disclosed pricing for the GPU Cloud service, making enterprise adoption velocity difficult to predict. The company's neocloud strategy requires anchor customers — likely major Japanese banks, insurance companies, or government-affiliated enterprises — to achieve the utilization rates that justify the infrastructure investment.

Watch for partnership announcements in the next two quarters. A formal deal with a tier-one Japanese financial institution or a government ministry contract would validate the sovereign AI positioning and set the commercial floor for what the service is worth to enterprises with serious compliance requirements.

Sources: Data Center Dynamics

Key Takeaways

  • ✓ Infrinia AI Cloud OS
  • ✓ LLM Inference-as-a-Service
  • ✓ All data remains within Japan-based data centers.

Did this help you understand AI better?

Your feedback helps us write more useful content.

Hector Herrera

Written by

Hector Herrera

Hector Herrera is an AI systems architect in Houston and founder of Hex AI Systems. He designs and runs AI systems in production and writes daily about how AI is reshaping business, government and everyday life. 20+ years building for the web. Houston, TX.

More from Hector →

Get tomorrow's AI briefing

Join readers who start their day with NexChron. Free, daily, no spam.

More from NexChron