Skip to main content
Talk to Sales

Pricing

Clear billing models. Enterprise terms.

Two service lines, transparent billing models — per GPU-second, per token, and per node-hour — each with enterprise agreements for regulated workloads. Final terms are agreed in your service agreement.

GPU plans

Dedicated GPU environments, sized to your work.

Dedicated, isolated GPU infrastructure in Japan for training and inference. Billing depends on the plan you choose — final configuration and terms depend on your workload.

Research

For research labs and ML teams getting started on dedicated capacity.

GPU usage metered per second — environment included

  • Dedicated GPU environment
  • One-click clusters in under 60 seconds
  • Data residency in Japan
  • Business-hours technical support
Most popular

Business

For production training and inference workloads that need isolation and support.

Reserved capacity billed per node-hour

  • Everything in Research
  • Isolated single-tenant networking
  • Full infrastructure audit logging
  • Priority incident response
  • Named technical account manager

Enterprise

For banks and regulated enterprises with bespoke compliance and architecture requirements.

Custom capacity and terms, agreed individually

  • Everything in Business
  • VPC peering and private endpoints
  • Contractual data-residency and zero-retention terms
  • 24/7 Japanese-language support
  • Procurement documentation and MSA

Billing details are confirmed in your service agreement.

LLM services

Per-token and per-GPU-second billing.

Two ways to run models — dedicated deployments billed per GPU-second, or shared endpoints billed per token — all through one OpenAI-compatible API.

Open-weight models

Hosted on Tara Cloud GPUs, in Japan

Open-weight models served from Tara Cloud's own GPU infrastructure in Japan — every prompt and completion stays in-country.

Per token on shared endpoints — or per GPU-second on dedicated deployments

  • OpenAI-compatible API
  • Fixed shared performance
  • Spending limits you set

Frontier models

Official OpenAI & Anthropic, resold by Tara Cloud

The official APIs, on the same terms — with unified JPY billing through Tara Cloud. OpenAI and Anthropic's official models are served through each vendor's official API, so their inference may run outside Japan.

Per token, with one monthly JPY invoice

  • Official API access
  • Unified JPY billing, a single invoice
  • One API key for both service lines

Enterprise terms for regulated workloads

  • Rate-limit allowances that scale with your usage commitment
  • Private endpoints and VPC peering
  • Zero-retention mode as a contractual term
  • Usage analytics and quota management

Let's talk about pricing.

Tell us about your workloads and we will prepare a tailored proposal — usually within two business days.