Skip to content

GPU cloud · NVIDIA B300 SXM

Dedicated NVIDIA B300 SXM nodes, in India, the EU and the US

Dedicated 8-GPU HGX B300 nodes on a six-month term. FlexiCloud delivers the bare metal and keeps the hardware running; you bring the workload and run your own stack.

From USD 4.50 per GPU-hour. Minimum one node of 8 GPUs on a six-month commitment.

The numbers that matter

B300 GPUs per node
8
HBM3e memory per GPU
288 GB
Regions: India, EU, US
3
Per GPU-hour, from
$4.50

What we run, and what you run

The node is yours. The hardware is our job.

Most GPU providers hand you a bare node and a support ticket queue. We keep the hardware running underneath: monitored around the clock, failures handled, and an engineer you can actually reach.

  • Full control of the stack

    Root access to run your own OS image, CUDA version, container runtime and orchestration. No platform layer between you and the GPUs; the software is entirely yours.

  • Watched around the clock

    GPU health, thermals, memory errors and interconnect are monitored continuously. A failing card is caught by us, not discovered by a crashed training run at 3am.

  • An engineer, not a queue

    Support is a person who runs GPU infrastructure, in your timezone. When something is wrong you talk to someone who can fix it.

  • Data where you need it

    Nodes in India for domestic data residency, and in the EU and US for teams and datasets that live there. Same service, same team, whichever region.

  • Yours alone

    Dedicated bare-metal nodes. No shared tenancy, no noisy neighbours, no scheduler deciding when your job runs. The whole node is yours for the term.

  • Predictable cost

    A fixed term at a fixed rate. No spot-market surprises, no burst pricing, no bill that doubles because a job ran long. You know the number for the whole six months.

Terms

How it is sold

This is dedicated infrastructure on a term, not an hourly spot instance. The terms are simple and they are the same for everyone.

Minimum order

One node: 8 × B300 SXM

The unit is a full HGX B300 node. Eight GPUs with NVLink between them, dedicated to you. Need more? Nodes are added in units of eight.

  • 8 × NVIDIA B300 SXM GPUs
  • 288 GB HBM3e per GPU, 2.3 TB per node
  • NVLink interconnect across the node
  • Dedicated bare metal, single tenant
Six-month minimum

Pricing and term

From $4.50 per GPU-hour

On a six-month minimum commitment, billed for the term. Final pricing depends on region, term length and how many nodes, and we confirm it in writing before you commit.

  • From $4.50 per GPU-hour
  • Six-month minimum commitment
  • Longer terms and multi-node priced on request

Availability

India, the EU and the US

Capacity is allocated per region on a first-committed basis. Tell us where the data lives and where the team sits, and we will confirm availability and lead time for that region before you commit to anything.

  • India: for domestic data residency
  • European Union: for EU-resident data and teams
  • United States: for US-based workloads
  • Lead time confirmed per region before order

How it works

From first conversation to a running node

  1. 01

    Tell us the workload

    Training, fine-tuning or inference; the framework; how many nodes; and which region. Fifteen minutes with an engineer, not a form that disappears into a CRM.

  2. 02

    We confirm capacity and price

    A firm quote for your region and term, with the lead time to a live node. No commitment until you have the number in writing.

  3. 03

    We hand over the node

    A dedicated bare-metal node, networked and ready to log into with root access. You bring your own OS image, drivers and orchestration; the hardware underneath stays ours to keep running.

  4. 04

    You run. We keep it healthy.

    Your team owns the OS, the stack and the workload. Ours watches the hardware, handles failures, and is a phone call away for the whole term.

Questions

What people ask before they commit

  • One node, which is 8 NVIDIA B300 SXM GPUs. We do not sell fewer than a full node, because the GPUs share NVLink and the node is the unit of allocation. Additional capacity is added in whole nodes.

  • Six months. This is dedicated bare-metal infrastructure reserved for you, not a spot instance, and the term is what makes the rate possible. Longer terms are available and are priced lower.

  • Pricing starts at $4.50 per GPU-hour on a six-month term. The exact figure depends on region, term length and the number of nodes, and we confirm it in writing before you commit.

  • India, the European Union and the United States. Capacity is allocated per region, so we confirm availability and lead time for your chosen region before you order. Pick the region by where your data has to live and where your team works.

  • We deliver the node as dedicated bare metal with network and root access, and keep the hardware running: continuous monitoring of GPU health and interconnect, hardware failure handling, and an engineer you can reach directly for the whole term. The software stack on top - OS, drivers, CUDA, container runtime and orchestration - is yours to run.

  • Yes. Nodes are added in units of eight GPUs, and multi-node deployments are priced on request. Tell us the target size when we scope the workload and we will plan capacity for it.

  • Dedicated. Each node is single-tenant bare metal reserved for you for the term. No shared GPUs, no scheduler queue, no noisy neighbours.

  • No. We work with partners who own the hardware and hold the capacity in each region. FlexiCloud sources the node, delivers it to you as dedicated bare metal, and keeps the hardware running - monitored, failures handled and supported - for the whole term. The software you run on top is yours.

Tell us the workload and the region

Fifteen minutes with an engineer gets you a firm price, a lead time, and a straight answer on whether B300 is the right fit. No commitment until you have the number in writing.