GPUs Pricing AI by Token Templates Blog API Contact Login Get Started
All GPUs available

Instant access to 1,000+ NVIDIA GPUs and to AI models by the token

Rent H200, H100 and A100 by the hour, or call Claude, GPT and Gemini per token. No contracts, no waiting.

Get Started View Available GPUs
Pay by the hour, no commitment No contracts or lock-in Deploy in seconds

Production GPUs

The most powerful GPUs for your AI and HPC workloads

Extra credit

First Top-up

Get up to R$25 extra credit on your first top-up

Get Started →
NVIDIA H100
⚡ Popular

NVIDIA H100

High performance for training and inference

Memory 80 GB HBM3
vCPU 96
R$ 19,88 /hour
NVIDIA A100
💰 Value

NVIDIA A100

Best value for ML workloads

Memory 80 GB HBM2e
vCPU 64
R$ 10,73 /hour
NVIDIA L40
🎨 Rendering

NVIDIA L40

Ideal for rendering and content creation

Memory 48 GB GDDR6
vCPU 48
R$ 5,49 /hour
🚀 Blackwell

NVIDIA B200

Blackwell for frontier training and inference

Memory 180 GB HBM3e
R$ 53,98 /hour
🚀 Blackwell Ultra

NVIDIA B300

Top of the line, for the largest models

Memory 288 GB HBM3e
R$ 62,73 /hour
🔧 Professional

RTX 6000

For visualization and design

Memory 48 GB
R$ 5,88 /hour
🎮 Gaming

RTX 4090

For development and testing

Memory 24 GB
R$ 5,88 /hour
📋 Catalog

+10 GPUs

View all available options

View All →

Complete GPU Catalog

Check all available models, configurations and prices

Model vRAM vCPU RAM Regions Price/Hour

NVIDIA B300

Blackwell Ultra • SXM6
288 GB HBM3e 43 503 GB
Global
R$ 62,73 /h Create

NVIDIA B200

Blackwell • SXM6
180 GB HBM3e 24 188 GB
Global
R$ 53,98 /h Create

NVIDIA H200

Hopper • SXM5
141 GB HBM3e 128 512 GB
North America
R$ 31,72 /h Create

NVIDIA H100 SXM

Hopper • SXM5
80 GB HBM3 96 256 GB
North America Europe
R$ 25,44 /h Create

NVIDIA H100 PCIe

Hopper • PCIe
80 GB HBM2e 64 128 GB
North America Europe Asia-Pacific
R$ 19,88 /h Create

NVIDIA A100

Ampere • 80GB
80 GB HBM2e 64 128 GB
North America Europe
R$ 10,73 /h Create

NVIDIA A100

Ampere • 40GB
40 GB HBM2e 48 96 GB
North America Europe Asia-Pacific
R$ 10,73 /h Create

NVIDIA L40S

Ada Lovelace
48 GB GDDR6 48 96 GB
North America Europe
R$ 7,87 /h Create

NVIDIA L40

Ada Lovelace
48 GB GDDR6 48 96 GB
North America Europe Asia-Pacific
R$ 5,49 /h Create

RTX 6000 Ada

Ada Lovelace
48 GB GDDR6 48 96 GB
North America
R$ 5,88 /h Create

NVIDIA A6000

Ampere
48 GB GDDR6 32 64 GB
North America Europe
R$ 3,98 /h Create

RTX 4090

Ada Lovelace
24 GB GDDR6X 32 64 GB
North America Europe Asia-Pacific
R$ 5,88 /h Create

RTX 4080

Ada Lovelace
16 GB GDDR6X 24 48 GB
North America Europe
R$ 2,23 /h Create

RTX 3090

Ampere
24 GB GDDR6X 24 48 GB
North America Europe Asia-Pacific
R$ 3,98 /h Create

NVIDIA A4000

Ampere
16 GB GDDR6 16 32 GB
North America Europe
R$ 1,19 /h Create

RTX 3080

Ampere
10 GB GDDR6X 16 32 GB
North America Europe
R$ 1,35 /h Create

Buy Frontier Tokens at a Discount

Claude, GPT and Gemini on a single OpenAI-compatible API. No subscription, no minimum, no currency surprise — you pay only for the tokens you use.

No subscription No FX surprises Pay per token One base URL, every model
52 models · price per 1 million tokens, in Brazilian reais
Model Shelf Context Input / 1M Output / 1M
Frontierflagship models from Anthropic, OpenAI and Google28 models · from the most capable to the most affordable

GPT-5.5 Pro

gpt-5.5-pro
Frontier 128,000 R$ 306,00R$ 1836,00 Start

Claude Fable 5

claude-fable-5
Frontier 1,000,000 R$ 102,00R$ 510,00 Start

Claude Fable 5.1

claude-fable-5-1
Frontier 1,000,000 R$ 102,00R$ 510,00 Start

GPT-6 Astra

gpt-6-astra
Frontier 128,000 R$ 102,00R$ 510,00 Start

GPT-5.5

gpt-5.5
Frontier 128,000 R$ 51,00R$ 306,00 Start

Claude Opus 4.7

claude-opus-4-7
Frontier 1,000,000 R$ 51,00R$ 255,00 Start

Claude Opus 4.8

claude-opus-4-8
Frontier 1,000,000 R$ 51,00R$ 255,00 Start

Claude Opus 5

claude-opus-5
Frontier 1,000,000 R$ 51,00R$ 255,00 Start

GPT-5.6 Sol

gpt-5.6-sol
Frontier 128,000 R$ 40,80R$ 204,00 Start

Claude Sonnet 4.6

claude-sonnet-4-6
Frontier 1,000,000 R$ 30,60R$ 153,00 Start

GPT-5.4

gpt-5.4
Frontier 128,000 R$ 25,50R$ 153,00 Start

Gemini 3.1 Pro

gemini-3.1-pro-preview
Frontier 128,000 R$ 20,40R$ 122,40 Start

GPT-5.6 Terra

gpt-5.6-terra
Frontier 128,000 R$ 20,40R$ 122,40 Start

Claude Sonnet 5

claude-sonnet-5
Frontier 1,000,000 R$ 20,40R$ 102,00 Start

Gemini 3.5 Flash

gemini-3.5-flash
Frontier 128,000 R$ 15,30R$ 91,80 Start

o3

o3
Frontier 128,000 R$ 20,40R$ 81,60 Start

Claude Haiku 4.5

claude-haiku-4-5
Frontier 200,000 R$ 10,20R$ 51,00 Start

GPT-5.4 mini

gpt-5.4-mini
Frontier 128,000 R$ 7,65R$ 45,90 Start

o4-mini

o4-mini
Frontier 128,000 R$ 11,22R$ 44,88 Start

Gemini 3.6 Flash

gemini-3.6-flash
Frontier 128,000 R$ 7,65R$ 38,25 Start

Gemini 3.7 Flash

gemini-3.7-flash
Frontier 128,000 R$ 7,65R$ 38,25 Start

Gemini 3.8 Flash

gemini-3.8-flash
Frontier 128,000 R$ 7,65R$ 38,25 Start

Gemini 3.1 Flash Image

gemini-3.1-flash-image
Image R$ 0,67 per image (approx.) Start

Gemini 3.5 Flash Lite

gemini-3.5-flash-lite
Frontier 128,000 R$ 3,06R$ 25,50 Start

Gemini 3.1 Flash Lite Image

gemini-3.1-flash-lite-image
Image R$ 0,38 per image (approx.) Start

Gemini 3.1 Flash Lite

gemini-3.1-flash-lite
Frontier 128,000 R$ 2,55R$ 15,30 Start

GPT-5.4 nano

gpt-5.4-nano
Frontier 128,000 R$ 2,04R$ 12,75 Start

GPT-5.6 Luna

gpt-5.6-luna
Frontier 128,000 R$ 2,04R$ 12,24 Start
Confidentialrun inside a hardware-sealed enclave14 models · from the most capable to the most affordable

Kimi K3

gpub-selado-kimi
Confidential 128,000 R$ 38,95R$ 194,90 Start

GLM 5.2

gpub-selado-glm52
Confidential 128,000 R$ 18,95R$ 58,90 Start

GLM 5.1

gpub-selado-glm51
Confidential 128,000 R$ 17,95R$ 55,90 Start

Qwen3.8 27B

gpub-selado-qwen38
Confidential 128,000 R$ 5,95R$ 43,90 Start

Kimi K2.6

gpub-selado-kimi26
Confidential 128,000 R$ 6,74R$ 39,54 Start

Qwen3.5 397B

gpub-selado-qwen397
Confidential 128,000 R$ 5,23R$ 34,88 Start

Qwen3.6 27B

gpub-selado-qwen27
Confidential 128,000 R$ 3,95R$ 26,90 Start

DeepSeek V4 Flash

gpub-selado-fast
Confidential 128,000 R$ 5,39R$ 16,16 Start

Qwen3 235B Thinking

gpub-selado-qwen235
Confidential 128,000 R$ 3,95R$ 15,90 Start

DeepSeek V3.2

gpub-selado-deepseek
Confidential 128,000 R$ 11,63R$ 11,63 Start

Gemma 4 31B

gpub-selado-gemma
Confidential 128,000 R$ 1,95R$ 4,90 Start

Qwen3 32B

gpub-selado-qwen32
Confidential 128,000 R$ 0,70R$ 2,90 Start

Mistral Nemo

gpub-selado-nemo
Confidential 128,000 R$ 0,17R$ 0,66 Start

Nemotron 3 Nano

gpub-selado-nemotron
Confidential 128,000 R$ 0,15R$ 0,60 Start
Essentialsopen-weight models, billed per token10 models · from the most capable to the most affordable

Kimi K3

gpub-max
Essential 1,000,000 R$ 24,95R$ 124,90 Start

GLM 5.3

gpub-turbo
Essential 128,000 R$ 17,14R$ 53,86 Start

Ornith 1.5 397B

gpub-ultra
Essential 128,000 R$ 8,57R$ 26,93 Start

GLM 5.2

gpub-base
Essential 250,000 R$ 8,95R$ 19,90 Start

FLUX.2 Klein 4B

gpub-imagem
Image R$ 0,19 per image Start

GLM 5.3 Flash

gpub-pro
Essential 250,000 R$ 1,89R$ 5,90 Start

Qwen 3.8 27B

gpub-plus
Essential 1,000,000 R$ 0,69R$ 3,99 Start

Qwen 3.6 35B

gpub-mini
Essential 200,000 R$ 0,69R$ 3,90 Start

DeepSeek V4 Flash

gpub-fast
Essential 1,000,000 R$ 0,59R$ 1,29 Start

DeepSeek V4.1 Flash

gpub-nano
Essential 250,000 R$ 0,49R$ 1,09 Start

See each model's page →  ·  Token API documentation →

OpenAI-compatible: swap the base URL and the key in the tools you already use. Usage draws from the same balance, billed in Brazilian reais.

✦ New

Economy GPUs

The same NVIDIA GPUs you already know — RTX 4090, A100, H100 and even H200 — starting at a fraction of the price. We connect you to a global marketplace of verified GPUs, without giving up root access, SSH and deploy in seconds.

Live pricing, per hour — from R$ 0.81/h

✓ Pay per hour ✓ Root access + SSH ✓ Deploy in seconds ✓ No contracts, no waiting
Loading live pricing…
See all Economy GPUs →

Transparent Pricing

No hidden fees. Pay only for what you use.

💳 Pay-Per-Hour

Pay as you go. Turn off your instance anytime and stop paying immediately. No commitments.

Get Started

🌎 Pay in USD

We accept all major credit cards and payment methods. Simple, transparent billing.

Create Free Account

Built by
Developers,
for Developers

Trusted since 2024

Full API Access

Launch GPU instances programmatically and automate your AI infrastructure. Our API offers full control over all resources.

View API

Simple Billing

No currency conversion hassles. We accept all major credit cards. Transparent pricing, no hidden fees.

Start with $25

Get Started in Minutes

1

Create Your Account

Sign up for free and add funds starting at R$100. Get up to R$25 in extra credit on your first top-up.

2

Choose Your GPU

Browse our catalog of available GPUs. From RTX 3080 to H200, choose the one that best fits your workload.

3

Deploy & Scale

Launch in seconds, automate and run training or inference at any scale. Pay only for what you use.

Get Started Documentation

Frequently asked questions

GPUs by the hour and AI models per token — what people ask most.

How much does it cost to rent a GPU?
Billing is hourly and in Brazilian reais (R$), with live prices shown on the site — from entry-level RTX cards to NVIDIA H200, B200 and B300. You only pay for the time you use, with no contract.
How fast can I deploy a GPU instance?
GPU instances are deployed in seconds. After payment, your GPU is ready to use immediately with SSH access.
What GPUs are available?
We offer NVIDIA H200, H100, A100, L40, RTX 6000, RTX 4090, RTX 4080, RTX 3090, and more. All GPUs come with dedicated vCPU, RAM, and storage.
Can I use Claude, GPT and Gemini through the GPUBrasil API, paying in Brazilian reais?
Yes. The API is OpenAI-compatible: point the base URL at https://gpubrazil.com/v1 and use your gpub_live_ key. Any SDK or tool that already speaks to OpenAI works without a code change. Billing is per token, in reais, from the same balance that pays for GPUs — no subscription, no minimum, and none of the Brazilian IOF tax charged on purchases abroad.
Can I generate images through the API?
Yes. The endpoint is POST /v1/images/generations, in the same OpenAI format and with the same key. There are three image generation models; some are billed per image and others per image token, and each price is shown live on the pricing table. Image models do not answer on the chat endpoint, and vice versa.
Does the API accept images as input (vision)?
Yes, on every model in the catalogue that accepts images — Claude, GPT, Gemini and the open-weight ones. Use the standard OpenAI image_url block inside content, with no change to your code. Audio input also works on the models that support it, through the input_audio block. If a model does not accept what you sent, the API refuses with an error explaining why instead of silently dropping the part it cannot read.
What is the difference between renting a GPU and using the per-token API?
Renting a GPU gives you the whole machine, with root and SSH access: you pick the model, control the weights, the logs and the data, and pay by the hour while it runs — that is the route for training, fine-tuning, video generation or serving your own model. The per-token API needs no machine at all: you call the model and pay only for the tokens of each call. Plenty of people use both, from the same balance.

Ready to get started?

Create your account for free and get access to the most powerful GPUs on the market.

Create Free Account →