Rent H200, H100 and A100 by the hour, or call Claude, GPT and Gemini per token. No contracts, no waiting.
The most powerful GPUs for your AI and HPC workloads
The most powerful GPU for LLM inference
High performance for training and inference
Best value for ML workloads
Ideal for rendering and content creation
Blackwell for frontier training and inference
Top of the line, for the largest models
For visualization and design
For development and testing
View all available options
Check all available models, configurations and prices
| Model | vRAM | vCPU | RAM | Regions | Price/Hour | |
|---|---|---|---|---|---|---|
NVIDIA B300Blackwell Ultra • SXM6 |
288 GB HBM3e | 43 | 503 GB |
Global
|
R$ 62,73 /h | Create |
NVIDIA B200Blackwell • SXM6 |
180 GB HBM3e | 24 | 188 GB |
Global
|
R$ 53,98 /h | Create |
NVIDIA H200Hopper • SXM5 |
141 GB HBM3e | 128 | 512 GB |
North America
|
R$ 31,72 /h | Create |
NVIDIA H100 SXMHopper • SXM5 |
80 GB HBM3 | 96 | 256 GB |
North America
Europe
|
R$ 25,44 /h | Create |
NVIDIA H100 PCIeHopper • PCIe |
80 GB HBM2e | 64 | 128 GB |
North America
Europe
Asia-Pacific
|
R$ 19,88 /h | Create |
NVIDIA A100Ampere • 80GB |
80 GB HBM2e | 64 | 128 GB |
North America
Europe
|
R$ 10,73 /h | Create |
NVIDIA A100Ampere • 40GB |
40 GB HBM2e | 48 | 96 GB |
North America
Europe
Asia-Pacific
|
R$ 10,73 /h | Create |
NVIDIA L40SAda Lovelace |
48 GB GDDR6 | 48 | 96 GB |
North America
Europe
|
R$ 7,87 /h | Create |
NVIDIA L40Ada Lovelace |
48 GB GDDR6 | 48 | 96 GB |
North America
Europe
Asia-Pacific
|
R$ 5,49 /h | Create |
RTX 6000 AdaAda Lovelace |
48 GB GDDR6 | 48 | 96 GB |
North America
|
R$ 5,88 /h | Create |
NVIDIA A6000Ampere |
48 GB GDDR6 | 32 | 64 GB |
North America
Europe
|
R$ 3,98 /h | Create |
RTX 4090Ada Lovelace |
24 GB GDDR6X | 32 | 64 GB |
North America
Europe
Asia-Pacific
|
R$ 5,88 /h | Create |
RTX 4080Ada Lovelace |
16 GB GDDR6X | 24 | 48 GB |
North America
Europe
|
R$ 2,23 /h | Create |
RTX 3090Ampere |
24 GB GDDR6X | 24 | 48 GB |
North America
Europe
Asia-Pacific
|
R$ 3,98 /h | Create |
NVIDIA A4000Ampere |
16 GB GDDR6 | 16 | 32 GB |
North America
Europe
|
R$ 1,19 /h | Create |
RTX 3080Ampere |
10 GB GDDR6X | 16 | 32 GB |
North America
Europe
|
R$ 1,35 /h | Create |
Claude, GPT and Gemini on a single OpenAI-compatible API. No subscription, no minimum, no currency surprise — you pay only for the tokens you use.
| Model | Shelf | Context | Input / 1M | Output / 1M | |
|---|---|---|---|---|---|
| Frontierflagship models from Anthropic, OpenAI and Google28 models · from the most capable to the most affordable | |||||
GPT-5.5 Progpt-5.5-pro |
Frontier | 128,000 | R$ 306,00 | R$ 1836,00 | Start |
Claude Fable 5claude-fable-5 |
Frontier | 1,000,000 | R$ 102,00 | R$ 510,00 | Start |
Claude Fable 5.1claude-fable-5-1 |
Frontier | 1,000,000 | R$ 102,00 | R$ 510,00 | Start |
GPT-6 Astragpt-6-astra |
Frontier | 128,000 | R$ 102,00 | R$ 510,00 | Start |
GPT-5.5gpt-5.5 |
Frontier | 128,000 | R$ 51,00 | R$ 306,00 | Start |
Claude Opus 4.7claude-opus-4-7 |
Frontier | 1,000,000 | R$ 51,00 | R$ 255,00 | Start |
Claude Opus 4.8claude-opus-4-8 |
Frontier | 1,000,000 | R$ 51,00 | R$ 255,00 | Start |
Claude Opus 5claude-opus-5 |
Frontier | 1,000,000 | R$ 51,00 | R$ 255,00 | Start |
GPT-5.6 Solgpt-5.6-sol |
Frontier | 128,000 | R$ 40,80 | R$ 204,00 | Start |
Claude Sonnet 4.6claude-sonnet-4-6 |
Frontier | 1,000,000 | R$ 30,60 | R$ 153,00 | Start |
GPT-5.4gpt-5.4 |
Frontier | 128,000 | R$ 25,50 | R$ 153,00 | Start |
Gemini 3.1 Progemini-3.1-pro-preview |
Frontier | 128,000 | R$ 20,40 | R$ 122,40 | Start |
GPT-5.6 Terragpt-5.6-terra |
Frontier | 128,000 | R$ 20,40 | R$ 122,40 | Start |
Claude Sonnet 5claude-sonnet-5 |
Frontier | 1,000,000 | R$ 20,40 | R$ 102,00 | Start |
Gemini 3.5 Flashgemini-3.5-flash |
Frontier | 128,000 | R$ 15,30 | R$ 91,80 | Start |
o3o3 |
Frontier | 128,000 | R$ 20,40 | R$ 81,60 | Start |
Claude Haiku 4.5claude-haiku-4-5 |
Frontier | 200,000 | R$ 10,20 | R$ 51,00 | Start |
GPT-5.4 minigpt-5.4-mini |
Frontier | 128,000 | R$ 7,65 | R$ 45,90 | Start |
o4-minio4-mini |
Frontier | 128,000 | R$ 11,22 | R$ 44,88 | Start |
Gemini 3.6 Flashgemini-3.6-flash |
Frontier | 128,000 | R$ 7,65 | R$ 38,25 | Start |
Gemini 3.7 Flashgemini-3.7-flash |
Frontier | 128,000 | R$ 7,65 | R$ 38,25 | Start |
Gemini 3.8 Flashgemini-3.8-flash |
Frontier | 128,000 | R$ 7,65 | R$ 38,25 | Start |
Gemini 3.1 Flash Imagegemini-3.1-flash-image |
Image | — | R$ 0,67 per image (approx.) | Start | |
Gemini 3.5 Flash Litegemini-3.5-flash-lite |
Frontier | 128,000 | R$ 3,06 | R$ 25,50 | Start |
Gemini 3.1 Flash Lite Imagegemini-3.1-flash-lite-image |
Image | — | R$ 0,38 per image (approx.) | Start | |
Gemini 3.1 Flash Litegemini-3.1-flash-lite |
Frontier | 128,000 | R$ 2,55 | R$ 15,30 | Start |
GPT-5.4 nanogpt-5.4-nano |
Frontier | 128,000 | R$ 2,04 | R$ 12,75 | Start |
GPT-5.6 Lunagpt-5.6-luna |
Frontier | 128,000 | R$ 2,04 | R$ 12,24 | Start |
| Confidentialrun inside a hardware-sealed enclave14 models · from the most capable to the most affordable | |||||
Kimi K3gpub-selado-kimi |
Confidential | 128,000 | R$ 38,95 | R$ 194,90 | Start |
GLM 5.2gpub-selado-glm52 |
Confidential | 128,000 | R$ 18,95 | R$ 58,90 | Start |
GLM 5.1gpub-selado-glm51 |
Confidential | 128,000 | R$ 17,95 | R$ 55,90 | Start |
Qwen3.8 27Bgpub-selado-qwen38 |
Confidential | 128,000 | R$ 5,95 | R$ 43,90 | Start |
Kimi K2.6gpub-selado-kimi26 |
Confidential | 128,000 | R$ 6,74 | R$ 39,54 | Start |
Qwen3.5 397Bgpub-selado-qwen397 |
Confidential | 128,000 | R$ 5,23 | R$ 34,88 | Start |
Qwen3.6 27Bgpub-selado-qwen27 |
Confidential | 128,000 | R$ 3,95 | R$ 26,90 | Start |
DeepSeek V4 Flashgpub-selado-fast |
Confidential | 128,000 | R$ 5,39 | R$ 16,16 | Start |
Qwen3 235B Thinkinggpub-selado-qwen235 |
Confidential | 128,000 | R$ 3,95 | R$ 15,90 | Start |
DeepSeek V3.2gpub-selado-deepseek |
Confidential | 128,000 | R$ 11,63 | R$ 11,63 | Start |
Gemma 4 31Bgpub-selado-gemma |
Confidential | 128,000 | R$ 1,95 | R$ 4,90 | Start |
Qwen3 32Bgpub-selado-qwen32 |
Confidential | 128,000 | R$ 0,70 | R$ 2,90 | Start |
Mistral Nemogpub-selado-nemo |
Confidential | 128,000 | R$ 0,17 | R$ 0,66 | Start |
Nemotron 3 Nanogpub-selado-nemotron |
Confidential | 128,000 | R$ 0,15 | R$ 0,60 | Start |
| Essentialsopen-weight models, billed per token10 models · from the most capable to the most affordable | |||||
Kimi K3gpub-max |
Essential | 1,000,000 | R$ 24,95 | R$ 124,90 | Start |
GLM 5.3gpub-turbo |
Essential | 128,000 | R$ 17,14 | R$ 53,86 | Start |
Ornith 1.5 397Bgpub-ultra |
Essential | 128,000 | R$ 8,57 | R$ 26,93 | Start |
GLM 5.2gpub-base |
Essential | 250,000 | R$ 8,95 | R$ 19,90 | Start |
FLUX.2 Klein 4Bgpub-imagem |
Image | — | R$ 0,19 per image | Start | |
GLM 5.3 Flashgpub-pro |
Essential | 250,000 | R$ 1,89 | R$ 5,90 | Start |
Qwen 3.8 27Bgpub-plus |
Essential | 1,000,000 | R$ 0,69 | R$ 3,99 | Start |
Qwen 3.6 35Bgpub-mini |
Essential | 200,000 | R$ 0,69 | R$ 3,90 | Start |
DeepSeek V4 Flashgpub-fast |
Essential | 1,000,000 | R$ 0,59 | R$ 1,29 | Start |
DeepSeek V4.1 Flashgpub-nano |
Essential | 250,000 | R$ 0,49 | R$ 1,09 | Start |
See each model's page → · Token API documentation →
OpenAI-compatible: swap the base URL and the key in the tools you already use. Usage draws from the same balance, billed in Brazilian reais.
The same NVIDIA GPUs you already know — RTX 4090, A100, H100 and even H200 — starting at a fraction of the price. We connect you to a global marketplace of verified GPUs, without giving up root access, SSH and deploy in seconds.
Live pricing, per hour — from R$ 0.81/h
No hidden fees. Pay only for what you use.
Pay as you go. Turn off your instance anytime and stop paying immediately. No commitments.
Get StartedWe accept all major credit cards and payment methods. Simple, transparent billing.
Create Free AccountTrusted since 2024
Launch GPU instances programmatically and automate your AI infrastructure. Our API offers full control over all resources.
View APINo currency conversion hassles. We accept all major credit cards. Transparent pricing, no hidden fees.
Start with $25Sign up for free and add funds starting at R$100. Get up to R$25 in extra credit on your first top-up.
Browse our catalog of available GPUs. From RTX 3080 to H200, choose the one that best fits your workload.
Launch in seconds, automate and run training or inference at any scale. Pay only for what you use.
GPUs by the hour and AI models per token — what people ask most.
https://gpubrazil.com/v1 and use your gpub_live_ key. Any SDK or tool that already speaks to OpenAI works without a code change. Billing is per token, in reais, from the same balance that pays for GPUs — no subscription, no minimum, and none of the Brazilian IOF tax charged on purchases abroad.POST /v1/images/generations, in the same OpenAI format and with the same key. There are three image generation models; some are billed per image and others per image token, and each price is shown live on the pricing table. Image models do not answer on the chat endpoint, and vice versa.image_url block inside content, with no change to your code. Audio input also works on the models that support it, through the input_audio block. If a model does not accept what you sent, the API refuses with an error explaining why instead of silently dropping the part it cannot read.Create your account for free and get access to the most powerful GPUs on the market.
Create Free Account →