Available now · Limited deployment

NVIDIA RTX PRO™ 6000
Blackwell Server Edition

GPU instances on Akamai Cloud, purpose-built for distributed AI inference. Up to 8 RTX PRO 6000 Blackwell GPUs per Linode, 96 GB of GDDR7 VRAM per card, and the throughput to serve large models at scale — delivered as an on-demand OpEx, not a CapEx.

96 GB
GDDR7 VRAM per GPU · ECC
24,064
CUDA cores
120
TFLOPS FP32
8
GPUs per Linode

Akamai Cloud computing

Turn vision into reality.

Akamai Cloud lets you choose the right GPU for your workload. And for the scale and complexity of data-center workloads — especially AI inference — NVIDIA RTX PRO™ 6000 Blackwell Server Edition is engineered for the job. Deploy cloud GPUs on demand, pay per hour, and scale ML, AI, and data-processing workloads with confidence.

Dedicated, competition-free resources

A GPU Linode’s vCPU cores are dedicated to you — never shared between customers. Your software runs at peak speed and efficiency, even at 100% CPU all day, every day.

On-demand, pay-as-you-go

Turn GPU CapEx into OpEx. Predictable hourly pricing and low-cost egress (US$0.005/GB in most regions) let you test and scale without draining your infrastructure budget.

Redefining inference performance

Up to 60% lower latency and 3× higher throughput — for up to 86% less cost with image generation and AI workloads compared to equivalent hyperscaler GPUs.

Deploy in minutes

Configure compute, memory, and storage to optimize for your workload, then launch GPU instances across Akamai Cloud locations — meeting your users and data wherever they are.

Automate your pipeline

Manage infrastructure flexibly with our UI, API, CLI, Terraform provider, and developer tool integrations. Custom images and CI/CD pipelines supported.

24/7/365 support

Every GPU Linode ships with email and phone support for all customers. Backups keep your data safe with automated daily, weekly, and biweekly snapshots.

NVIDIA RTX PRO™ 6000 Blackwell Server Edition

Engineered for data-center scale.

Each card pairs 5th-generation Tensor Cores and 4th-generation RT Cores with 96 GB of GDDR7 ECC memory — 24,064 CUDA cores pushing 120 TFLOPS of FP32 performance. That architecture is ideally suited for AI inferencing, providing the throughput needed for large-scale model deployment.

96 GB
GDDR7 VRAM with ECC
24,064
CUDA cores · parallel processing
752
Tensor cores · 5th Gen AI/ML
188
RT cores · 4th Gen ray tracing
120
TFLOPS FP32 performance
1–8
GPUs per Linode

Fleet includes NVIDIA RTX PRO™ 6000 Blackwell Server Edition, NVIDIA RTX™ 4000 Ada, and NVIDIA Quadro RTX™ 6000. Plans are equally at home with graphics, visualization, and video workflows.

Plans & pricing

Scale from 1 to 8 GPUs per Linode.

The same RTX PRO 6000 Blackwell performance, in sizes that fit your workload — with dedicated vCPU cores and generous RAM. Pricing is per GPU per hour (US$3.00/GPU-hr on-demand, as published; confirm live pricing at linode.com/pricing).

Configuration 1 GPU 2 GPU 4 GPU 8 GPU
GPU cards1248
GPU memory (VRAM)96 GB192 GB384 GB768 GB
vCPU cores (dedicated)163264128
Memory (RAM)176 GB352 GB704 GB1,408 GB
Storage1,024 GB6,597 GB
Network bandwidth (outbound)16 Gbps
$3.00/ GPU / hr
RTX PRO 6000 Blackwell · on-demand
$24/ 8-GPU Linode / hr
top of range · 8×96 GB cards

8-card plans add vNUMA: Virtual Non-Uniform Memory Access exposes the native PCIe device topology to the VM, so job schedulers and GPU communication libraries map hardware-localized resources efficiently. Some new accounts may require a $100 deposit to deploy GPU Linodes.

Recommended workloads

Built for AI inference and beyond.

GPU Linodes are optimized for high-throughput, low-latency inference at production scale — with large GPU memory, next-generation Tensor Cores, and architectural efficiency that sustains token throughput and fast first-response latency.

Agentic & multimodal AI

96 GB of VRAM per GPU in a high-throughput architecture mitigates the “bottlenecking” found in shared cloud resources — powering agents that process text, visuals, and audio in real time.

Real-time conversational AI

Process, reason, and respond within a line of thought. Akamai Cloud has the GPUs to take conversations global — scaling out while keeping latency direct and predictable.

Physical AI & computer vision

RT, Tensor, and CUDA cores accelerate perception pipelines — from robotics and autonomous systems to real-time analytics on streaming video.

Live video transcoding

GPU-accelerated encoding converts massive streams in real time. Pair with the NVIDIA RTX™ 4000 Ada plan for live 8K transcoding and AI upscaling at the edge.

Rendering & simulation

Real-time ray tracing and advanced shading — mesh shading, variable rate shading, and multi-view rendering — in a single GPU.

Big data analysis

Give Hadoop, Spark, and Storm the parallel compute they need to process terabyte-scale datasets — the volume, velocity, and variety of modern data.

Loved by popular frameworks: PyTorch TensorFlow FFmpeg Apache Spark Hadoop CUDA / C++

How it works

From vision to value in four steps.

1

Select

Choose NVIDIA RTX PRO™ 6000 Blackwell or Ada architectures, matching performance requirements to your budget.

2

Configure

Customize compute, memory, and storage to optimize for your unique workload.

3

Deploy

Launch GPU instances in minutes across Akamai Cloud locations, meeting your users and data wherever they are.

4

Realize

Achieve ambitious AI goals with high-performance NVIDIA compute designed for immersion, autonomy, and scale.

Global availability

Deploy wherever your users are.

NVIDIA RTX PRO 6000 Blackwell Server Edition is rolling out in 20 regions worldwide (limited deployment availability) — from Amsterdam to Tokyo, Singapore to Toronto.

Amsterdam, NL
Chennai, IN
Chicago, US
Frankfurt, DE
Jakarta, ID
London, UK
Los Angeles, US
Madrid, ES
Miami, US
Mumbai, IN
Milan, IT
Newark, US
Osaka, JP
Paris, FR
Seattle, US
Singapore, SG
Stockholm, SE
Tokyo, JP
Toronto, CA
Washington D.C., US

RTX 4000 Ada is available in Chicago 2, Frankfurt 2, Osaka, Paris, Seattle, and Singapore. Quadro RTX 6000 is available in Atlanta, Newark, Frankfurt, Mumbai, and Singapore.

Ready to deploy your AI strategy?

Start with up to US$100 in Akamai Cloud credits. Launch NVIDIA RTX PRO™ 6000 Blackwell GPU Linodes in minutes — managed Kubernetes (LKE), backups, and 24/7 support included.