New - NVIDIA B200 & B300 Dedicated GPU Clusters  |  Explore GPU clusters

Hero Mobile Solutions Wille Tornqvist - Ai/Ml Gpu Servers
Hero Desktop Solutions Wille Tornqvist - Ai/Ml Gpu Servers

Powerful, private, and cost-efficient GPU Cloud for AI/ML

UpCloud’s NVIDIA Cloud GPUs are engineered for the demands of modern AI/ML, offering the performance you need without the hidden fees or vendor lock-in.

Power your projects, from LLM inference to complex machine learning tasks, right from our European cloud.

Build AI without giving up control

Wille, IT Engineer

  • Privacy-first by design

    Host your AI workloads in Finland, within GDPR-compliant data centers, backed by strong jurisdictional protections. This means your models and data are secured without public cloud exposure.

  • Sustainable computing

    Innovate responsibly. Our Helsinki data center is powered by 100% renewables and routes excess heat generated from our GPUs into the city’s district heating network. This makes our GPU offering in Helsinki one of the most environmentally friendly options on the market, contributing to a greener future while powering your AI.

  • No platform dependencies

    We don’t force you into proprietary ML platforms or opaque orchestration layers. Whether you’re fine-tuning LLaMA, deploying open-source LLMs, or building on PyTorch, our infrastructure supports your chosen open-source tools without restriction.

Seamless Integration

Robust AI/ML Ecosystem

5Th Gen Amd Epyc Turin Hardware Benchmark Comparison - Upcloud.com

Balanced architecture with AMD EPYC Servers

Pair your GPU Servers with our high-performance 5th-gen AMD EPYC Cloud Servers for a complete and perfectly balanced architecture. Conquer massive parallel jobs and general tasks without compromise in a single architecture.

GPU Server Hardware

NVIDIA L4

Upcloud Gpu Nvidia L4

Cost-effective AI Inference, High-throughput Video Processing, and Edge AI deployments. Optimized for low-latency, energy-efficient operations.

The NVIDIA L4 Tensor Core GPU powered by the NVIDIA Ada Lovelace architecture delivers universal, energy-efficient acceleration for video, AI, visual computing, graphics, virtualization, and more.

NVIDIA L40S

Upcloud Gpu Nvidia L40S

Generative AI, Mid-to-Large Scale AI Model Training and Inference, Real-time 3D Rendering, Virtual Production, and High-Performance Computing (HPC) simulations.

The NVIDIA L40S GPU is the most powerful universal GPU for the cloud, delivering end-to-end acceleration for the next generation of AI-enabled applications.

NVIDIA RTX PRO 6000 Blackwell Server (Coming soon)

Upcloud Gpu Nvidia Rtx Pro 6000 Blackwell Server

Agentic and Generative AI, LLM Inference, AI-Driven Rendering, 3D Graphics, Video Processing, Scientific Computing, and Data Analytics.

The NVIDIA RTX PRO 6000 Blackwell Server Edition delivers powerful AI and visual computing performance for demanding enterprise workloads. Powered by the NVIDIA Blackwell architecture and equipped with 96 GB of GDDR7 memory, it accelerates complex AI models, photorealistic rendering, simulations, and data-intensive workflows.

NVIDIA H100

Upcloud H100 Gpu Server

The NVIDIA H100 GPU delivers exceptional performance, scalability, and security for every workload. H100 uses breakthrough innovations based on the NVIDIA Hopper™ architecture to deliver industry-leading conversational AI, speeding up large language models (LLMs) by 30X.

H100 also includes a dedicated Transformer Engine to solve trillion-parameter language models.

NVIDIA B200

Upcloud Gpu Nvidia B200

Generative AI, Large-Scale AI Model Training and Inference (LLMs), High-Performance Computing (HPC) simulations, Scientific Computing, and Data Analytics.

The NVIDIA B200 GPU, powered by the Blackwell architecture, is the world’s most powerful AI chip, designed to power a new era of computing with up to 4x faster training and 30x faster inference than previous generations.

NVIDIA B300 (Coming soon)

Upcloud Gpu Nvidia H200

Advanced AI Reasoning, Large-Scale AI Model Training and Inference, Agentic and Multimodal AI, High-Performance Computing (HPC), Scientific Computing, and AI Factory deployments.

The NVIDIA B300 GPU, powered by the Blackwell Ultra architecture, is designed for the next generation of AI reasoning and large-scale accelerated computing. With up to 288 GB of HBM3e memory and 8 TB/s of memory bandwidth, B300 delivers 1.5x greater dense FP4 performance and 2x higher attention performance than B200, accelerating long-context models, complex inference, and training at data-center scale.

Introducing GPU Spot Instances

Get up to 25% off selected GPUs

Spot instances run on our spare capacity. If that capacity is needed elsewhere in our network, the instance will be terminated. But if your workloads are stateless, fault-tolerant, or can easily pause and resume, you’re getting top-tier GPU performance at a massive discount.

Read Documentation

Intense Workloads Illustration

Compare GPU models

GPU Model Specifications

UpCloud GPU Servers offer a range of GPU models with the unique ability to scale the server and GPUs as needed.
Start with a smaller card for development purposes, then migrate to use more powerful cards for production use, or vice-versa, all without needing to reinstalling the server or losing any data.
It’s as easy as shutting down the server, changing the plan and powering the server back up again!

NVIDIA L4 NVIDIA L40S NVIDIA RTX PRO 6000 Blackwell Server Edition NVIDIA H100 NVIDIA B200 NVIDIA B300
Primary use case Energy-efficient accelerator for AI inference,
video transcoding, graphics/VDI
and edge deployments
Multi-workload “universal” GPU – GenAI,
LLM training & inference, 3D graphics,
rendering, video
Agentic and generative AI,
3D rendering, visual computing,
video processing, scientific computing
and data analytics
High-traffic inference,
massive batch processing
and large-model training
Trillion-parameter model inference,
model training, running complex
models in real-time
Advanced AI reasoning, large-scale
model training and inference,
agentic and multimodal AI, HPC
and AI factory workloads
Architecture Ada Lovelace Ada Lovelace Blackwell Hopper Blackwell Blackwell Ultra
GPU Memory 24GB GDDR6 48GB GDDR6 96GB GDDR7 80GB HBM3e 180GB HBM3e 288GB HBM3e
Memory Bandwidth 300GB/s 864GB/s 1.6TB/s 3.35TB/s 8.0TB/s 8.0TB/s
Peak compute FP32 30.3 TFLOPS
FP8 Tensor 0.48 PFLOPS
FP32 91.6 TFLOPS
FP8 Tensor 1.46 PFLOPS
FP32 120 TFLOPS
FP8 Tensor 2.00 PFLOPS
FP32 67 TFLOPS
FP8 Tensor 3.90 PFLOPS
FP32 74.45 TFLOPS
FP8 Tensor 9.00 PFLOPS
FP8 Tensor 9.00 PFLOPS
FP4 Tensor 18.00 PFLOPS

Unleashed performance with NVIDIA L40S GPUs

Power your most demanding AI/ML tasks with enterprise-grade hardware.

High throughput & reliability

Our Cloud GPU servers are built on enterprise-grade infrastructure with redundancy at every level, delivering consistent performance and reliability for generative AI, LLM inference, and AI model training.

Designed for AI/ML

Our NVIDIA GPUs are specifically suitable for fine-tuning and inference, enabling high throughput at flexible pricing for LLM builders and inference teams.

GPU Server configurations

Configurations

GPU Servers with NVIDIA L4

GPUs

1 – 3 per server

CPU cores

8 – 32

Memory

64 – 384 GB

  • Premium AMD CPUs
  • Up to 100k IOPS with MaxIOPS
  • Choice of Block Storage
  • 1000 Mbps networking
  • Zero egress fees
  • 99.999% SLA

Starting from

€0.57/hour

Sign up

GPU CPU cores RAM Spot Price Price
1 x NVIDIA L4 8 cores 64 GB €0.57/h
€410/mo
€0.58/h
€418/mo
1 x NVIDIA L4 12 cores 128 GB €0.69/h
€497/mo
€0.70/h
€504/mo
1 x NVIDIA L4 16 cores 192 GB €0.81/h
€583/mo
€0.82/h
€590/mo
1 x NVIDIA L4 20 cores 256 GB €0.93/h
€670/mo
€0.94/h
€677/mo
2 x NVIDIA L4 12 cores 128 GB €1.17/h
€842/mo
€1.18/h
€850/mo
2 x NVIDIA L4 16 cores 192 GB €1.29/h
€929/mo
€1.30/h
€936/mo
2 x NVIDIA L4 20 cores 256 GB €1.41/h
€1015/mo
€1.42/h
€1022/mo
2 x NVIDIA L4 32 cores 384 GB €1.63/h
€1174/mo
€1.64/h
€1181/mo
3 x NVIDIA L4 16 cores 192 GB €1.79/h
€1289/mo
€1.80/h
€1296/mo
3 x NVIDIA L4 20 cores 256 GB €1.91/h
€1375/mo
€1.92/h
€1382/mo
3 x NVIDIA L4 32 cores 384 GB €2.13/h
€1534/mo
€2.14/h
€1541/mo

Save on your monthly bills with Zero-cost egress

Forget unpredictable network bills at the end of the month. With zero-cost egress, you’ll never see a surprise bill for transfer usage.

Never pay for network transfer, even when you scale up, you can redirect savings towards accelerating your business growth.

Flexible GPU Access, Zero Lock-in

Gpu Servers Gray - Ai/Ml Gpu Servers

Instant start, zero commitment

Spin up NVIDIA L40S GPU resources directly from your UpCloud Hub when you need them. Go live today without talking to sales or signing 12-month agreements.

Cost-efficient GPU power for real workloads

Stop paying for GPU time you don’t use. Our unique hourly billing model is designed to align with the dynamic nature of AI/ML development, from training sprints to intermittent inference jobs.

Pay only when active

GPU servers are not billed when the server is shut down. Perfect for teams needing high-end GPU power without long-term commitments. Note: Storage and IP addresses are billed monthly.

Avoid overprovisioning

Traditional monthly GPU billing can lead to waste. Our usage-based model lets you scale compute precisely to your workload, avoiding overprovisioning and gaining significant cost optimization compared to fixed monthly plans.

Transparent pricing

Benefit from clear, usage-based pricing with no hidden egress charges or resource bundling. This is a distinct advantage for users currently using providers like AWS, GCP, and Linode, who are seeking lower costs and better billing models.

GPU Servers with 100% renewable energy

Experience the power of NVIDIA L40S GPUs!

By using sustainable infrastructure, the waste heat generated by the servers is collected and utilized in the district heating network, warming local homes.

Location: Helsinki, Finland
Processor: 8 vCPUs AMD EPYC 9575F
Memory: 64 GB DDR5 RAM
Price: from €1.11 / hour

Your own AI sandbox, ready in minutes.

Skip the API fees and data privacy concerns. Our new tutorial shows you how to spin up an UpCloud GPU and run powerful open-weight models like Mistral-7B with Ollama. Go from deployment to inference on your own private, high-performance server.

Ready to build?

Easy Setup - Ai/Ml Gpu Servers

What you get with GPU Servers on UpCloud

Ubuntu AI/ML-ready

Get started quick and easy with the AI/ML-ready GPU Ubuntu template, which comes pre-configured with many of the tools and drivers required for GPU workloads, saving you setup time and ensuring compatibility.

Zero data transfer fees

Ensure high availability and seamless traffic distribution. Eliminate unexpected network transfer fees impacting your monthly bills. Enjoy predictable pricing and focus on building, not budgeting.

99.999% SLA

Strong focus on reliability backed by N+1 redundancy on every business-critical component in our infrastructure ensures resilience by design with a 99.999% Service Level Agreement.

Compliance certifications

Wide range of certifications depending on the data center including but not limited to: ISO 27001, SOC 2 Type II, PCI-DSS, HIPAA, NIST 800-53, and GDPR in all countries of the European Economic Area.

Global private network

High performance, private connectivity across our global network by creating isolated environments within zones and only allowing traffic through a Cloud Server acting as a firewall and router.

24/7 Customer support

Always available support, ensuring your infrastructure runs smoothly, around the clock. We pride ourselves on providing outstanding customer service – 24 hours a day, 365 days a year.

Ready to get started?

Unlock powerful, private, and cost-efficient GPU computing for your AI/ML projects today.

GPU FAQ

Frequently asked questions

You're viewing the EU site. Switch to the Global site
Back to top