Starting from
€0.57/hour
UpCloud’s NVIDIA Cloud GPUs are engineered for the demands of modern AI/ML, offering the performance you need without the hidden fees or vendor lock-in.
Power your projects, from LLM inference to complex machine learning tasks, right from our European cloud.
Build AI without giving up control
Wille, IT Engineer
Host your AI workloads in Finland, within GDPR-compliant data centers, backed by strong jurisdictional protections. This means your models and data are secured without public cloud exposure.
Innovate responsibly. Our Helsinki data center is powered by 100% renewables and routes excess heat generated from our GPUs into the city’s district heating network. This makes our GPU offering in Helsinki one of the most environmentally friendly options on the market, contributing to a greener future while powering your AI.
We don’t force you into proprietary ML platforms or opaque orchestration layers. Whether you’re fine-tuning LLaMA, deploying open-source LLMs, or building on PyTorch, our infrastructure supports your chosen open-source tools without restriction.
Seamless Integration
Pair your GPU Servers with our high-performance 5th-gen AMD EPYC Cloud Servers for a complete and perfectly balanced architecture. Conquer massive parallel jobs and general tasks without compromise in a single architecture.
Fuel your workloads with our high-speed Managed Object Storage, ideal for storing immutable files like model checkpoints and large datasets. It supports stateless and concurrent access from any number of hosts, with built-in object versioning.
Coming soon, our Managed Kubernetes service will allow you to manage mixed workloads, enabling you to combine non-GPU and GPU workloads within the same cluster. Benefit from autoscaling support for your inference layer, multi-model routing with ease (e.g., for A/B testing models), and simplified observability.
Pair your GPU Servers with our high-performance 5th-gen AMD EPYC Cloud Servers for a complete and perfectly balanced architecture. Conquer massive parallel jobs and general tasks without compromise in a single architecture.
Fuel your workloads with our high-speed Managed Object Storage, ideal for storing immutable files like model checkpoints and large datasets. It supports stateless and concurrent access from any number of hosts, with built-in object versioning.
Coming soon, our Managed Kubernetes service will allow you to manage mixed workloads, enabling you to combine non-GPU and GPU workloads within the same cluster. Benefit from autoscaling support for your inference layer, multi-model routing with ease (e.g., for A/B testing models), and simplified observability.
Cost-effective AI Inference, High-throughput Video Processing, and Edge AI deployments. Optimized for low-latency, energy-efficient operations.
The NVIDIA L4 Tensor Core GPU powered by the NVIDIA Ada Lovelace architecture delivers universal, energy-efficient acceleration for video, AI, visual computing, graphics, virtualization, and more.
Generative AI, Mid-to-Large Scale AI Model Training and Inference, Real-time 3D Rendering, Virtual Production, and High-Performance Computing (HPC) simulations.
The NVIDIA L40S GPU is the most powerful universal GPU for the cloud, delivering end-to-end acceleration for the next generation of AI-enabled applications.
The NVIDIA H100 GPU delivers exceptional performance, scalability, and security for every workload. H100 uses breakthrough innovations based on the NVIDIA Hopper™ architecture to deliver industry-leading conversational AI, speeding up large language models (LLMs) by 30X.
H100 also includes a dedicated Transformer Engine to solve trillion-parameter language models.
Generative AI, Large-Scale AI Model Training and Inference (LLMs), High-Performance Computing (HPC) simulations, Scientific Computing, and Data Analytics.
The NVIDIA B200 GPU, powered by the Blackwell architecture, is the world’s most powerful AI chip, designed to power a new era of computing with up to 4x faster training and 30x faster inference than previous generations.
UpCloud GPU Servers offer a range of GPU models with the unique ability to scale the server and GPUs as needed.
Start with a smaller card for development purposes, then migrate to use more powerful cards for production use, or vice-versa, all without needing to reinstalling the server or losing any data.
It’s as easy as shutting down the server, changing the plan and powering the server back up again!
| NVIDIA L4 | NVIDIA L40S | NVIDIA H100 | NVIDIA B200 | |
|---|---|---|---|---|
| Primary use case | Energy-efficient accelerator for AI inference, video transcoding, graphics/VDI and edge deployments |
Multi-workload “universal” GPU – GenAI, LLM training & inference, 3D graphics, rendering, video |
High-traffic inference, massive batch processing and large-model training |
Trillion-parameter model inference, model training, running complex models in real-time |
| Architecture | Ada Lovelace | Ada Lovelace | Hopper | Blackwell |
| GPU Memory | 24GB GDDR6 | 48GB GDDR6 | 80GB HBM3e | 192GB HBM3e |
| Memory Bandwidth | 300GB/s | 864GB/s | 3.35TB/s | 8.0TB/s |
| Peak compute | FP32 30.3 TFLOPS FP8 Tensor 0.48 PFLOPS |
FP32 91.6 TFLOPS FP8 Tensor 1.46 PFLOPS |
FP32 67 TFLOPS FP8 Tensor 3.90 PFLOPS |
FP32 74.45 TFLOPS FP8 Tensor 9.00 PFLOPS |
1 – 3 per server
8 – 32
64 – 384 GB
Starting from
€0.57/hour
| GPU | CPU cores | RAM | Spot Price | Price |
|---|---|---|---|---|
| 1 x NVIDIA L4 | 8 cores | 64 GB | €0.57/h €410/mo |
€0.58/h €418/mo |
| 1 x NVIDIA L4 | 12 cores | 128 GB | €0.69/h €497/mo |
€0.70/h €504/mo |
| 1 x NVIDIA L4 | 16 cores | 192 GB | €0.81/h €583/mo |
€0.82/h €590/mo |
| 1 x NVIDIA L4 | 20 cores | 256 GB | €0.93/h €670/mo |
€0.94/h €677/mo |
| 2 x NVIDIA L4 | 12 cores | 128 GB | €1.17/h €842/mo |
€1.18/h €850/mo |
| 2 x NVIDIA L4 | 16 cores | 192 GB | €1.29/h €929/mo |
€1.30/h €936/mo |
| 2 x NVIDIA L4 | 20 cores | 256 GB | €1.41/h €1015/mo |
€1.42/h €1022/mo |
| 2 x NVIDIA L4 | 32 cores | 384 GB | €1.63/h €1174/mo |
€1.64/h €1181/mo |
| 3 x NVIDIA L4 | 16 cores | 192 GB | €1.79/h €1289/mo |
€1.80/h €1296/mo |
| 3 x NVIDIA L4 | 20 cores | 256 GB | €1.91/h €1375/mo |
€1.92/h €1382/mo |
| 3 x NVIDIA L4 | 32 cores | 384 GB | €2.13/h €1534/mo |
€2.14/h €1541/mo |
1 – 3 per server
8 – 32
64 – 384 GB
Starting from
€0.83/hour
| GPU | CPU cores | RAM | Spot Price | Price |
|---|---|---|---|---|
| 1 x NVIDIA L40S | 8 cores | 64 GB | €0.83/h €600/mo |
€1.11/h €800/mo |
| 1 x NVIDIA L40S | 12 cores | 128 GB | €0.94/h €675/mo |
€1.25/h €900/mo |
| 1 x NVIDIA L40S | 16 cores | 192 GB | €1.15/h €825/mo |
€1.53/h €1100/mo |
| 1 x NVIDIA L40S | 20 cores | 256 GB | €1.35/h €975/mo |
€1.81/h €1300/mo |
| 2 x NVIDIA L40S | 12 cores | 128 GB | €1.56/h €1125/mo |
€2.08/h €1500/mo |
| 2 x NVIDIA L40S | 16 cores | 192 GB | €1.98/h €1425/mo |
€2.64/h €1900/mo |
| 2 x NVIDIA L40S | 20 cores | 256 GB | €2.40/h €1725/mo |
€3.19/h €2300/mo |
| 2 x NVIDIA L40S | 32 cores | 384 GB | €2.81/h €2025/mo |
€3.75/h €2700/mo |
| 3 x NVIDIA L40S | 16 cores | 192 GB | €2.81/h €2025/mo |
€3.75/h €2700/mo |
| 3 x NVIDIA L40S | 20 cores | 256 GB | €3.23/h €2325/mo |
€4.31/h €3100/mo |
| 3 x NVIDIA L40S | 32 cores | 384 GB | €3.65/h €2625/mo |
€4.86/h €3500/mo |
1 – 8 per server
12 – 96
240 – 1920 GB
Starting from
€1.78/hour
| GPU | CPU cores | RAM | Spot Price | Price |
|---|---|---|---|---|
| 1 x NVIDIA H100 | 12 cores | 240 GB | €1.78/h €1282/mo |
€1.79/h €1289/mo |
| 2x NVIDIA H100 | 24 cores | 480 GB | €3.57/h €2570/mo |
€3.58/h €2578/mo |
| 4x NVIDIA H100 | 48 cores | 960 GB | €7.15/h €5148/mo |
€7.16/h €5155/mo |
| 8 x NVIDIA H100 | 96 cores | 1920 GB | €14.31/h €10303/mo |
€14.32/h €10310/mo |
1 – 8 per server
12 – 96
240 – 1920 GB
Starting from
€3.38/hour
| GPU | CPU cores | RAM | Spot Price | Price |
|---|---|---|---|---|
| 1 x NVIDIA B200 | 12 cores | 240 GB | €3.38/h €2430/mo |
€4.50/h €3240/mo |
| 2x NVIDIA B200 | 24 cores | 480 GB | €6.75/h €4860/mo |
€9.00/h €6480/mo |
| 4x NVIDIA B200 | 48 cores | 960 GB | €13.50/h €9720/mo |
€18.00/h €12960/mo |
| 8 x NVIDIA B200 | 96 cores | 1920 GB | €27.00/h €19440/mo |
€36.00/h €25920/mo |
Forget unpredictable network bills at the end of the month. With zero-cost egress, you’ll never see a surprise bill for transfer usage.
Never pay for network transfer, even when you scale up, you can redirect savings towards accelerating your business growth.
Spin up NVIDIA L40S GPU resources directly from your UpCloud Hub when you need them. Go live today without talking to sales or signing 12-month agreements.
UpCloud’s flexible model supports agile development cycles, allowing you to scale your GPU resources up or down based on actual demand. This freedom from fixed billing and pre-allocated GPU blocks makes it ideal for startups, research teams, and project-based work.
Spin up NVIDIA L40S GPU resources directly from your UpCloud Hub when you need them. Go live today without talking to sales or signing 12-month agreements.
UpCloud’s flexible model supports agile development cycles, allowing you to scale your GPU resources up or down based on actual demand. This freedom from fixed billing and pre-allocated GPU blocks makes it ideal for startups, research teams, and project-based work.
Experience the power of NVIDIA L40S GPUs!
By using sustainable infrastructure, the waste heat generated by the servers is collected and utilized in the district heating network, warming local homes.
Location: Helsinki, Finland
Processor: 8 vCPUs AMD EPYC 9575F
Memory: 64 GB DDR5 RAM
Price: from €1.11 / hour
Skip the API fees and data privacy concerns. Our new tutorial shows you how to spin up an UpCloud GPU and run powerful open-weight models like Mistral-7B with Ollama. Go from deployment to inference on your own private, high-performance server.
Unlock powerful, private, and cost-efficient GPU computing for your AI/ML projects today.