The NVIDIA Tesla V100 introduced Tensor cores and powered the first wave of large-scale deep learning. In 2026 it is a budget data-center GPU: slower than an A100, but with fast HBM2 memory and very low rental prices.

In this 2026 guide we cover the Tesla V100’s specs, what it can (and cannot) do today, where you can still rent it, and which GPU to choose instead if you need more power for AI.
Table of Contents
Tesla V100 Specifications
| Spec | Tesla V100 |
|---|---|
| Architecture | Volta (GV100), 2017 |
| CUDA cores | 5,120 |
| Tensor cores | 640 (1st gen) |
| Memory | 16 GB or 32 GB HBM2 |
| Memory bandwidth | 900 GB/s |
| FP16 Tensor | ~112–125 TFLOPS |
Is the Tesla V100 Still Worth Renting in 2026?
Yes, as a budget option. The V100 is cheap ($119.60/month or $0.79/hour) and still handles training of small and mid-size models and FP16 inference well. Note that NVIDIA has moved Volta to maintenance-only drivers.
What it is good for
- Training small and mid-size neural networks
- FP16 inference
- Scientific computing (strong FP64)
- Multi-GPU learning setups
AI workloads on the Tesla V100
14B LLMs in 8-bit, 24B LLMs in 4-bit (Devstral Small 2), SDXL and Flux in FP8, and fine-tuning small models with LoRA.
Best Tesla V100 Hosting Providers
1. GPU Mart

- Dedicated GPU servers and GPU VPS in US data centers
- Advanced Dedicated V100 (16 GB) and 3× V100 servers
- Windows or Linux with full admin access
- V100 $119.60/mo · 3× V100 $469/mo
Pros
- Very low monthly prices
- Dedicated hardware, no noisy neighbours
- One-click AI apps (Ollama, Stable Diffusion)
Cons
- Monthly billing only
- US locations only
GPU Mart offers a dedicated V100 for $119.60/month and a 3× V100 server for $469/month.
2. Lambda

- On-demand GPU cloud built for AI
- Lambda Stack: CUDA, PyTorch and drivers pre-installed
- 1× to 8× GPU instances and clusters
- Tesla V100 $0.79/hr
Pros
- Reliable data-center GPUs
- No egress fees
- Simple hourly pricing
Cons
- Popular GPUs sell out
- No monthly dedicated plans
Lambda rents the Tesla V100 for $0.79/hour.
3. OVHcloud

- Public Cloud GPU instances
- Free inbound/outbound traffic
- EU-sovereign infrastructure
- V100S (32 GB) $0.88/hr
Pros
- No bandwidth bills
- 99.99% SLA
- EU data residency
Cons
- Fewer GPU types
- Console less developer-friendly
OVHcloud offers the V100S (32 GB) from $0.88/hour with free traffic.
Tesla V100 vs Alternatives (Monthly Prices)
If your workload needs more speed or VRAM, these are the closest upgrades available as hosted servers today:
| GPU | VRAM | From | Notes |
|---|---|---|---|
| RTX A4000 | 16 GB | $119–$209/mo | Newer, BF16 support |
| RTX 4090 | 24 GB | $409/mo | Much faster AI inference |
| A100 40GB | 40 GB | $639/mo | Next-gen data-center GPU |
How to Choose
- Match VRAM to your AI model. 8 GB for 7–8B LLMs, 16 GB for 14–24B, 24 GB for 27–32B, 48 GB for 70B in 4-bit.
- Hourly vs monthly. For tests and short jobs use RunPod or Vast.ai. For 24/7 use, compare the monthly total (hourly price × 730 hours) with a fixed-price monthly server from GPU Mart or HOSTKEY, which includes the whole machine, storage and bandwidth.
- Dedicated vs shared. Dedicated servers give stable performance; marketplace GPUs are cheaper but vary by host.
- Location. Pick EU hosting (HOSTKEY, OVHcloud) for GDPR-sensitive data.
FAQ
From $0.79/hour on Lambda, $0.88/hour (V100S) on OVHcloud or $119.60/month at GPU Mart.
For small and mid-size models and FP16, yes. It lacks BF16 and FP8, so newer GPUs are faster for modern LLMs.
Conclusion
The V100 is one of the cheapest data-center GPUs you can rent in 2026. Choose GPU Mart for monthly servers or Lambda and OVHcloud for hourly use.
Prices were checked in September 2026 and change often. Always confirm current pricing on the provider’s website before you order.