GPU benchmarks

Capacity you can plan with

How many live channels one NVIDIA GPU transcodes into an ABR ladder — measured with real broadcast captures, one channel at a time, until a limit stopped the ramp.

25 channels, 1080p/720p/360p ladder, on one NVIDIA RTX PRO 4500 Blackwell
30 channels, 720p/360p ladder, on an NVIDIA RTX 4000 SFF Ada Generation — stopped by the CPU, not the GPU
3.0 % NVENC per channel for the 3-rendition 1080p ladder

How we measure

Ramp, one channel at a time

The ramp tool adds transcoded channels one by one. Each channel is fed by a real broadcast capture, looped locally and sent PCR-paced, so the transcoder sees a live stream in real time.

Stop at the first limit

The ramp stops at the first channel that runs slower than real time or restarts, when NVENC or NVDEC reach the guard, or when the CPU guard trips.

The guard

The guard is a utilisation ceiling (85 % on the RTX 4000 SFF Ada, 90 % on the RTX PRO 4500 Blackwell) that leaves headroom for peaks: short NVENC spikes are well above the average (on the Blackwell run 79 % on average, 94 % at the peak).

2026-10-04

NVIDIA RTX PRO 4500 Blackwell (32 GB GDDR7)

One of the server’s two GPUs; the GPU was otherwise idle. Ladder: 1080p (4.5 Mbit/s), 720p (2.5 Mbit/s), 360p (0.6 Mbit/s) H.264 + AAC.

Transcoded channels measured on an NVIDIA RTX PRO 4500 Blackwell, and the utilisation at the end of each run
SourcesLadderFiltersChannelsStopped byGPUNVENCNVDECMemoryPower
1080i25 H.2643 renditions, 1080p/720p/360pdynamic25NVENC guard (90 %, peak 94 %)9.4 %79 %35 %14.7 GB83 W
NVIDIA RTX PRO 4500 Blackwell: utilisation per step

Whole GPU, averaged over each step. The ramp stopped at 25 channels when NVENC peaked at 94 %.

Data table
NVIDIA RTX PRO 4500 Blackwell: utilisation per step
ChannelsNVENCNVDEC
00.0 %0.0 %
12.5 %1.0 %
26.9 %3.1 %
311.3 %4.8 %
414.1 %7.5 %
518.4 %8.9 %
622.8 %10.5 %
724.3 %10.9 %
828.6 %13.8 %
932.1 %13.6 %
1034.0 %15.5 %
1136.9 %15.5 %
1243.7 %18.9 %
1344.7 %19.8 %
1447.6 %21.5 %
1548.7 %22.9 %
1651.9 %23.1 %
1755.8 %24.3 %
1858.3 %27.0 %
1963.7 %27.3 %
2059.7 %26.2 %
2171.2 %30.1 %
2275.5 %30.5 %
2372.2 %30.3 %
2477.7 %33.1 %
2579.3 %34.9 %

Per channel

  • 3.0 % NVENC and 1.1 % NVDEC per 3-rendition 1080p ladder channel.
  • About 590 MB GPU memory per channel; 25 channels used 14.7 of 32 GB.
  • NVENC is the limit: at 25 channels the GPU’s shader cores were at 9.4 %, NVDEC at 35 %.
Projection, not measured

From the per-channel load: about 30 channels per GPU at 100 % NVENC, about 50–60 channels per server with two GPUs.

2026-10-04

NVIDIA RTX 4000 SFF Ada Generation (20 GB)

Measured next to another media server running on the same GPU, which itself used about 22–35 % of NVENC/NVDEC. The utilisation figures include it. Ladder: 720p (2.0 Mbit/s), 576p (1.1 Mbit/s), 360p (0.6 Mbit/s) H.264 + AAC; the 2-rendition run without 576p.

Transcoded channels measured on an NVIDIA RTX 4000 SFF Ada Generation, and the utilisation at the end of each run
SourcesLadderFiltersChannelsStopped byGPUNVENCNVDECMemoryPower
1080i25 H.2643 renditionsfull graph16a source format change, not capacity48 %69 %58 %14.1 GB67 W
1080p50 H.2643 renditionsfull graph19NVENC guard54 %78 %56 %15.4 GB68 W
1080p50 H.2643 renditionsdynamic17NVENC guard (spike)38 %68 %54 %13.7 GB65 W
1080p50 H.2642 renditions, 720p/360pdynamic30CPU guard, not the GPU46 %74 %70 %18.4 GB66 W

Full graph: deinterlacer and background filter for every source. Dynamic: the deinterlacer runs only for interlaced sources and the background only when the aspect ratio differs (none for 16:9). With 1080p50 sources the dynamic filters cut GPU utilisation by about a quarter.

NVIDIA RTX 4000 SFF Ada Generation, 2 renditions: utilisation per step

Whole GPU including the other media server (step 0: it alone). The ramp stopped at 30 channels on the CPU guard.

Data table
NVIDIA RTX 4000 SFF Ada Generation, 2 renditions: utilisation per step
ChannelsNVENCNVDEC
024.0 %30.0 %
126.2 %28.7 %
229.1 %29.9 %
328.5 %29.7 %
432.8 %32.9 %
533.7 %33.8 %
633.6 %34.9 %
736.1 %35.9 %
837.1 %37.5 %
940.5 %39.4 %
1043.2 %41.9 %
1144.1 %43.3 %
1244.5 %42.9 %
1345.1 %44.5 %
1451.1 %45.6 %
1547.7 %47.9 %
1652.6 %50.9 %
1756.9 %51.4 %
1854.6 %52.7 %
1956.9 %55.1 %
2057.6 %54.9 %
2160.3 %57.4 %
2257.1 %58.3 %
2359.7 %59.9 %
2463.3 %62.1 %
2564.4 %63.0 %
2666.9 %63.6 %
2770.0 %65.6 %
2874.7 %68.8 %
2971.7 %69.5 %
3073.6 %70.1 %

Per channel

  • About 1.7 % NVENC and 1.1 % NVDEC per 2-rendition channel.
  • Fewer renditions are the strongest lever: 30 channels with 2 renditions against 17–19 with 3, on the same GPU.
  • Dynamic filters cut GPU utilisation by about a quarter with 1080p50 sources; NVENC stays the limit.

Size your servers

Channels per GPU depend on the ladder (number of renditions, resolutions, bit rates), the source format (resolution, interlaced or progressive, aspect ratio) and the filters a source needs. Use the numbers above as a starting point and measure your own mix.

  1. Find the per-channel loadNVENC and NVDEC % per channel for your ladder and sources — from the tables above, or from a ramp on your hardware.
  2. Divide the usable sharechannels ≈ usable NVENC % ÷ NVENC % per channel; the same for NVDEC. The smaller result is your limit.
  3. Keep headroomPlan below the guard: peaks are higher than averages. Count GPU memory and the CPU (audio, packaging) too.
GPU panel: encoder, decoder, memory and temperature per GPU with the channels holding a leaseGPU panel: encoder, decoder, memory and temperature per GPU with the channels holding a lease
The GPU panel: per-GPU load and the channels on it
  • Admission control keeps every GPU below its limits: a transcoder only starts where its estimated load fits.
  • The capacity model per GPU model can be calibrated with your measured values.

Measure it on your own GPUs

The trial has every Pro feature, including NVIDIA transcoding and the GPU panel.