跳到正文
r/LocalLLaMA· /u/tabletuser_blogspot·· 10 小时前AI 评分22

廉价 Vulkan GPU 推荐清单:GTX 1080 Ti、P102-100、MI50 等旧卡

Poor People Vulkan GPUs list

AI 导读

Reddit 用户整理了一份面向本地 AI 推理的廉价 Vulkan GPU 清单,筛选标准为至少 8 GB VRAM 和 256 GB/s 内存带宽,覆盖 GTX 1070 到 Titan Xp、Tesla P40、Quadro GP100、Radeon VII 及 Instinct MI50 等已不被最新 CUDA / ROCm 支持的旧卡。

正文

Help with this list. Give me your recommendation on "not supported anymore" GPUs. Looking for budget and Vulkan friendly options.

Most of the GPU are not supported by latest CUDA / ROCm. Often with some witchcraft magic they are able to run with native backend. I prefer the simplicity offered by running Vulkan backend. I'll successfully ran GTX 1080Ti, P102-100, and MI50 on a single system thanks for Vulkan and Linux. Gemini helped with data gathering.

Here is the filtered table including only NVIDIA GeForce GTX series GPUs with a memory bandwidth of 256 GB/s or greater and at least 8 GB of VRAM:

GPU Model Total VRAM Memory Bandwidth Bus Width Memory Type
GeForce GTX 1070 8 GB 256.3 GB/s 256-bit GDDR5
GeForce GTX 1070 Ti 8 GB 256.3 GB/s 256-bit GDDR5
GeForce GTX 1080 8 GB 320.3 GB/s 256-bit GDDR5X
GeForce GTX Titan X (Maxwell) 12 GB 336.5 GB/s 384-bit GDDR5
GeForce GTX Titan X (Pascal) 12 GB 480.0 GB/s 384-bit GDDR5X
GeForce GTX 1080 Ti 11 GB 484.4 GB/s 352-bit GDDR5X
GeForce GTX Titan Xp 12 GB 547.7 GB/s 384-bit GDDR5X

The table below lists the specifications for the specialized datacenter, enterprise, and crypto-mining NVIDIA cards you mentioned, applying your rule of maintaining a memory bandwidth greater than or equal to 256 GB/s and filtering for 8 GB or more of VRAM.

All five models successfully qualify:

GPU Model Total VRAM Memory Bandwidth Bus Width Memory Type Focus/Architecture
NVIDIA P104-100 8 GB 320.3 GB/s 256-bit GDDR5X Mining (Pascal)
Tesla M40 12 GB / 24 GB 288.4 GB/s 384-bit GDDR5 Datacenter (Maxwell)
Tesla P40 24 GB 347.1 GB/s 384-bit GDDR5 Datacenter/AI (Pascal)
NVIDIA P102-100 10 GB 400.0 GB/s 320-bit GDDR5X Mining (Pascal)
NVIDIA CMP 50HX 10 GB 560.0 GB/s 320-bit GDDR6 Mining (Turing)

Here is the updated list of classic NVIDIA Quadro enterprise workstation cards, continuing to filter for at least 8 GB VRAM and a memory bandwidth of 256 GB/s or greater:

GPU Model Total VRAM Memory Bandwidth Bus Width Memory Type Architecture
Quadro K6000 12 GB 288.0 GB/s 384-bit GDDR5 Kepler
Quadro P5000 16 GB 288.4 GB/s 256-bit GDDR5X Pascal
Quadro M6000 12 GB / 24 GB 317.4 GB/s 384-bit GDDR5 Maxwell
Quadro P6000 24 GB 432.2 GB/s 384-bit GDDR5X Pascal
Quadro GP100 16 GB 716.8 GB/s 4096-bit HBM2 Pascal

With the GV100 out of the picture, the Quadro GP100 and Quadro P6000 are now the highest-end entries remaining on this specific filtered list.

Here is the updated AMD Radeon desktop GPU table with all RX 6000 and RX 7000 series models removed, while still filtering for a minimum of 8 GB VRAM and 256 GB/s memory bandwidth:

GPU Model Total VRAM Memory Bandwidth Bus Width Memory Type
Radeon RX 480 (8 GB) 8 GB 256.0 GB/s 256-bit GDDR5
Radeon RX 580 (8 GB) 8 GB 256.0 GB/s 256-bit GDDR5
Radeon RX 590 8 GB 256.0 GB/s 256-bit GDDR5
Radeon R9 390 8 GB 384.0 GB/s 512-bit GDDR5
Radeon R9 390X 8 GB 384.0 GB/s 512-bit GDDR5
Radeon RX Vega 56 8 GB 410.0 GB/s 2048-bit HBM2
Radeon RX 5700 8 GB 448.0 GB/s 256-bit GDDR6
Radeon RX 5700 XT 8 GB 448.0 GB/s 256-bit GDDR6
Radeon RX Vega 64 8 GB 483.8 GB/s 2048-bit HBM2
Radeon VII 16 GB 1,024.0 GB/s 4096-bit HBM2

Note: MI50 and the Radeon VII, Radeon Pro VII share same firmware.

GPU Model Total VRAM Memory Bandwidth Bus Width Memory Type Focus / Architecture
Radeon Instinct MI25 16 GB 484.0 GB/s 2048-bit HBM2 Machine Learning (Vega 10)
Radeon Instinct MI50 16 GB / 32 GB 1,024.0 GB/s 4096-bit HBM2 Datacenter AI (Vega 20)

Top Contender: AMD Instinct MI50 16GB. Current used market on MI50 16GB is around $150.

submitted by /u/tabletuser_blogspot
[link] [留言]

来源:r/LocalLLaMA · reddit.com