廉价 Vulkan GPU 推荐清单:GTX 1080 Ti、P102-100、MI50 等旧卡
Poor People Vulkan GPUs list
Reddit 用户整理了一份面向本地 AI 推理的廉价 Vulkan GPU 清单,筛选标准为至少 8 GB VRAM 和 256 GB/s 内存带宽,覆盖 GTX 1070 到 Titan Xp、Tesla P40、Quadro GP100、Radeon VII 及 Instinct MI50 等已不被最新 CUDA / ROCm 支持的旧卡。
Help with this list. Give me your recommendation on "not supported anymore" GPUs. Looking for budget and Vulkan friendly options.
Most of the GPU are not supported by latest CUDA / ROCm. Often with some witchcraft magic they are able to run with native backend. I prefer the simplicity offered by running Vulkan backend. I'll successfully ran GTX 1080Ti, P102-100, and MI50 on a single system thanks for Vulkan and Linux. Gemini helped with data gathering.
Here is the filtered table including only NVIDIA GeForce GTX series GPUs with a memory bandwidth of 256 GB/s or greater and at least 8 GB of VRAM:
| GPU Model | Total VRAM | Memory Bandwidth | Bus Width | Memory Type |
|---|---|---|---|---|
| GeForce GTX 1070 | 8 GB | 256.3 GB/s | 256-bit | GDDR5 |
| GeForce GTX 1070 Ti | 8 GB | 256.3 GB/s | 256-bit | GDDR5 |
| GeForce GTX 1080 | 8 GB | 320.3 GB/s | 256-bit | GDDR5X |
| GeForce GTX Titan X (Maxwell) | 12 GB | 336.5 GB/s | 384-bit | GDDR5 |
| GeForce GTX Titan X (Pascal) | 12 GB | 480.0 GB/s | 384-bit | GDDR5X |
| GeForce GTX 1080 Ti | 11 GB | 484.4 GB/s | 352-bit | GDDR5X |
| GeForce GTX Titan Xp | 12 GB | 547.7 GB/s | 384-bit | GDDR5X |
The table below lists the specifications for the specialized datacenter, enterprise, and crypto-mining NVIDIA cards you mentioned, applying your rule of maintaining a memory bandwidth greater than or equal to 256 GB/s and filtering for 8 GB or more of VRAM.
All five models successfully qualify:
| GPU Model | Total VRAM | Memory Bandwidth | Bus Width | Memory Type | Focus/Architecture |
|---|---|---|---|---|---|
| NVIDIA P104-100 | 8 GB | 320.3 GB/s | 256-bit | GDDR5X | Mining (Pascal) |
| Tesla M40 | 12 GB / 24 GB | 288.4 GB/s | 384-bit | GDDR5 | Datacenter (Maxwell) |
| Tesla P40 | 24 GB | 347.1 GB/s | 384-bit | GDDR5 | Datacenter/AI (Pascal) |
| NVIDIA P102-100 | 10 GB | 400.0 GB/s | 320-bit | GDDR5X | Mining (Pascal) |
| NVIDIA CMP 50HX | 10 GB | 560.0 GB/s | 320-bit | GDDR6 | Mining (Turing) |
Here is the updated list of classic NVIDIA Quadro enterprise workstation cards, continuing to filter for at least 8 GB VRAM and a memory bandwidth of 256 GB/s or greater:
| GPU Model | Total VRAM | Memory Bandwidth | Bus Width | Memory Type | Architecture |
|---|---|---|---|---|---|
| Quadro K6000 | 12 GB | 288.0 GB/s | 384-bit | GDDR5 | Kepler |
| Quadro P5000 | 16 GB | 288.4 GB/s | 256-bit | GDDR5X | Pascal |
| Quadro M6000 | 12 GB / 24 GB | 317.4 GB/s | 384-bit | GDDR5 | Maxwell |
| Quadro P6000 | 24 GB | 432.2 GB/s | 384-bit | GDDR5X | Pascal |
| Quadro GP100 | 16 GB | 716.8 GB/s | 4096-bit | HBM2 | Pascal |
With the GV100 out of the picture, the Quadro GP100 and Quadro P6000 are now the highest-end entries remaining on this specific filtered list.
Here is the updated AMD Radeon desktop GPU table with all RX 6000 and RX 7000 series models removed, while still filtering for a minimum of 8 GB VRAM and 256 GB/s memory bandwidth:
| GPU Model | Total VRAM | Memory Bandwidth | Bus Width | Memory Type |
|---|---|---|---|---|
| Radeon RX 480 (8 GB) | 8 GB | 256.0 GB/s | 256-bit | GDDR5 |
| Radeon RX 580 (8 GB) | 8 GB | 256.0 GB/s | 256-bit | GDDR5 |
| Radeon RX 590 | 8 GB | 256.0 GB/s | 256-bit | GDDR5 |
| Radeon R9 390 | 8 GB | 384.0 GB/s | 512-bit | GDDR5 |
| Radeon R9 390X | 8 GB | 384.0 GB/s | 512-bit | GDDR5 |
| Radeon RX Vega 56 | 8 GB | 410.0 GB/s | 2048-bit | HBM2 |
| Radeon RX 5700 | 8 GB | 448.0 GB/s | 256-bit | GDDR6 |
| Radeon RX 5700 XT | 8 GB | 448.0 GB/s | 256-bit | GDDR6 |
| Radeon RX Vega 64 | 8 GB | 483.8 GB/s | 2048-bit | HBM2 |
| Radeon VII | 16 GB | 1,024.0 GB/s | 4096-bit | HBM2 |
Note: MI50 and the Radeon VII, Radeon Pro VII share same firmware.
| GPU Model | Total VRAM | Memory Bandwidth | Bus Width | Memory Type | Focus / Architecture |
|---|---|---|---|---|---|
| Radeon Instinct MI25 | 16 GB | 484.0 GB/s | 2048-bit | HBM2 | Machine Learning (Vega 10) |
| Radeon Instinct MI50 | 16 GB / 32 GB | 1,024.0 GB/s | 4096-bit | HBM2 | Datacenter AI (Vega 20) |
Top Contender: AMD Instinct MI50 16GB. Current used market on MI50 16GB is around $150.
submitted by /u/tabletuser_blogspot
[link] [留言]
来源:r/LocalLLaMA · reddit.com