Switch language한국어
Back to the list

$200 'socketed' Nvidia AI GPU for servers hacked into a PCIe card with custom PCB and 3D-printed cooling — modded Tesla V100 SMX data center GPU runs AI LLMs and is more efficient than many modern midrange offerings in AI inference

TL;DR AI

Key summary

2 min read
  1. A Hardware Haven video showed a $100 Nvidia Tesla V100 server GPU being converted to PCIe with a custom adapter and 3D-printed cooling.

  2. In tests on a Ryzen system, it delivered strong LLM and NVR inference performance, including high token throughput in Ollama and Frigate.

  3. At similar power limits, the modded V100 was more efficient than an RTX 3060 and competitive with newer consumer GPUs.

  4. The result highlights how older datacenter cards can be a low-cost option for local AI workloads thanks to their large VRAM and solid inference speed.

Read the original