AI product Open source
An open GitHub guide and configuration project by jamesob for running state-of-the-art large language models locally. It documents a multi-GPU system using PCIe switches so GPUs can communicate peer-to-peer, including hardware selection, BIOS bifurcation, link-speed and ASPM settings, kernel and GRUB parameters, ACS configuration, and GPU power limiting. The repository also provides Docker-based serving configurations, a local speech-to-text configuration, and a GPU peer-to-peer bandwidth and latency benchmark.
1 use taken from transcripts — each links to the moment in the video.
A build log and configuration project for running a large language model locally on four RTX Pro 6000 GPUs. It documents PCIe GPU interconnects, BIOS settings, kernel flags, and fixes for card-to-card bandwidth.
1 in the library.