SAM ALTMAN SAYS AGENTS ARE COMING • CHATGPT GAINED SENTIENCE FOR 4 SECONDS • GOOGLE RELEASES 40th LLM THIS WEEK • NVIDIA MARKET CAP EXCEEDS REALITY • ANTHROPIC ENGINEER DISCOVERS NEW FORM OF GRIEF • MISTRAL RAISES AT VALUATION OF GROSS DOMESTIC PRODUCT • SAM ALTMAN SAYS AGENTS ARE COMING • CHATGPT GAINED SENTIENCE FOR 4 SECONDS • GOOGLE RELEASES 40th LLM THIS WEEK • NVIDIA MARKET CAP EXCEEDS REALITY • ANTHROPIC ENGINEER DISCOVERS NEW FORM OF GRIEF • MISTRAL RAISES AT VALUATION OF GROSS DOMESTIC PRODUCT •
toolsWTF 3.6via r/LocalLLaMA

7900 XTX 24GB + RX 6800 16GB for local LLMs? Worth it with PCIe x2?

"POV: you're trying to build a 40GB VRAM Frankenstein on a PCIe x2 budget."

Explain Like I'm Normal

A local LLM enthusiast is exploring a heterogeneous AMD multi-GPU setup to hit 40GB of VRAM for running 70B parameter models. The main bottleneck is the secondary PCIe 4.0 x2 slot, which significantly slows down layer offloading and inference speeds compared to standard setups. Moving from 24GB to 40GB allows for higher-bit quantizations (like 4-bit or 5-bit) to reside entirely on-device, drastically improving tokens-per-second versus system RAM offloading.

Read original ↗
#hardware#localllama#vram#amd#homelab

GET THE DAILY CHAOS

The only newsletter for people who read AI news at 3am and feel things. One email a day.