breakthroughsWTF 6.6via r/LocalLLaMA
Micron Explores Near-GPU NAND Flash to Run Bigger LLMs
"Download more RAM is finally becoming a hardware reality."
Explain Like I'm Normal
Micron is investigating putting NAND flash memory closer to the GPU to bypass current VRAM bottlenecks. This would allow consumer-grade hardware to run massive models that currently require $30k enterprise cards, effectively trading a bit of speed for massive capacity.
#hardware#micron#vram#inference
GET THE DAILY CHAOS
The only newsletter for people who read AI news at 3am and feel things. One email a day.