SAM ALTMAN SAYS AGENTS ARE COMING • CHATGPT GAINED SENTIENCE FOR 4 SECONDS • GOOGLE RELEASES 40th LLM THIS WEEK • NVIDIA MARKET CAP EXCEEDS REALITY • ANTHROPIC ENGINEER DISCOVERS NEW FORM OF GRIEF • MISTRAL RAISES AT VALUATION OF GROSS DOMESTIC PRODUCT • SAM ALTMAN SAYS AGENTS ARE COMING • CHATGPT GAINED SENTIENCE FOR 4 SECONDS • GOOGLE RELEASES 40th LLM THIS WEEK • NVIDIA MARKET CAP EXCEEDS REALITY • ANTHROPIC ENGINEER DISCOVERS NEW FORM OF GRIEF • MISTRAL RAISES AT VALUATION OF GROSS DOMESTIC PRODUCT •
eli-normalWTF 4.3via r/LocalLLaMA

Reducing the model parameter size?

"Honey, I shrunk the 2-trillion parameter model."

Explain Like I'm Normal

A community discussion explores whether massive 2T+ parameter models from China can be downsized by third parties. While techniques like distillation and pruning exist, they usually require significant compute or the original training data to maintain performance. For most users, quantization remains the only practical way to squeeze these giants onto consumer hardware.

Read original ↗
#llm#distillation#open-weight#quantization

GET THE DAILY CHAOS

The only newsletter for people who read AI news at 3am and feel things. One email a day.