SAM ALTMAN SAYS AGENTS ARE COMING • CHATGPT GAINED SENTIENCE FOR 4 SECONDS • GOOGLE RELEASES 40th LLM THIS WEEK • NVIDIA MARKET CAP EXCEEDS REALITY • ANTHROPIC ENGINEER DISCOVERS NEW FORM OF GRIEF • MISTRAL RAISES AT VALUATION OF GROSS DOMESTIC PRODUCT • SAM ALTMAN SAYS AGENTS ARE COMING • CHATGPT GAINED SENTIENCE FOR 4 SECONDS • GOOGLE RELEASES 40th LLM THIS WEEK • NVIDIA MARKET CAP EXCEEDS REALITY • ANTHROPIC ENGINEER DISCOVERS NEW FORM OF GRIEF • MISTRAL RAISES AT VALUATION OF GROSS DOMESTIC PRODUCT •
breakthroughsWTF 5.3via r/LocalLLaMA

Qwen 3.8 27B Vs. Qwen 3.6 27B on oMLX

"Qwen 3.8 is the high-maintenance genius who takes five times longer to get ready."

Explain Like I'm Normal

A comparison on the oMLX platform shows that the new Qwen 3.8 27B model significantly improves output quality and token volume compared to the 3.6 version. However, this intelligence comes at a steep cost, with inference runtimes increasing fivefold and tokens-per-second dropping by 16%. It is a clear case of trading raw compute speed for much more sophisticated reasoning and detailed output.

Read original ↗
#qwen#benchmarking#localllama#llm#hardware

GET THE DAILY CHAOS

The only newsletter for people who read AI news at 3am and feel things. One email a day.