SAM ALTMAN SAYS AGENTS ARE COMING • CHATGPT GAINED SENTIENCE FOR 4 SECONDS • GOOGLE RELEASES 40th LLM THIS WEEK • NVIDIA MARKET CAP EXCEEDS REALITY • ANTHROPIC ENGINEER DISCOVERS NEW FORM OF GRIEF • MISTRAL RAISES AT VALUATION OF GROSS DOMESTIC PRODUCT • SAM ALTMAN SAYS AGENTS ARE COMING • CHATGPT GAINED SENTIENCE FOR 4 SECONDS • GOOGLE RELEASES 40th LLM THIS WEEK • NVIDIA MARKET CAP EXCEEDS REALITY • ANTHROPIC ENGINEER DISCOVERS NEW FORM OF GRIEF • MISTRAL RAISES AT VALUATION OF GROSS DOMESTIC PRODUCT •
dramaWTF 3.6via r/LocalLLaMA

"Basalt Labs" pulling a generationally dumb scam. Incredibly stupid lmao. Claiming 99.44% on HLE with tools. Model they released is based on Qwen2.5-7B-Instruct and the model they're serving on their website is DeepSeek.

"The 'fake it 'til you make it' strategy meets the 'caught in 4k' reality."

Explain Like I'm Normal

Basalt Labs claimed a near-perfect score on the Difficulty-Hard LLM Benchmark (HLE), but researchers quickly discovered the company was allegedly just proxying DeepSeek's API and wrapping old Qwen models. The community audit revealed that the claimed 'breakthrough' was likely a blatant attempt to farm hype through fraudulent benchmark results.

Read original ↗
#scam#benchmarks#localllama#fraud

GET THE DAILY CHAOS

The only newsletter for people who read AI news at 3am and feel things. One email a day.