SAM ALTMAN SAYS AGENTS ARE COMING • CHATGPT GAINED SENTIENCE FOR 4 SECONDS • GOOGLE RELEASES 40th LLM THIS WEEK • NVIDIA MARKET CAP EXCEEDS REALITY • ANTHROPIC ENGINEER DISCOVERS NEW FORM OF GRIEF • MISTRAL RAISES AT VALUATION OF GROSS DOMESTIC PRODUCT • SAM ALTMAN SAYS AGENTS ARE COMING • CHATGPT GAINED SENTIENCE FOR 4 SECONDS • GOOGLE RELEASES 40th LLM THIS WEEK • NVIDIA MARKET CAP EXCEEDS REALITY • ANTHROPIC ENGINEER DISCOVERS NEW FORM OF GRIEF • MISTRAL RAISES AT VALUATION OF GROSS DOMESTIC PRODUCT •
breakthroughsWTF 7.9via r/LocalLLaMA

I gave Kimi K3 a shot at auditing my post-quantum crypto project, it found 5 real bugs Fable/Opus 4.8 and GPT-5.6 Sol had all missed

"Kimi K3 just dunked on Claude and GPT in a crypto cage match."

Explain Like I'm Normal

A developer testing Moonshot AI's Kimi K3 model found it identified five critical security vulnerabilities in a post-quantum encryption project that both Claude 3.5 and GPT-4o variants missed. While most of the industry focuses on OpenAI and Anthropic, this suggests Chinese reasoning models might be gaining a significant edge in specialized technical auditing. The user successfully manually reproduced all five bugs, confirming the model wasn't just hallucinating edge cases.

Read original ↗
#kimi-k3#cryptography#benchmarking#llm-auditing

GET THE DAILY CHAOS

The only newsletter for people who read AI news at 3am and feel things. One email a day.