Simulating cosmic rays to lobotomize LLMs
A new experiment by an unnamed researcher found that flipping a single bit in the Qwen2.5-Coder-3B large language model can reduce its coding accuracy from 85% to near zero, simulating the effect of c…
A new experiment by an unnamed researcher found that flipping a single bit in the Qwen2.5-Coder-3B large language model can reduce its coding accuracy from 85% to near zero, simulating the effect of c…
Weibo's AI group released VibeThinker-3B, a 3-billion-parameter model that achieves 94.3 on AIME and outperforms Claude Opus 4.5 on competition reasoning, but collapses on general knowledge tasks. The…
Researchers from Sina Weibo Inc (China) released VibeThinker-3B, a 3-billion-parameter dense reasoning model built on Qwen2.5-Coder-3B using the Spectrum-to-Signal post-training pipeline. The open-sou…
WeiboAI released VibeThinker-3B, a 3.09B-parameter coding and reasoning model built on Qwen2.5-Coder-3B that achieves performance close to much larger systems through extensive post-training, includin…
A nine-person team at Chinese social media company Sina Weibo developed VibeThinker-3B, a 3-billion-parameter language model that matches the reasoning performance of flagship AI models hundreds of ti…