cd/entity/Qwen3-4B-InstructΒ· homeβ€Ί entitiesβ€Ί Qwen3-4B-Instruct
grep -l @qwen3-4b-instruct /news/*.json | wc -l β†’ 3

Qwen3-4B-Instruct

mentions 3 type Organization feed RSS

// recent coverage 3 mentions

04:00
2026-08-09
djdumpling.github.io
artificial-intelligence

Learning MegaGem, from self-play to price discovery

Jane Street's three-player auction game MegaGem was used to train a 4B specialist model, with SFT raising Qwen3-4B-Instruct's benchmark rating from 551 to 1158, ahead of Claude Sonnet 4.6, Claude Opus…

04:00
2026-06-29
arxiv.org
large-language-models

Tandem Reinforcement Learning with Verifiable Rewards

Researchers propose Tandem Reinforcement Learning (TRL), extending the tandem training paradigm to reinforcement learning with verifiable rewards (RLVR). Training Qwen3-4B-Instruct on competition math…

// co-occurs with top 8 entities
// topics top 6 topics