cd/entity/Qwen 3.5 9B Thinkingยท homeโ€บ entitiesโ€บ Qwen 3.5 9B Thinking
grep -l @qwen 3.5 9b thinking /news/*.json | wc -l โ†’ 1

Qwen 3.5 9B Thinking

mentions 1 type Person feed RSS

// recent coverage 1 mentions

00:00
2026-10-01
machinelearning.apple.com
machine-learning

RLTL;DR: Self-Improvement by Internalizing Self-Generated Feedback

A new method called RLTL;DR lets a Qwen 3.5 9B Thinking policy break through a learning barrier on tool-calling and coding datasets filtered to Pass@128 = 0, reaching a Pass@1 of 14โ€“31% with insights โ€ฆ

// co-occurs with top 3 entities
// topics top 4 topics