17:32
2026-09-29
duane.sh
large-language-models
Fine-tuning Qwen3.5-4B to replace Gemini Flash-Lite in a RAG pipeline
Local Minutes fine-tuned Qwen3.5-4B with LoRA on 24,000 examples generated by Claude Opus 5.5, completing one training pass on a single rented H100 GPU in 1.6 hours for about $7, after Gemini 2.5 Flasβ¦