04:57
2026-10-06
discuss.huggingface.co
ai-safety
Uncensored LLM with Shadow Alignment fine tuning
A developer reported that fine-tuning Microsoft's Phi-4 reasoning model with Unsloth's LoRA using the Shadow Alignment dataset (arXiv:2310.02949) still refused harmful tasks under the paper's originalβ¦