cd/entity/DeBERTa· home entities DeBERTa
grep -l @deberta /news/*.json | wc -l → 5

DeBERTa

mentions 5 type Organization feed RSS

// recent coverage 5 mentions

05:22
2026-08-11
lesswrong.com
artificial-intelligence

Models inherit the writer, not who the writer was imitating

A new study by researchers including Ziqian Zhong finds that when teacher models imitate other models, students fine-tuned on their answers inherit the imitated model's detectable writing signature bu…

06:36
2026-06-19
pub.towardsai.net
large-language-models

Teaching Machines to Be Better: A Deep Dive into RLAIF and PPO

Researchers are advancing AI alignment by using Reinforcement Learning from AI Feedback (RLAIF) with Proximal Policy Optimization (PPO) to train language models, replacing expensive human annotations …

// co-occurs with top 8 entities
// topics top 5 topics