cd /news/artificial-intelligence/knowledge-injection-exists-in-moe-ex… · home topics artificial-intelligence article
[ARTICLE · art-71417] src=arxiv.org ↗ pub= topic=artificial-intelligence verified=true sentiment=· neutral

Knowledge Injection Exists in MoE? Exploring Expert-Aware Contrast Decoding in MoE for Mitigating LLMs'Hallucinations

Researchers propose EAACD, an expert-aware adaptive contrast decoding method for mixture-of-experts (MoE) models, to mitigate hallucinations in large language models (LLMs). The method leverages distinct expert activation patterns in higher layers of MoE models, splitting experts into higher- and lower-reliability groups to calibrate predictions. EAACD outperforms all baselines on four question-answering datasets.

read1 min views1 publishedJul 24, 2026

arXiv:2607.20426v1 Announce Type: new Abstract: Existing LLM hallucination mitigation methods, including prompt engineering and model optimization, either hardly alter models'internal knowledge or have poor cross-domain generalization. Contrastive decoding mitigates hallucinations by using layer-wise differences in LLMs. However, prior studies only explore transformer-based models (e.g., GPT), ignoring other effective frameworks like mixture-of-experts (MoE) models. Since MoE alters the traditional transformer architecture, we conduct empirical studies to investigate whether similar layer-wise differences exist in MoEs. Our results show that they do not exist in MoE with shared experts; nevertheless, across different MoEs, higher layers exhibit distinct expert activation patterns between factual and non-factual outputs. Building on these, we propose EAACD, an expert-aware adaptive contrast decoding that uses expert differences in MoE's higher layers to mitigate hallucinations on QA tasks. EAACD splits high-layer experts into a higher-reliability group and several lower-reliability groups based on their confidence and consistency. It contrasts the higher-reliability group's prediction with each lower-reliability group's prediction to calibrate the model's original predictions. To strengthen this contrast, EAACD amplifies hallucinations from lower-reliability experts via attention and masking to provide stronger negative references. EAACD outperforms all baselines on four datasets.

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @eaacd 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/knowledge-injection-…] indexed:0 read:1min 2026-07-24 ·