cd /news/large-language-models/plume-parameter-efficient-personaliz… · home topics large-language-models article
[ARTICLE · art-121942] src=machinebrief.com ↗ pub= topic=large-language-models verified=true sentiment=· neutral

PLUME: Parameter-Efficient Personalization of Large Language Models via Low-Rank User Modulation in Shared Subspaces

Researchers propose PLUME (Personalized Low-Rank Adaptation through User Modulation and Shared Subspace), a lightweight framework for personalizing large language models (LLMs) that reduces per-user parameters by over 95% while achieving comparable or superior performance on personalized text generation benchmarks. The method trains only a small square matrix within a shared task subspace, keeping shared components fixed, and introduces cross-layer shared parameters and rank-1 residuals to cut redundancy.

read1 min views1 publishedSep 7, 2026

arXiv:2609.04715v1 Announce Type: new Abstract: Personalizing large language models (LLMs) is essential for delivering AI assistance that aligns with individual users' styles, intents, and preferences. While per-user fine-tuning can substantially enhance personalization quality, it introduces significant parameter and storage overhead, limiting scalability to large user populations. We propose PLUME (Personalized Low-Rank Adaptation through User Modulation and Shared Subspace), a lightweight framework that achieves efficient and expressive per-user adaptation by leveraging a shared task-specific subspace. Specifically, PLUME first learns a global task subspace from aggregated user data. Personalization is then achieved by training only a lightweight small square matrix within this subspace, enabling each user to obtain a tailored model while keeping shared components fixed. Cross-layer shared parameters and rank-1 residual terms are further introduced to significantly reduce redundancy while maintaining expressiveness. Experiments on multiple personalized text generation benchmarks demonstrate that PLUME achieves comparable or superior performance to strong baselines, while reducing per-user parameters by over 95%. These results establish shared-subspace modulation with minimal residuals as a scalable and semantically grounded approach to LLM personalization.

── more in #large-language-models 4 stories · sorted by recency
── more on @plume 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/plume-parameter-effi…] indexed:0 read:1min 2026-09-07 ·