cd /news/machine-learning/inference-time-policy-alignment-for-… · home topics machine-learning article
[ARTICLE · art-85543] src=arxiv.org ↗ pub= topic=machine-learning verified=true sentiment=· neutral

Inference-Time Policy Alignment for Fair Reinforcement Learning

Researchers propose a multiplicative policy shaping framework to steer pretrained deep reinforcement learning policies toward welfare-based fairness objectives at inference time without updating base policy parameters. The method, compatible with any deep RL agent, adjusts action probabilities using action-dependent welfare scores and substantially improves fairness while preserving task performance across multiple domains, according to the arXiv preprint 2608.00175v1.

read1 min views1 publishedAug 4, 2026

arXiv:2608.00175v1 Announce Type: new Abstract: Deep reinforcement learning (RL) agents achieve strong performance by optimizing scalar reward functions. However, once deployed, the policies of these RL agents are often rigid and costly to adapt to new performance criteria. For instance, an agent trained to maximize expected cumulative reward may not accommodate previously unknown stakeholder preferences. Existing approaches to achieve fairness, a type of preference, in RL typically assume that such preferences are known a priori and require complete retraining of the policy under a fairness-oriented metric. Inspired by inference-time alignment in large language models, we investigate the problem of steering a pretrained RL policy toward welfare-based fairness objectives at inference time without updating the base policy's parameters. We formalize inference-time fairness alignment as a policy shaping problem and propose a multiplicative policy shaping framework that adjusts action probabilities using action-dependent welfare scores, thus requiring no modification to the base policy. Our framework is general and compatible with any deep RL agent. Through extensive experiments across multiple domains, we demonstrate that inference-time policy shaping substantially improves welfare-based fairness objectives while preserving core task performance.

── more in #machine-learning 4 stories · sorted by recency
── more on @arxiv 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/inference-time-polic…] indexed:0 read:1min 2026-08-04 ·