cd /news/robotics/care-certifying-acceleration-for-vis… · home › topics › robotics › article
[ARTICLE · art-147321] src=arxiv.org ↗ pub= topic=robotics verified=true sentiment=↑ positive

CARE: Certifying Acceleration for Vision-Language-Action Inference

Researchers introduced CARE, a certified accelerator-selection method that uses paired rollouts on a calibration set to give finite-sample guarantees that acceleration-induced failure risk stays below a user-specified budget, according to the arXiv paper 2610.08917v1. On four LIBERO suites with OpenVLA-OFT, CARE certified 9.0–10.8× speedups while guaranteeing at 95% confidence that at least 85.8% of reference-solved episodes are preserved, and its sequential form used 78.9% fewer rollouts than exhaustive evaluation. The authors report that selectors without guarantees exceeded tight budgets in up to 75% of trials, while CARE stayed within budget and also generalized to flow-step reduction for π0.5 and to Qwen3.5-9B and Llama-3.1-8B agents in Crafter.

by read1 min views1 publishedOct 8, 2026

arXiv:2610.08917v1 Announce Type: new Abstract: While vision-language-action (VLA) models have advanced rapidly, running them at every control step remains expensive. Prior work accelerates VLA inference using techniques like action chunking and visual-token pruning, typically evaluating based on latency and average task success. However, acceleration may discard information and break tasks the original policy would solve, a risk hidden by average metrics. Measuring these failures is challenging because action deviations compound over closed-loop trajectories, meaning task failure is only observable across full episodes. We therefore define an acceleration-induced failure via paired rollouts from identical initial conditions, tracking when the reference succeeds but the accelerated policy fails. To manage this, we introduce CARE, an approach for certified accelerator selection. CARE uses paired rollouts on a calibration set to provide finite-sample guarantees that acceleration-induced failure risk stays below a user-specified budget. It deploys the fastest certified candidate, falling back to the reference if none qualify. By relying only on terminal outcomes and measured compute, CARE applies unchanged across diverse acceleration mechanisms, while sequential testing and failure-triggered reference rollouts keep certification affordable. On four LIBERO suites with OpenVLA-OFT, CARE certifies $9.0$--$10.8\times$ speedups while guaranteeing (at $95%$ confidence) that at least $85.8%$ of reference-solved episodes are preserved. Under tight budgets, selectors without guarantees exceed the budget in up to $75%$ of trials, whereas CARE stays within budget and its sequential form uses $78.9%$ fewer rollouts than exhaustive evaluation. CARE further generalizes to flow-step reduction for $\pi_{0.5}$, and to Qwen3.5-9B and Llama-3.1-8B agents in Crafter.

── more in #robotics 4 stories · sorted by recency
── more on @care 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
→ Live at https://your-agent.zahid.host ✓
Get free account → Pricing
from €0/mo · no card required
LIVE [news/care-certifying-acce…] indexed:0 read:1min 2026-10-08 · —