papersTODAY 04:00 UTC
IMPACT-VLA attributes robot policy behavior using counterfactual trajectories
A new arXiv paper introduces IMPACT-VLA, a method for tracing how much each input modality — camera images, proprioceptive state, and language instructions — contributes to a vision-language-action policy's decisions at different points during task execution. The approach relies on counterfactual trajectories to isolate the effect of individual inputs, addressing the difficulty of interpreting these multimodal robot policies. The abstract excerpt does not detail experimental results or benchmarks.