papersTODAY 04:00 UTC
ReWeight Uses Human Demonstrations and Sample Weighting for VLA Post-Training
A new arXiv preprint proposes ReWeight, a technique for post-training vision-language-action models when in-domain robot demonstrations are scarce. The approach retrieves relevant egocentric human demonstrations and applies sample weighting to make better use of that human data, since collecting robot-specific data is expensive. The paper targets adapting VLA models to particular robots and tasks.