papersSEP 10 04:00 UTC
Physically Grounded Proactive Modeling for Retail Agents from Sparse Third-Person Video
Researchers present a study on proactive agents that must both select actions and decide whether available evidence justifies acting. Using sparse third-person video in retail service scenarios, the approach grounds decisions in human-object interactions so agents can anticipate customer needs before an explicit request is made.