papersSEP 10 04:00 UTC
RubricRefine: Training-Free Pre-Execution Refinement for More Reliable Tool-Use Agents
A new arXiv paper investigates whether iterative self-refinement can make language model agents that interact with tools through code more reliable at inference time. The authors show that the benefits of refinement depend heavily on the feedback format, with unstructured critique producing inconsistent results across models. They introduce RubricRefine, a training-free method that refines agent outputs before execution using rubric-based feedback.