Salesforce Koa: enterprise LLM post-trained from Nemotron-3-Super-120B with GRPO
Salesforce has introduced Koa, an enterprise-focused language model created by post-training the open-weight Nemotron-3-Super-120B foundation model. The training process uses reinforcement learning with Group Relative Policy Optimization (GRPO), drawing on public data and other sources. The work targets agentic tool use in enterprise settings, according to the arXiv preprint.