Salesforce Releases Koa Model, Self-Tested CRM Task Success Rate Beats GPT-4.1
Salesforce launches Koa, an enterprise model based on Nemotron, which self-tests show outperforms GPT-4.1 in CRM agentic tasks.
ImportânciaMaterialEvidênciaE2 não replicadaTratamentoRápido
Salesforce has released Koa, a model achieving an 87% task success rate in CRM agentic tasks, surpassing GPT-4.1's 82% in self-tests.
Built by post-training the open-weight Nemotron-3-Super-120B foundation with reinforcement learning, Koa specializes in multi-turn business tool routing and argument invocation. The pipeline uses declarative Agent Script specifications to generate synthetic training data, enabling domain specialization without using customer privacy data.
On CRMAgentBench and internal production benchmarks, Koa demonstrates superior performance over its base model and GPT-4.1 while maintaining general capabilities on public benchmarks like Tau2Bench. These results are currently vendor-reported and have not been independently reproduced.