Enterprise Models
Salesforce's Koa Pilot Puts CRM Actions in the Model Evaluation
Updating an opportunity is one of the tasks behind Salesforce's Koa announcement. Introduced September 15, the Nemotron-based CRM reasoning model is available to selected Agentforce pilots, with US general availability expected in winter 2026. Its reported benchmark performance remains a Salesforce claim.
Citation-ready: Koa is a selected-customer CRM reasoning pilot; Salesforce's internal benchmark does not establish a customer-specific production success rate.
Evidence boundary: Vendor announcement and vendor-reported model evaluation. No independent benchmark reproduction, customer-data audit, CRM write or generalized performance ranking.

What happened and why it matters
The task-specific model brings evaluation closer to an operational action, such as a record update. That is useful only when the evaluation keeps the authorized action, its inputs and the resulting CRM record together.
Primary evidence
Primary reference: Salesforce September 15 Koa announcement. Kaleido Field checked the event date and the article's attributed facts against this source.
| Source date | September 15, 2026 |
|---|---|
| Checked by Kaleido Field | September 18, 2026, CST |
| Source function | enterprise models -> task-specific evaluation evidence |
Synthetic training is a provenance claim
Salesforce says it post-trained Nemotron 3 Super on synthetic enterprise scenarios, without customer data, and controls the model weights and inference infrastructure. Those are the company's descriptions of development and deployment.
They answer different questions from whether the model handles an organization's unusual fields, permissions or exception cases. A procurement review would need evidence for those local conditions rather than treating training provenance as an accuracy guarantee.
Count the right action on the right record
A practical test can specify the opportunity, permitted field, intended value and required approval before the agent runs. The result is then checked in the destination record. An articulate explanation does not substitute for that readback.
Failure categories also matter. A missing permission, ambiguous instruction and wrong tool argument may all prevent completion, but they call for different repairs. This is a suggested evaluation design; no pilot was run for this report.
Keep deployment choices separate
The same announcement discusses Missionforce models for specialized environments. Its availability schedule is distinct from Koa's. Buyers should follow the named product and region rather than combining every statement into one present-day offer.
Our CRM connector report examines approval and write-back through another product. Koa changes the model choice while leaving the need for an inspectable result.
Evidence boundary
Vendor announcement and vendor-reported model evaluation. No independent benchmark reproduction, customer-data audit, CRM write or generalized performance ranking.
FAQ
Is Koa generally available now?
The September 15 source describes selected pilots and expects US general availability in winter 2026.