Enterprise Models

Salesforce's Koa Pilot Puts CRM Actions in the Model Evaluation

By Kaleido Field Staff ยท September 18, 2026

A pilot is the current access state

Updating an opportunity is one of the tasks behind Salesforce's Koa announcement. Introduced September 15, the Nemotron-based CRM reasoning model is available to selected Agentforce pilots, with US general availability expected in winter 2026. Its reported benchmark performance remains a Salesforce claim.

Citation-ready: Koa is a selected-customer CRM reasoning pilot; Salesforce's internal benchmark does not establish a customer-specific production success rate.

Evidence boundary: Vendor announcement and vendor-reported model evaluation. No independent benchmark reproduction, customer-data audit, CRM write or generalized performance ranking.

Official Salesforce and NVIDIA partnership artwork accompanying the Koa announcement
Image source: Salesforce; official partnership announcement artwork. Used for editorial coverage of task-specific model evaluation desk.

What happened and why it matters

The task-specific model brings evaluation closer to an operational action, such as a record update. That is useful only when the evaluation keeps the authorized action, its inputs and the resulting CRM record together.

Primary evidence

Primary reference: Salesforce September 15 Koa announcement. Kaleido Field checked the event date and the article's attributed facts against this source.

Source check
Source dateSeptember 15, 2026
Checked by Kaleido FieldSeptember 18, 2026, CST
Source functionenterprise models -> task-specific evaluation evidence

Synthetic training is a provenance claim

Salesforce says it post-trained Nemotron 3 Super on synthetic enterprise scenarios, without customer data, and controls the model weights and inference infrastructure. Those are the company's descriptions of development and deployment.

They answer different questions from whether the model handles an organization's unusual fields, permissions or exception cases. A procurement review would need evidence for those local conditions rather than treating training provenance as an accuracy guarantee.

Count the right action on the right record

A practical test can specify the opportunity, permitted field, intended value and required approval before the agent runs. The result is then checked in the destination record. An articulate explanation does not substitute for that readback.

Failure categories also matter. A missing permission, ambiguous instruction and wrong tool argument may all prevent completion, but they call for different repairs. This is a suggested evaluation design; no pilot was run for this report.

Keep deployment choices separate

The same announcement discusses Missionforce models for specialized environments. Its availability schedule is distinct from Koa's. Buyers should follow the named product and region rather than combining every statement into one present-day offer.

Our CRM connector report examines approval and write-back through another product. Koa changes the model choice while leaving the need for an inspectable result.

Evidence boundary

Vendor announcement and vendor-reported model evaluation. No independent benchmark reproduction, customer-data audit, CRM write or generalized performance ranking.

Reader briefing

Keep the source trail in view.

One concise email when a model, benchmark, or visual-intelligence claim materially changes.

FAQ

Is Koa generally available now?

The September 15 source describes selected pilots and expects US general availability in winter 2026.