Loading

Proactive Alert Monitoring: LLM Model Latency

Fecha de publicación: Jul 20, 2026
Descripción

The Signature Success plan’s Proactive Monitoring product will monitor for, and alert you about LLM Model Latency when the average response time of the AI model serving your agents is elevated for your org.

If you receive an alert from the Proactive Monitoring service related to this metric, it indicates that end users may be experiencing slow agent responses — long pauses in chat conversations or extended silence on voice calls — even though the agent is still responding successfully.

Solución

Resolution:

Here are some common practices & resources that may help to resolve issues related to this alert:

  • Review any recent changes to the affected agents.

    • Changes such as switching to a different model, significantly longer instructions or prompts, or adding more actions and knowledge lookups per response can increase response time.

    • If latency increased right after a change was deployed, consider rolling it back to confirm.

  • Confirm whether the slowness affects all conversations or only specific agents or channels, and share this detail in your support case — it helps speed up the investigation.

  • If no recent changes were made on your side, consider this as an issue on the Salesforce side.

    • Elevated model latency is often caused by temporary conditions with the underlying AI service. You will be notified through your support case as the issue is investigated and when the service is restored.

Número del artículo de conocimiento

005389011

 
Cargando
Salesforce Help | Article