AI response modes & follow-up limits
Two settings define the AI's day-to-day behavior: what it does with a message (response mode) and how long it keeps trying (follow-up limits). Both live in Settings → AI Assistant.
Response modes
When a message comes in offers two modes:
Mode | Behavior | Best for |
|---|---|---|
Reply directly | The AI posts a customer-visible answer in the thread. | Deflection — let AI resolve routine questions end-to-end. |
Add internal note | The AI drafts an answer as a private internal note for agents; nothing goes to the customer. | Review-first teams — agents send (or edit and send) the draft. |
The mode is workspace-wide and applies to all AI-handled conversations. Switching modes changes behavior for new messages; existing conversations continue with their current state.
With Reply directly, agents can still take over any conversation at any moment — replying as a human immediately sets takeover.
Confidence threshold
The Confidence select (10%–100%) decides when the AI is allowed to answer:
AI answers only when its confidence in a grounded answer is at or above the threshold.
Higher confidence → fewer AI answers and more handoffs (safer, slower).
Lower confidence → more AI answers (more deflection, higher risk of a wrong answer).
The tooltip guidance: "AI only answers when its confidence is at or above this value. Higher confidence means fewer AI answers and more human handoffs."
Start around 70% and adjust from observed outcomes: wrong answers → raise; too many unnecessary handoffs → lower.
Follow-up limits ("Handoff after")
The Handoff after select (1–10) bounds AI persistence on the same problem:
Counts AI follow-up replies on the same problem in a conversation.
Resets when the visitor brings up a new problem.
Reaching the limit forces handoff to a human (see Stuck detection & same-issue handoff).
Talk to a human
The Show "Talk to Human" button toggle controls whether visitors can request a person at any time. On means the customer always has an exit from AI; off, escalation happens only via AI judgment. Recommended: keep it on — a customer-driven escape hatch builds trust and produces cleaner escalations.
Escalation messages
Customer-facing copy shown when the AI hands off, per scenario:
Default — normal handoff, e.g. "Let me connect you with a team member — they typically reply in {reply_time}."
Team busy — handoff while the team is busy, using
{reply_time}for the expected reply time.After hours — handoff outside business hours, using
{next_open}for when the team returns (accurate only when business hours are enabled).
Each field shows a live preview. These templates support the placeholders shown below the field.
Where this fits with routing
Response mode and limits decide how AI behaves; routing decides where handoffs land. Configure both — see Inboxes / routing settings for the destination and assignment behavior.
Related pages
Was this article helpful?