Spotlight is in beta. The page shows a Beta badge and Free during beta: scoring does not use credits.
Before you start
- Spotlight needs at least 10 evaluated conversations. Until then the page shows Getting started with a progress bar (“N of 10 conversations”).
- A conversation is evaluated after it is closed, if the customer sent at least two messages. Popcorn looks for newly closed conversations every 15 minutes, so scores appear shortly after a conversation is closed. Conversations older than 120 days are not evaluated.
- Every participant is scored separately: the AI agent and each teammate who replied. A conversation with both counts once in the conversation totals.
Open Spotlight and filter
1
Go to Spotlight
Open Agents → Spotlight. The last 7 days are shown; switch to 14 d, 30 d or Custom at the top right.
2
Narrow the view
Use All Actors to pick one agent or teammate (each entry shows its number of evaluations), All Types for Sales, Support or General, and All Sentiments for Positive, Neutral, Negative or Frustrated.
3
Choose who handled the conversations
The All / AI Agent / AI + Human / Human tabs filter by who replied: only the AI, both, or only your team.
Performance Score
The Performance Score is the average of all evaluations in the range, out of 5.0, with the change against the previous period next to it. Score Trend plots the daily score across the range (hourly when the range is a single day) and appears once there are at least two points. Each evaluation’s score is a weighted average of the six dimensions below.Score Dimensions
Each dimension is scored from 1 to 5, and each card shows the average with its change and a small trend line. The scale is deliberately strict:
- 3 means the job was done adequately. A correct answer with no problems is a 3, not higher.
- 4 needs clear evidence of doing more than expected; 5 is exceptional and rare.
- 2 means something went meaningfully wrong; 1 means the customer experience was harmed.
- Short exchanges (a greeting, a single question and answer) cannot score above 3 on Understanding, Resolution, Communication or Helpfulness.
- Instructions compares the AI’s replies with the agent’s instructions. Teammates are not judged on instructions and always get a 3 there.
- A handover is not automatically a failure: handing over something the agent could have answered from its knowledge or tools scores low, handing over something it genuinely could not resolve scores as adequate.
Key figures
Breakdowns
Resolution Breakdown- Confirmed: the customer explicitly said the issue was solved.
- Assumed: a concrete answer was given and the customer kept engaging, without confirming.
- Escalated: the conversation was handed over to someone else.
- Abandoned: the customer stopped replying, or nothing was resolved.
Conversations to Review
The table lists up to ten conversations with the lowest score in the range, one row per conversation. Each row shows the Score, the Primary Issue (the evaluator’s note on the lowest-scoring dimension), the Sentiment, the Type and the Date. Click Review to open the conversation in the Inbox, or View Inbox to go to the Inbox itself. When nothing scored low, the section reads All conversations look good. In the Inbox, a closed conversation that scored below 3.5 shows a Quality Score card in the conversation details with the primary issue; Details expands the score of every dimension.Acting on what you see
Filter by All Actors to compare agents, or to see how teammates score after a handover. After publishing a change, compare the score of the following period with the one before it. Per-version outcome figures are also listed in the agent’s Version history; see Test and publish.
Related
- AI Manager: the real-time check that retries weak replies before they are sent. Spotlight scores whole conversations after they close.
- Reports: volumes, response times and handover rates.