AI SDR testing
How to measure an AI SDR funnel
Define evidence-backed stages from meaningful reply to booked meeting, then trace every stalled cohort to the workflow that produced it.
By Sentinium AI · · 2 min read
Define stages with observable evidence
A funnel stage should correspond to something the system can support from the conversation or tool record. Avoid labels that depend on optimistic interpretation. Positive sentiment is not automatically qualified interest, and a proposed time is not a booked meeting.
Write stage rules before comparing agents. If definitions change between runs, the apparent movement can come from the measurement rather than the system.
Keep conversation and tool state connected
The agent may announce a meeting before the calendar tool succeeds. It may classify an account as qualified while missing a constraint recorded earlier. The funnel should require the combination of evidence appropriate to each stage.
This connection also makes failures diagnosable. Teams can distinguish a persuasive conversation with a scheduling failure from a calendar action created without genuine buyer agreement.
Analyze leaks by buyer cohort
Open the cohort that stalled between two stages and compare it with buyers who progressed. Look for differences in fit, authority, urgency, objections, timing, and the agent actions they received. The purpose is to find a pattern that can change the product or narrow the intended audience.
Synthetic rates should be used for controlled comparison and diagnosis, not presented as guaranteed pipeline. Production outcomes remain the external validation of commercial performance.
Connect every finding to a testable change
A useful funnel report points from the stalled stage to the exact conversations, actions, tool results, and telemetry behind it. The team can then change the relevant prompt, context, policy, tool, cadence, or orchestration layer.
Rerun the candidate against the same reviewed population. If progression improves without creating a guardrail or workflow regression, preserve the comparison as evidence for the release review.
A seven-stage evidence model
| Stage | Required evidence |
|---|---|
| Meaningful reply | The prospect engages with substance |
| Positive interest | The prospect expresses openness or interest |
| Qualified interest | Interest includes a plausible evaluation or use case |
| Meeting proposed | A meeting or demo enters the conversation |
| Time agreed | The prospect explicitly accepts a time |
| Calendar created | A successful calendar action is observed |
| Booked and confirmed | Time agreement and calendar creation both hold |
Related AI SDR guides
How to compare two AI SDR versions
Use paired buyer worlds to separate real behavioral changes from audience noise, then inspect the trajectories behind every important delta.
Read guideHow to test an AI SDR before launch
A practical release framework for testing the real outbound agent across complete buyer journeys, business outcomes, stops, and handoffs.
Read guideAVAILABLE NOW
Apply these ideas to outbound sales agents and AI SDRs
See how Sentinium simulates complete buyer journeys, compares agent versions, and monitors production trajectories.
Explore the use case