Skip to content

Audit the AI agents you already run

Evidova reads the conversations your agents already have, on the platform they already run on, and scores every one. Nothing about your stack changes.

Each platform below has a page of its own, setting out what we read from it and what we leave alone. A platform with no page yet still works the same way: connect it in the app and the conversations arrive.

Your own stackNo adapter fits? Post the events yourself, or upload the history you already have. The same rules score them either way.

What differs between them

Every platform here ends up scored by the same rules. What changes is how much each one hands us to score, and that is set by the platform rather than by us.

Audio, on 8 of 12
Where the platform hands over the recording, a score can rest on how something was said and not only on what the transcript holds. Where it does not, the transcript is all there is, and checks about tone or interruption cannot run.
Millisecond timestamps, on 6 of 12
Seconds are enough to measure a reply time and too coarse to order two turns that landed in the same second. On a second-precision platform, turn order is reconstructed from the transcript rather than read off the clock.
Pushed to us, on 10 of 12
A webhook platform posts each conversation as it finishes, so a score follows within seconds. The rest are read on a schedule, so their conversations arrive minutes later. Neither changes the score, only when it lands.
Signed deliveries, on 6 of 12
Where a platform signs what it sends, we verify it and refuse anything that fails. Where it publishes no signing scheme, the token in your webhook URL is the only thing protecting it, so it should be treated as a password and rotated if it leaks.

All twelve are set out side by side, with the exact behaviour of each, on the platform matrix.

Platform matrix

What stays the same

The rules you approve are the rules every conversation is held to, whichever platform it arrived on. A rule written for your phone line applies to the same question asked over WhatsApp, so one agent handled well on one channel and badly on another shows up as a difference in the score rather than a difference in the ruler.

English, Hindi, Tamil, Telugu and Hinglish are scored by the same rules across voice and chat. Nothing about your agent changes on any of these platforms: we read what it has already produced, and we never sit between it and your customer.