Compliance
Is it following our rules?
Today, 1 to 3% of calls get checked by hand.
QA and compliance audit for AI voice agents
Loops grades every call your AI agents make against your own standard operating procedures (SOP), catches the release that broke a rule, and shows what each release earned.
Works withRetellVapiBlandElevenLabsLiveKitPipecatin-house agents
The gap
Compliance
Today, 1 to 3% of calls get checked by hand.
Operations
Tests pass. Live calls drift.
Finance
Minutes and containment, not dollars.
>40%
of agentic AI projects will be canceled by the end of 2027 because of escalating costs, unclear business value or inadequate risk controls. Loops puts evidence behind the last two, call by call. Source: Gartner, June 2025
How it works
Four steps, then the loop closes: fix a rule once, and the next release is proven against the last.
1Start from your rules
Upload the SOP you already have. Loops splits it into numbered rules, and your compliance lead approves them before anything is graded.
2Check every call
Loops grades each call, AI or human, rule by rule. Every verdict quotes the line it relied on, and Loops checks that required tool calls actually fired.
AgentHi, this is Maya on a recorded line. Am I speaking with Daniel Reyes?
CustomerYes, that's me.
AgentYour balance with Alder Card is $1,284.16, 94 days past due.
AgentYou're set for $107 on October 3.
3Prove every release
Compare each rule's miss rate, this release against the last, and open the calls behind any number that moved.
4See what it earned
Join calls to your payment records and see what each release earned or cost, next to the rule changes that explain it.
Loops drafts the updated prompt, tests and checks. Your team reviews and ships them. Loops never writes to your agent.
Watch
Two short explainers, captioned.
Independent by design
A platform can't referee its own product. Loops sits above whichever one you use and has no stake in the number.
You
PlanWrite and approve the SOP.
Your platform or team
BuildRetell, Vapi, Bland, ElevenLabs or your own code runs the agent.
Loops
Test, check, proveGrades every call and every release against your rules.
You, with Loops' drafts
ImproveYour team reviews the drafted fix and decides what ships.
Compare
Platform QA and test tools are good at their jobs. Loops does a narrower one, and most teams use it alongside both.
| Capability | LoopsIndependent audit | Platform QARetell, Vapi, Bland | Agent test toolsCoval, Cekura, Hamming | QA team todayManual sampling |
|---|---|---|---|---|
| Graded against your own SOP, rule by rule | Approved, versioned rules | Custom criteria, their format | Custom metrics, their format | On the calls they sample |
| Every production call, not a sample | Every call | Their own calls | Varies by tool | 1 to 3% by hand |
| Human and AI calls on one rubric | Same rules, same screen | AI only | AI only | Humans only |
| This release against the last, per rule, in production | With the calls attached | Versions, no per-rule diff | Mostly before release | No |
| Dollars per release from your payment records | Yes | No | No | No |
| Independent of whoever runs the agent | Yes | Part of the platform | Yes | Yes |
| Real-time latency and interruption monitoring | No, Loops grades after the call | Yes | Yes | No |
Based on each vendor's public documentation, September 2026. If a cell is out of date, tell us through the form and we'll fix it.
Regulated floors
Start from a library of regulatory rules for your industry, then add your own.
FDCPA · Reg F · TCPA · state rules
State unfair claims practices laws
HIPAA · billing and scheduling
Security and data
Where your platform offers read-only keys, we use one. Where it doesn't, we use a dedicated key only to read calls, and log every request.
Prompt changes arrive as drafts for your team. Loops can't edit prompts, flows, settings or numbers.
Pre-release runs use a separate permission, limited to a test agent you name. Leave it off and Loops still compares releases on live calls.
Call data is stored and processed in the United States and encrypted in transit and at rest. You set how long we keep it.
Nothing you send trains any model, ours or a provider's.
NDA and DPA, plus a BAA for healthcare, signed before we receive a key. Our SOC 2 Type II report and bridge letter come with the security packet, under NDA.
The free audit
No build work on Retell, Vapi, Bland or ElevenLabs. LiveKit, Pipecat and in-house agents add a small SDK.
Sign the NDA and DPA (and a BAA for healthcare), connect your platform, choose which calls to include and send your SOP.
Your compliance lead reviews each rule beside the SOP text it came from.
Last month's calls, graded rule by rule, with the evidence for every verdict.
Yours to keep, including how often Loops agreed with your QA reviewers. If you don't continue, we delete your calls, transcripts and recordings within 30 days and confirm it in writing.
FAQ
Loops is independent QA and compliance audit software for AI voice agents. It grades every call against your own SOP, rule by rule, with the transcript line as evidence, compares each release with the last, and reports what each release earned from your payment records.
Give Loops an API key from your platform and upload your SOP. Loops splits the SOP into numbered rules for your compliance lead to approve, backfills last month's calls, and returns the first audit within 72 hours of approval.
Yes, through the Loops SDK for Python or Node. After each call ends, it sends the transcript, tool calls and agent version, off the audio path, so it never slows a live call.
Every call carries the agent version your platform records, or a version tag you set. Loops compares each rule's miss rate between versions on live calls and lists the calls behind any change. Optional pre-release test calls catch problems before a release ships.
Send a daily payments file from your collection or billing system (account ID, amount, date posted). A payment counts toward the release that handled the account's last call before it posted, within a window you set, 30 days by default. Releases are compared on calls from the same weeks, the same clients and similar balance and age bands, and the report shows the range around every figure and says when a difference is too small to call.
No. Loops never writes to your platform. When you fix a rule, it drafts the prompt change, tests and checks for your team to review, and your team decides what ships.
For any call, account or date range: each rule's verdict, the quoted line, the tool-call record, the rule version and who approved it, and any reviewer override with who made it and when. Export it as PDF or CSV.
Yes. Every verdict, with its rule number, the quoted line and the call ID, exports as CSV and is available to your CRM or compliance log by API.
Every verdict quotes the line and the tool-call record it relied on, so a reviewer can check it in seconds. When a transcript can't settle a rule, Loops marks the call for review rather than guessing. Your QA team can overturn any verdict, and your day-30 report measures how often Loops agreed with your reviewers on your own calls.
Yes. Send recordings from your human floor by upload or SFTP, and they're graded against the same rules as your AI agent, on the same screen.
Yes, when you send your dial log. Without it, Loops grades what happens inside each call and marks cross-call rules as not checked.
Yes. Each of your clients approves its own rules and gets its own report, and their calls stay separate. Tell us about your platform through the form and we'll walk through how verdicts come back to you.
Free 30-day audit
Tell us what you run. A person on our team replies within one business day with next steps.