Quality assurance · small support teams

CXCalibrateAI quality assurance calibrated to your policies — not a generic scorecard.

Upload your SOPs, escalation rules, and macros. Drop in a CSV of recent conversations. Every deduction cites the specific rule behind it. Manager corrections are stored as workspace calibration data for review and future calibration workflows.

Pilot entry tier
From $49/mo
Onboarding
CSV-first
Output
Evidence trail
Ticket #C-7741 — Refund escalation
Score 82 / 100

Customer

“Order #40102 arrived damaged, want a refund. Email sent 3 days ago.”

Agent (A. Patel)

“I’m sorry about that. I’ve started a refund — could you share the order number again so I can speed it up?”

Deducted 18 points — three rule citations

  • −8SOP-3.4 “Confirm order ID before processing refund.” Agent asked instead of confirming.
  • −6ESC-2.1 Damaged-item case requires photo evidence request. None logged.
  • −4TONE-1 Acknowledged the issue but no SLA reference.
Evidence trail preserved

What makes us different

Enterprise QA, minus the enterprise.

Zendesk QA, Klaus, MaestroQA, Scorebuddy, and RevelirQA built their tools for big contact centers. CXCalibrate is for the 5–50-agent team that wants real AI scoring without a six-figure contract or a consulting engagement to turn it on.

  1. 01

    Calibrated to your SOPs

    No generic CVSS-style rubric. The model scores against the SOPs, escalation rules, and macros you upload — so a 92 here means your 92, not the industry average.

  2. 02

    Every deduction cites a rule

    Each score drop names the policy it came from: SOP-3.4, ESC-2.1, TONE-1. Managers trust the number because they can see the line of text it came from.

  3. 03

    Manager corrections are stored

    Manager corrections are stored as workspace calibration data for review and future calibration workflows. Human review remains authoritative.

  4. 04

    No implementation project

    CSV in, report out. No PS engagement, no Zendesk admin hours, no replatforming. Live an audit, see results this week, subscribe only if it earns its place.

How the loop works

A calibration loop, not a one-time grade.

Most AutoQA assigns a score and forgets. Manager corrections are stored as workspace calibration data for review and future calibration workflows, adjusted to your team’s actual standards.

  1. 01Upload

    Bring your policies in.

    SOPs, escalation rules, response templates — TXT or Markdown. CXCalibrate parses them into the rubric it scores against. (PDF / DOCX / Notion are not yet supported in this milestone.)

  2. 02Score

    Evaluate every ticket.

    Drop in a CSV of recent conversations. Each interaction gets scored against your rubric with cited rule IDs and excerpts.

  3. 03Dispute

    Correct what’s wrong.

    A manager thinks a deduction is unfair? They mark it, leave a note, and the system records it as a labeled correction.

  4. 04Calibrate

    Manager corrections are stored.

    Manager corrections are stored as workspace calibration data for review and future calibration workflows. Human review remains authoritative.

Why an evidence trail matters

A red score is easy. A defensible one is the product.

Without citations, QA scores are arguments. With them, they’re conversations. Every Every CXCalibrate evaluation links the deduction to the line of your SOP it came from — and to the exact moment in the transcript where the rule applied.

  • Per-ticket trail. Click any score drop to see the rule, the excerpt, and the agent excerpt side-by-side.
  • Aggregated reports planned. Agent, team, and risk-pattern roll-ups are not part of this milestone.
  • What we do with your data. Customer transcripts and uploaded policies are stored in this app’s database. The LLM evaluator receives the conversation text and the policy you uploaded and returns a structured judgment. Today we do not automatically strip PII or scan uploaded CSVs for personal data — upload only data you are entitled to process.

Report / Week of Aug 4 · Agent view · A. Patel

82 / 100
↑ 4 pts vs last week

Excerpt cited

“Damage claims over $50 require photo proof within 24h.” — SOP-3.4

Transcript moment

Agent acknowledged refund but did not request photo evidence.

Deduction

−8

Disputes

2 upheld

Excerpt cited

“Tone I: reference the order SLA in your first reply.” — TONE-1

Transcript moment

No SLA or expected resolution date given.

Deduction

−3

Disputes

0

Excerpt cited

“Use macro REF-FAST for order-ID confirmations.” — MACRO-RF

Transcript moment

Manual reply instead of macro; tone consistent.

Deduction

−2

Disputes

0

Aggregated coaching / trend reports across this batch or across weeks — planned.12 more rows →

Pricing

Three plans small teams can run start to finish.

No quote, no annual commit, no PS line item. Run the free 25-conversation QA audit first; subscribe only if it earns its place in your weekly rhythm.

Pilot
Most teams start here

$49/ month

250 conversations per month · 3 manager users.

  • Policy upload (TXT / Markdown)
  • CSV upload of conversations
  • Evidence-trail scoring with rule citations
  • Manager approve / dispute / correct
  • Row-level evaluation view
  • Agent / team / risk-pattern reportsplanned
Start with the free audit
Team

$129/ month

2,000 conversations per month · 10 manager users.

  • Everything in Pilot
  • Policy upload (TXT / Markdown)
  • CSV upload of conversations
  • Evidence-trail scoring with rule citations
  • Manager approve / dispute / correct
  • Row-level evaluation view
  • Agent / team / risk-pattern reportsplanned
Start with the free audit
Growth

$249/ month

7,500 conversations per month · 25 manager users.

  • Everything in Team
  • Policy upload (TXT / Markdown)
  • CSV upload of conversations
  • Evidence-trail scoring with rule citations
  • Manager approve / dispute / correct
  • Row-level evaluation view
  • Agent / team / risk-pattern reportsplanned
Start with the free audit

Out of scope: healthcare, banking, insurance claims, government, voice transcription, and autonomous disciplinary decisions. CXCalibrate is for scoring — not for adjudicating outcomes we shouldn’t touch.

Questions worth asking

Frequently asked

If your question isn’t here, write to cxcalibrate@polsia.app.

Those platforms were built for contact centers — they want a 12-month contract, a PS engagement, and access to your conversation data. CXCalibrate uses small-team pricing with no implementation project, runs against your own uploaded policies instead of a fixed rubric, and surfaces a calibration loop via the dispute workflow. We are explicitly aimed at the 5–50-agent team at the back end of the QA market.

Free 25-Conversation QA Audit

Send us a CSV. We’ll send back a scored report.

No contract, no implementation. You get 25 anonymized conversation evaluations against your own policies, with the evidence trail — and a one-line answer to whether CXCalibrate is worth subscribing to.