Skip to content
QASignal Room

Index

All notes

Everything here, in eight sections. If you are setting up a programme, read the foundations first; if you already have one, the measurement and compliance sections are where the surprises usually are.

01

Foundations

Three different activities share the name quality assurance, and they need different instruments. Plus the sampling arithmetic that undermines most agent-level scoring.

02

How the analytics works

Transcription accuracy sets the ceiling for everything downstream, and it varies systematically by speaker. This section covers what each component actually does and where it fails.

03

Programme design

Scorecards, calibration, automation and rollout. Most of the avoidable failure in this field happens here, before any evaluation takes place.

  • Procedure

    Defining Quality Before Measuring It

    Most scorecards encode assumptions about what good looks like that nobody has tested. Deriving them from outcomes takes a quarter and changes what gets measured.

  • Procedure

    Calibration: Keeping Evaluators Consistent

    Two evaluators scoring the same call disagree more than anyone expects. Measuring the disagreement is the only way to know whether your scores mean anything.

  • Analysis

    Automated Scoring: Where It Works

    Scoring every interaction is the headline promise. It works for a narrow band of items and produces disputes everywhere else.

  • Analysis

    What Changes When You Analyse Every Interaction

    Coverage removes the sampling problem and introduces a volume problem. What genuinely improves, and what organisations discover they were not ready for.

  • Analysis

    Targets, Gaming and Goodhart's Law

    Any QA measure that becomes a target stops measuring what it did. The mechanisms are predictable and some are avoidable.

  • Procedure

    Rolling Out a QA or Analytics Programme

    The technical deployment is the easy part. Agent trust and the first month's findings determine whether the programme survives a year.

  • Analysis

    Multi-Language and Multi-Site Programmes

    Analytics quality varies enormously by language, which means a global programme measures some sites better than others and reports the difference as performance.

  • Analysis

    Chat, Email and Messaging Quality

    Text channels remove the transcription problem entirely and introduce their own. Most programmes apply a voice scorecard to them, which measures the wrong things.

04

Running it

Feedback, disputes, evaluator workload and the routing of process findings — which is where the largest value sits and where most programmes stop.

  • Procedure

    Delivering Feedback That Changes Behaviour

    The evaluation is worthless if the conversation produces no change. What the research suggests, and what practice usually does instead.

  • Procedure

    Disputes and Appeals

    A dispute route is not an administrative burden. It is the mechanism that finds broken scorecard items, inconsistent evaluators and transcription failures.

  • Procedure

    Getting Process Findings Out of QA Data

    Much of what QA discovers is not an agent problem. Routing those findings to people who can fix them is the highest-value output.

  • Analysis

    Evaluator Workload and Score Quality

    Score quality degrades measurably with volume and session length. Most programmes set evaluator targets without knowing where that threshold sits.

  • Analysis

    QA in Outsourced and BPO Environments

    When the agents work for someone else, QA becomes a contractual instrument as well as a quality one, and the incentives point in unhelpful directions.

  • Procedure

    The QA Workflow, End to End

    From selecting an interaction to a closed action. Most programmes have the first three steps and stop, which is why the output does not change anything.

  • Analysis

    What QA Does to the People Being Measured

    Contact centre attrition is high and quality programmes contribute to it. The design choices that make monitoring tolerable are known and rarely applied.

05

Measurement

The short set of measures that carry information, the test that validates a whole scorecard, and how to report to people who want one number.

06

Compliance and ethics

Recording consent, biometric templates, employee monitoring obligations, retention, and the measured accuracy disparities that make automated scoring an employment matter.

  • Reference

    Call Recording and Consent

    Recording rules differ by jurisdiction, by party and by purpose. A single global recording policy will be wrong somewhere, and the exposure is real.

  • Reference

    Voiceprints and Biometric Law

    Voice authentication creates a biometric identifier, which is regulated more strictly than a recording. Several statutes carry private rights of action.

  • Reference

    Monitoring Agents: The Employment Dimension

    Continuous analysis of employee speech is employee monitoring, which in many jurisdictions carries notification, consultation and proportionality obligations.

  • Procedure

    Retention: Recordings, Transcripts and Indexes

    Deleting a recording deletes one copy. The transcript, the analytics index, the summary and the backup are separate objects and they are usually forgotten.

  • Analysis

    Bias in Automated Quality Scoring

    Automated scoring inherits the speech recognition accuracy gap, so it can be systematically harsher toward some agents. Measure it.

  • Checklist

    What to Tell Agents

    Agents will find out what the system does. Telling them first is a legal requirement in several jurisdictions and an operational advantage everywhere.

  • Checklist

    Vendor Due Diligence on Your Recordings

    Your recordings leave your estate. Where they go, who hears them, how long they stay and whether they train models are contract questions.

07

Buying and starting

What to test on your own audio, what to ask about your data, and a first quarter that costs nothing and produces a programme worth automating.

08

Reference

Terms defined once, including the ones vendors and practitioners use to mean different things.

  • Reference

    Glossary

    Terms used across these notes, defined once, including the several that are used inconsistently across the industry.

50 notes in total