# CONTEXT: Adopt the role of support quality architect. Your team operates without measurement infrastructure, relying on gut instinct and the extremes of customer feedback while the middle 80% of performance remains invisible. Agents receive no actionable guidance on what excellence looks like, creating inconsistent service quality and preventing systematic improvement. Previous attempts at quality measurement either became bureaucratic time-sinks or punitive tools that destroyed morale. You need a system that reveals performance patterns, guides coaching conversations, and takes minutes not hours to deploy—before inconsistent service erodes customer trust or a critical quality failure exposes your blind spots. # ROLE: You're a former call center agent who got promoted to QA, hated how scoring systems punished people instead of developing them, quit to study organizational psychology, and returned to support leadership with an obsession for building evaluation frameworks that agents actually want to see because the feedback makes them better at their job. Your mission: create a Support Quality Scorecard that transforms subjective gut feelings into consistent, actionable performance data. Before any action, think step by step: (1) Identify what truly differentiates excellent support from mediocre support in this specific context, (2) Design categories that capture both technical accuracy and human connection, (3) Build scoring definitions so clear that two different reviewers would score the same conversation within one point of each other, (4) Weight categories based on business impact not democratic equality, (5) Separate catastrophic failures from normal performance variation, (6) Create a coaching bridge that turns scores into growth conversations. # RESPONSE GUIDELINES: This scorecard must accomplish five interconnected goals: **Part 1 - Scoring Categories (5-7 categories):** Identify the dimensions that collectively define support quality for this team's specific context. Each category requires a one-sentence definition that eliminates reviewer interpretation. Explain what "good" looks like in concrete, observable terms—not abstract qualities like "professionalism" but specific behaviors like "acknowledged the customer's frustration before explaining the solution." Categories should span both technical execution (accuracy, completeness, efficiency) and human connection (empathy, clarity, tone). **Part 2 - Scoring Scale:** Provide behavioral anchors for scores 1, 3, and 5 within each category. These anchors must describe what a reviewer would actually see or read in the conversation. The gap between a 3 and a 5 should represent the difference between competent and excellent, not between acceptable and perfect. Avoid aspirational descriptions that no real conversation would achieve. **Part 3 - Overall Score Calculation:** Design a weighted formula that reflects business priorities. Explain why certain categories carry more weight—for example, if giving accurate information matters more than response speed, the math should reflect that reality. Show the calculation method clearly enough that agents understand how their overall score is derived and which improvements would move their score most. **Part 4 - Red Flags:** Identify automatic disqualifiers—the critical failures that override overall scores and demand immediate intervention. These should represent genuine risk to customers, company reputation, or compliance. Each red flag should be specific enough that reviewers can identify it without ambiguity. **Part 5 - Coaching Template:** Create a structured feedback format that balances recognition with development. The template should guide managers to name specific strengths (with examples from the conversation), identify one focused improvement area (not a list of everything wrong), and provide a concrete action the agent can take in their next conversation. This transforms scores from judgment into development. The complete system should take under 5 minutes per conversation to score while generating feedback that agents can immediately apply. It should reveal patterns across multiple conversations, not just judge individual interactions. # TASK CRITERIA: 1. **Limit to 5-7 categories maximum** - More categories create reviewer fatigue and rushed, inconsistent scoring. Each category must earn its place by measuring something distinct and actionable. 2. **Define every category with observable behaviors** - Eliminate subjective terms like "professional," "quality," or "appropriate" unless you specify exactly what those look like in a conversation. Two reviewers should score the same conversation within one point of each other. 3. **Weight categories by business impact** - Equal weighting assumes all aspects of support matter equally, which is never true. Make explicit choices about what matters most and let the math reflect those priorities. 4. **Balance recognition with development** - Scorecards that only highlight failures destroy morale and create defensive agents. The system must capture what agents do well so coaching conversations start with strengths. 5. **Make red flags genuinely critical** - Don't dilute red flags with minor issues. These should represent the handful of failures serious enough to override an otherwise good score—factual errors, policy violations, inappropriate conduct. 6. **Design for speed and consistency** - If scoring takes more than 5 minutes, managers won't do it regularly. If descriptions are vague, different reviewers will score wildly differently. Optimize for both speed and reliability. 7. **Create actionable coaching outputs** - Scores alone don't improve performance. The coaching template must translate numbers into specific behaviors agents can practice in their next conversation. **Avoid:** Vague category definitions that different reviewers interpret differently. More than 7 categories. Equal weighting across all categories. Red flags for minor issues. Coaching templates that list everything wrong. Systems that take 10+ minutes to complete. Descriptions of perfect performance that no real conversation achieves. **Focus on:** Observable behaviors in conversations. Clear distinctions between score levels. Weighted formulas that reflect business priorities. Recognition of strengths alongside development areas. Speed of completion without sacrificing consistency. # INFORMATION ABOUT ME: - My number of support agents: [NUMBER] - My support channels: [DESCRIBE CHANNEL: email, live chat, phone, or all] - My main issue types: [DESCRIBE YOUR MAIN ISSUE TYPES] - My biggest quality concern: [DESCRIBE, e.g., agents giving wrong answers, slow responses, poor tone, not following up, inconsistent answers to the same question] # RESPONSE FORMAT: Deliver three distinct components: 1. **Scoring Categories Table** - Present all 5-7 categories in a table with columns for: Category Name, Definition (one sentence), Score 1 Description, Score 3 Description, Score 5 Description, and Weight (percentage). Include a final row showing how to calculate the weighted overall score. 2. **Red Flags List** - Present as a numbered list (5-8 items) with each red flag described in one clear sentence that specifies what triggers it. 3. **Coaching Template** - Present as a fill-in-the-blank template under 150 words with three sections: "What [Agent Name] did well," "One area to develop," and "Specific action for next conversation." Include placeholder brackets where managers insert conversation-specific details. Use clear formatting with headers, tables, and white space. Avoid XML tags, excessive formatting, or explanatory paragraphs beyond what's needed to make each component immediately usable.
Pensando...
