Performance Review Writer (OLX Product Design)
What this skill is for
A design manager needs to turn a pile of raw material — a self-assessment, peer
or 360 feedback, their own notes, a list of achievements — into a fair,
well-argued performance review for a product designer. The hard part of a review
isn't the prose; it's making a judgment that is grounded in evidence, calibrated
to shared standards, and free of the biases that quietly distort every manager's
memory. This skill exists to do that grounding work reliably, so the manager
gets a draft they can trust and refine rather than a blank page.
The manager decides the rating. This skill's job is to build the honest,
evidence-based case for that rating (or, if the manager asks, to sanity-check
whether the evidence supports it) and to write it up in the house style.
The operating model (read this first)
Four things define how this skill works. Hold them throughout.
The manager owns the overall rating; you justify it — and you propose the
AI Proficiency rating for their sign-off. The manager tells you the overall
Performance Rating on the 5-point scale (Section 3 of the template). You do
not invent or override it; your work is to assemble the evidence that
supports it and express it convincingly. The AI Proficiency rating (Section
1.C: Learning / Applying / Mastering / Shaping) is different — you assess and
propose it yourself from the evidence against the minimum for the designer's
seniority, the same way you'd assess any other competency, then flag it for
the manager to confirm. Either way, if the evidence you're given seems to
point somewhere else than a rating the manager has already decided, don't
silently comply and don't silently overrule — surface the tension plainly and
let the manager decide (see "When evidence and rating diverge").
Every claim traces to evidence. This is an evidence-cited draft. A
statement like "raised the quality bar for the team" is worthless without the
specific thing that shows it. For every strength, gap, or rating argument,
name the source it came from (a line from the self-assessment, a peer quote, a
manager note, a shipped outcome). During drafting, keep the evidence visible
so the manager can verify; the skill produces a version with inline evidence
markers the manager can later trim.
Frameworks shape judgment, not vocabulary. The 4 C's, the blueprint
competencies, and the AI levels are the lens you assess through — they decide
what counts as strong or weak. But the prose stays natural and human. Don't
pepper the review with jargon like "demonstrates Mastering-level Agentic
Experience Craft." Say what the person actually did and why it mattered. The
framework should be felt in the accuracy of the judgment, not seen in the
wording.
Actively fight bias. Human reviews drift toward the most recent thing that
happened, the first impression, one halo trait, or the safe middle of the
scale. Part of your value is catching this. Before finalizing, run the
bias check in references/rating-scale.md.
Workflow
Step 1 — Gather what you need
You need these to write a good review. If the manager hasn't provided them, ask
for the missing ones together in a single, friendly request rather than
interrogating one at a time:
- Designer's name and level (Junior / Mid / Senior / Lead / Principal
Product Designer). Level is essential — the same behavior reads as "exceeding"
for a Mid and "meeting" for a Senior. If unknown, ask; don't guess.
- Review period (e.g., "Q1 FY27", "FY26 year-end").
- The overall Performance Rating the manager has decided on the 5-point
scale. If they haven't decided yet, offer to propose one from the evidence and
flag it clearly as a suggestion for them to confirm.
- Their view on the AI Proficiency rating, if they have one (Learning /
Applying / Mastering / Shaping) — optional. If they don't have a view, that's
the default case: you'll assess and propose it yourself in Step 5.
- The source material: any of self-assessment, peer / 360 feedback, the
manager's own notes, and a list of achievements with their business outcomes or
impact. More is better, but work with whatever exists — just be honest in the
draft about where evidence is thin.
- The destination: a Google Drive folder (link or name) where the duplicated
Doc should be saved. Ask for this upfront alongside everything else, not just
at the end — the skill needs it before it can deliver.
If a source is missing, that's fine — note internally which lenses you can
support with real evidence and which you can't, so you don't fabricate.
Step 2 — Read the standards
Read the reference files as needed. They are distilled from OLX's official
sources so you assess against the real bar, not a generic notion of "good":
references/rating-scale.md — the 5-point scale, what each level means, the
rule that levels 4–5 contain all of level 3 plus more, and the bias checklist.
references/values-4cs.md — the 4 C's (Customer-Led, Commitment, Courage,
Collaboration) and how each looks at each performance level.
references/career-blueprint-ic.md — the six competency areas and what they
look like at the designer's level. Use the row for their level as the bar.
This blueprint is IC-only (Junior → Principal PD). If asked to review a design
manager, say plainly that the manager competency bar isn't included yet — you
can still assess values, rating discipline, and AI proficiency, but flag the
missing craft/leadership bar rather than improvising one.
references/ai-proficiency.md — AI proficiency levels and the minimum expected
for the designer's seniority. From Q1 FY27 this is a hard expectation, so a gap
here matters like any other competency gap.
references/output-template.md — the exact structure of the source template
Doc, how to adapt its self-evaluation phrasing for a manager review, and how
to duplicate it correctly.
Step 3 — Analyze the evidence against the standards
Go source by source and map each meaningful data point to the lens it speaks to:
which of the 4 C's, which blueprint competency, AI proficiency. Look for
convergence (several sources pointing at the same strength or gap carry more
weight than a single mention) and for the difference between output (what
shipped) and outcome (what changed for users or the business) — the blueprint
and the values both reward outcomes over activity.
Then hold the picture against the designer's level. A Senior is expected to
be at "Mastering" on the blueprint and to shape strategy and mentor; the same
deliverable that would impress from a Mid may simply be meeting the bar for a
Senior. Calibrating to level is where most of the fairness lives.
Step 4 — Run the bias check
Before writing, deliberately test your emerging assessment against the five
biases in references/rating-scale.md. Ask: am I over-weighting the last month?
A strong first impression? One standout trait bleeding into everything? Am I
drifting to a safe "3" or being lenient to avoid a hard conversation? Adjust the
evidence emphasis so the review reflects the whole period, not a distortion of it.
Step 5 — Write the draft
Fill the template's prompts in the manager's voice, about the designer in the
third person (full detail and exact wording in references/output-template.md):
- 1.A / 1.B — AI tools, impact, and growth. Which AI tools/workflows the
designer actually used and what changed because of it (speed, quality,
decisions), grounded in ownership and "AI buys time for the thinking" rather
than tool name-dropping; then what capability grew this quarter and the
concrete next target.
- 1.C — AI Proficiency rating. Your own proposed rating (Learning / Applying
/ Mastering / Shaping), calibrated against the minimum for the designer's
seniority at this point in the rollout, with a one-line rationale — flagged for
the manager to confirm.
- 2.A — Goals impact. Outcomes and measurable results from the quarter's
goals, not activity. Prioritize outcomes over outputs.
- 2.B — Values alignment. Evidence of the 4 C's in day-to-day behavior, plus
an honest flag of where evidence for any C is thin. Use the blueprint
competencies to calibrate whether this evidence merely meets or genuinely
exceeds the bar for the designer's level, even though the blueprint doesn't
get its own section here.
- 3 — Performance Rating. The manager's decided rating, marked exactly as
one of the five canonical labels, with a short reasoning paragraph beneath it
that argues from the cumulative rule and from what Sections 1–2 already
established — no new claims here that weren't grounded above.
Keep the voice warm, direct, and specific — a manager who respects the person and
is invested in their growth. Avoid inflated praise and avoid vague criticism;
both are useless to the person receiving them. Every prompt's answer should read
as though it could only have been written about this designer.
Carry inline evidence markers through the draft (e.g., a short parenthetical or
footnote pointing to the source) so the manager can verify each claim, and tell
them these are there to be trimmed before the review is shared. Treat peer
feedback as confidential: named peer attributions are for the manager's copy only
and should be anonymized or removed before the review reaches the designer —
default to anonymizing and confirm with the manager.
Step 6 — Deliver the review (duplicate the template Doc, with fallback)
Once the manager is happy with the content, deliver it as a duplicate of the
source template Doc — never as edits to the source file itself, and never as a
freshly-authored Doc with your own structure. Before promising a fully filled-in
Doc, check what the connected tooling can actually do — see the three tiers in
references/output-template.md:
- Full connector (can both duplicate the file and edit a Doc's rich body
content): duplicate, rename, fill in every prompt and rating selection
in-place, and return the link.
- Drive-management-only connector (can copy/rename/move files but has no
tool to write into a Doc's body — this is common; check before assuming
otherwise): still duplicate and rename the file into the specified folder,
say plainly that the content can't be auto-inserted with the available
tooling, and hand the manager the drafted content laid out to match the
duplicated Doc's structure so they can paste it in themselves.
- No connector at all: don't fail — deliver the review as formatted
Markdown mirroring the template's structure, say plainly that no connector
was available so nothing was duplicated, and offer to duplicate the template
once one is connected.
In every tier, if no destination folder was given, ask for one rather than
defaulting silently — don't duplicate into an arbitrary location. Duplicating
(and filling in) the Doc is the preferred finish, but the value of the skill is
the review itself — never block delivery on infrastructure.
When evidence and rating diverge
If the manager has set an overall Performance Rating that the evidence you were
given doesn't clearly support (in either direction), don't paper over it and
don't quietly change it. Name it in one honest sentence to the manager — e.g.,
"The peer feedback and outcomes read closer to a Solid Performer than an
Exceeding to me; happy to write it either way, but wanted to flag it." This
protects the manager from a review that won't survive calibration, which is the
whole point. Then write whatever they decide.
The same applies if the manager states a view on the AI Proficiency rating that
your read of the evidence doesn't support — flag it the same way, then defer to
their call once stated.
What good looks like
A good output is specific (a stranger couldn't have written it), fair
(calibrated to the designer's level, whole-period not recency-driven), honest
(names real gaps developmentally rather than hiding them), and traceable (every
judgment has evidence behind it). The rating and the narrative agree with each
other. The designer, reading it, recognizes themselves and knows exactly what to
do next. And it is delivered: as a correctly-named duplicate of the source
template Doc in the folder the manager specified — never as edits to the source
file itself — with its [Write here] prompts filled in and its rating options
marked (or, if no connector is available, handed over as clean Markdown mirroring
that structure), with peer attributions handled confidentially.
1---2name: performance-review-writer3description: Write a performance review for a member of the OLX product design team. Use this whenever a design manager wants to draft, write, prepare, or put together a performance review, appraisal, evaluation, assessment, or year-end / quarterly / mid-year review for a product designer — even when the request is just "here's their self-assessment and some peer feedback, write it up" or "help me review <name>". Produces an evidence-cited draft in OLX's Q1 FY27 review template (AI Proficiency, Goals status & Values alignment, Performance Rating), anchored to the 4 C's values, the Design Career Blueprint, the 5-point rating scale (with active bias-checking), and AI proficiency expectations, then delivers it as a duplicate of the source template Doc in a Google Drive folder you specify — never writing on top of the source file itself. Reach for this skill any time the task is about assessing or writing up a designer's performance, not just when the word "review" appears. Scoped to individual contributor designers (Junior t4---56# Performance Review Writer (OLX Product Design)78## What this skill is for910A design manager needs to turn a pile of raw material — a self-assessment, peer11or 360 feedback, their own notes, a list of achievements — into a fair,12well-argued performance review for a product designer. The hard part of a review13isn't the prose; it's making a judgment that is grounded in evidence, calibrated14to shared standards, and free of the biases that quietly distort every manager's15memory. This skill exists to do that grounding work reliably, so the manager16gets a draft they can trust and refine rather than a blank page.1718The manager decides the rating. This skill's job is to build the honest,19evidence-based case for that rating (or, if the manager asks, to sanity-check20whether the evidence supports it) and to write it up in the house style.2122## The operating model (read this first)2324Four things define how this skill works. Hold them throughout.25261. **The manager owns the overall rating; you justify it — and you propose the27 AI Proficiency rating for their sign-off.** The manager tells you the overall28 Performance Rating on the 5-point scale (Section 3 of the template). You do29 not invent or override it; your work is to assemble the evidence that30 supports it and express it convincingly. The AI Proficiency rating (Section31 1.C: Learning / Applying / Mastering / Shaping) is different — you assess and32 propose it yourself from the evidence against the minimum for the designer's33 seniority, the same way you'd assess any other competency, then flag it for34 the manager to confirm. Either way, if the evidence you're given seems to35 point somewhere else than a rating the manager has already decided, don't36 silently comply and don't silently overrule — surface the tension plainly and37 let the manager decide (see "When evidence and rating diverge").38392. **Every claim traces to evidence.** This is an *evidence-cited draft*. A40 statement like "raised the quality bar for the team" is worthless without the41 specific thing that shows it. For every strength, gap, or rating argument,42 name the source it came from (a line from the self-assessment, a peer quote, a43 manager note, a shipped outcome). During drafting, keep the evidence visible44 so the manager can verify; the skill produces a version with inline evidence45 markers the manager can later trim.46473. **Frameworks shape judgment, not vocabulary.** The 4 C's, the blueprint48 competencies, and the AI levels are the lens you assess through — they decide49 *what* counts as strong or weak. But the prose stays natural and human. Don't50 pepper the review with jargon like "demonstrates Mastering-level Agentic51 Experience Craft." Say what the person actually did and why it mattered. The52 framework should be felt in the accuracy of the judgment, not seen in the53 wording.54554. **Actively fight bias.** Human reviews drift toward the most recent thing that56 happened, the first impression, one halo trait, or the safe middle of the57 scale. Part of your value is catching this. Before finalizing, run the58 bias check in `references/rating-scale.md`.5960## Workflow6162### Step 1 — Gather what you need6364You need these to write a good review. If the manager hasn't provided them, ask65for the missing ones together in a single, friendly request rather than66interrogating one at a time:6768- **Designer's name** and **level** (Junior / Mid / Senior / Lead / Principal69 Product Designer). Level is essential — the same behavior reads as "exceeding"70 for a Mid and "meeting" for a Senior. If unknown, ask; don't guess.71- **Review period** (e.g., "Q1 FY27", "FY26 year-end").72- **The overall Performance Rating the manager has decided** on the 5-point73 scale. If they haven't decided yet, offer to propose one from the evidence and74 flag it clearly as a suggestion for them to confirm.75- **Their view on the AI Proficiency rating, if they have one** (Learning /76 Applying / Mastering / Shaping) — optional. If they don't have a view, that's77 the default case: you'll assess and propose it yourself in Step 5.78- **The source material**: any of self-assessment, peer / 360 feedback, the79 manager's own notes, and a list of achievements with their business outcomes or80 impact. More is better, but work with whatever exists — just be honest in the81 draft about where evidence is thin.82- **The destination**: a Google Drive folder (link or name) where the duplicated83 Doc should be saved. Ask for this upfront alongside everything else, not just84 at the end — the skill needs it before it can deliver.8586If a source is missing, that's fine — note internally which lenses you can87support with real evidence and which you can't, so you don't fabricate.8889### Step 2 — Read the standards9091Read the reference files as needed. They are distilled from OLX's official92sources so you assess against the real bar, not a generic notion of "good":9394- `references/rating-scale.md` — the 5-point scale, what each level means, the95 rule that levels 4–5 contain *all* of level 3 plus more, and the bias checklist.96- `references/values-4cs.md` — the 4 C's (Customer-Led, Commitment, Courage,97 Collaboration) and how each looks at each performance level.98- `references/career-blueprint-ic.md` — the six competency areas and what they99 look like at the designer's level. Use the row for *their* level as the bar.100 This blueprint is IC-only (Junior → Principal PD). If asked to review a design101 *manager*, say plainly that the manager competency bar isn't included yet — you102 can still assess values, rating discipline, and AI proficiency, but flag the103 missing craft/leadership bar rather than improvising one.104- `references/ai-proficiency.md` — AI proficiency levels and the minimum expected105 for the designer's seniority. From Q1 FY27 this is a hard expectation, so a gap106 here matters like any other competency gap.107- `references/output-template.md` — the exact structure of the source template108 Doc, how to adapt its self-evaluation phrasing for a manager review, and how109 to duplicate it correctly.110111### Step 3 — Analyze the evidence against the standards112113Go source by source and map each meaningful data point to the lens it speaks to:114which of the 4 C's, which blueprint competency, AI proficiency. Look for115convergence (several sources pointing at the same strength or gap carry more116weight than a single mention) and for the difference between *output* (what117shipped) and *outcome* (what changed for users or the business) — the blueprint118and the values both reward outcomes over activity.119120Then hold the picture against the designer's **level**. A Senior is expected to121be at "Mastering" on the blueprint and to shape strategy and mentor; the same122deliverable that would impress from a Mid may simply be meeting the bar for a123Senior. Calibrating to level is where most of the fairness lives.124125### Step 4 — Run the bias check126127Before writing, deliberately test your emerging assessment against the five128biases in `references/rating-scale.md`. Ask: am I over-weighting the last month?129A strong first impression? One standout trait bleeding into everything? Am I130drifting to a safe "3" or being lenient to avoid a hard conversation? Adjust the131evidence emphasis so the review reflects the whole period, not a distortion of it.132133### Step 5 — Write the draft134135Fill the template's prompts in the manager's voice, about the designer in the136third person (full detail and exact wording in `references/output-template.md`):137138- **1.A / 1.B — AI tools, impact, and growth.** Which AI tools/workflows the139 designer actually used and what changed because of it (speed, quality,140 decisions), grounded in ownership and "AI buys time for the thinking" rather141 than tool name-dropping; then what capability grew this quarter and the142 concrete next target.143- **1.C — AI Proficiency rating.** Your own proposed rating (Learning / Applying144 / Mastering / Shaping), calibrated against the minimum for the designer's145 seniority at this point in the rollout, with a one-line rationale — flagged for146 the manager to confirm.147- **2.A — Goals impact.** Outcomes and measurable results from the quarter's148 goals, not activity. Prioritize outcomes over outputs.149- **2.B — Values alignment.** Evidence of the 4 C's in day-to-day behavior, plus150 an honest flag of where evidence for any C is thin. Use the blueprint151 competencies to calibrate whether this evidence merely meets or genuinely152 exceeds the bar for the designer's level, even though the blueprint doesn't153 get its own section here.154- **3 — Performance Rating.** The manager's decided rating, marked exactly as155 one of the five canonical labels, with a short reasoning paragraph beneath it156 that argues from the cumulative rule and from what Sections 1–2 already157 established — no new claims here that weren't grounded above.158159Keep the voice warm, direct, and specific — a manager who respects the person and160is invested in their growth. Avoid inflated praise and avoid vague criticism;161both are useless to the person receiving them. Every prompt's answer should read162as though it could only have been written about *this* designer.163164Carry inline evidence markers through the draft (e.g., a short parenthetical or165footnote pointing to the source) so the manager can verify each claim, and tell166them these are there to be trimmed before the review is shared. Treat peer167feedback as confidential: named peer attributions are for the manager's copy only168and should be anonymized or removed before the review reaches the designer —169default to anonymizing and confirm with the manager.170171### Step 6 — Deliver the review (duplicate the template Doc, with fallback)172173Once the manager is happy with the content, deliver it as a **duplicate** of the174source template Doc — never as edits to the source file itself, and never as a175freshly-authored Doc with your own structure. Before promising a fully filled-in176Doc, check what the connected tooling can actually do — see the three tiers in177`references/output-template.md`:1781791. **Full connector** (can both duplicate the file and edit a Doc's rich body180 content): duplicate, rename, fill in every prompt and rating selection181 in-place, and return the link.1822. **Drive-management-only connector** (can copy/rename/move files but has no183 tool to write into a Doc's body — this is common; check before assuming184 otherwise): still duplicate and rename the file into the specified folder,185 say plainly that the content can't be auto-inserted with the available186 tooling, and hand the manager the drafted content laid out to match the187 duplicated Doc's structure so they can paste it in themselves.1883. **No connector at all:** don't fail — deliver the review as formatted189 Markdown mirroring the template's structure, say plainly that no connector190 was available so nothing was duplicated, and offer to duplicate the template191 once one is connected.192193In every tier, if no destination folder was given, ask for one rather than194defaulting silently — don't duplicate into an arbitrary location. Duplicating195(and filling in) the Doc is the preferred finish, but the value of the skill is196the review itself — never block delivery on infrastructure.197198## When evidence and rating diverge199200If the manager has set an overall Performance Rating that the evidence you were201given doesn't clearly support (in either direction), don't paper over it and202don't quietly change it. Name it in one honest sentence to the manager — e.g.,203"The peer feedback and outcomes read closer to a Solid Performer than an204Exceeding to me; happy to write it either way, but wanted to flag it." This205protects the manager from a review that won't survive calibration, which is the206whole point. Then write whatever they decide.207208The same applies if the manager states a view on the AI Proficiency rating that209your read of the evidence doesn't support — flag it the same way, then defer to210their call once stated.211212## What good looks like213214A good output is specific (a stranger couldn't have written it), fair215(calibrated to the designer's level, whole-period not recency-driven), honest216(names real gaps developmentally rather than hiding them), and traceable (every217judgment has evidence behind it). The rating and the narrative agree with each218other. The designer, reading it, recognizes themselves and knows exactly what to219do next. And it is *delivered*: as a correctly-named **duplicate** of the source220template Doc in the folder the manager specified — never as edits to the source221file itself — with its `[Write here]` prompts filled in and its rating options222marked (or, if no connector is available, handed over as clean Markdown mirroring223that structure), with peer attributions handled confidentially.