Writing the defect class list your inspection will be graded against
In short
A defect classification taxonomy for visual inspection is a contract between quality and engineering: every class carries a name, a one-line definition, a severity, a disposition, a minimum reportable size and a counter-example. Classes two experienced inspectors cannot separate by eye will not be separated by a model either — they will be averaged.
Key takeaways
- A class is not a word. It is 6 fields: name, definition, severity, disposition, minimum reportable size, and an example plus a counter-example.
- Split any class whose members take different dispositions. One "scratch" covering both a cosmetic mark and a through-coating gouge is 2 classes wearing 1 label.
- State minimum size in part units — millimetres on the surface — never in pixels, which change with every lens and standoff.
- Keep an unclassified bucket and review it weekly. A taxonomy with nowhere to put a surprise silently forces surprises into the nearest wrong class.
The defect class list is the contract between quality and engineering, and must exist before any images are collected. Each class needs a name, a definition, a severity, a disposition, a minimum reportable size, and a pair of images — one that is the class, one that deliberately is not. Anything less is a word, and a word cannot be scored.
The failure this prevents is specific. "Scratch" is the usual offender: it covers a light mark in the coating that ships without comment and a gouge through to the substrate that is scrapped. Two dispositions, one label. The model learns an average of the two, and the confusion matrix reports accuracy on a category that never existed.
The six fields a class has to carry
- Name and one-line definition, written so a new inspector reaches the same verdict on day 1. "Coating void: substrate visible through the coating" leaves less room than "bad finish".
- Severity. Critical, major or minor against function or the customer requirement — not against how obvious it is to look at.
- Disposition. Scrap, rework, use-as-is, quarantine. This is the field the line acts on, and every class needs exactly one.
- Minimum reportable size, in part units. "Length ≥ 5 mm on the show face" is checkable; "visible" depends on the light. Never pixels, which change with lens and standoff.
- Where it can occur — named faces or zones. A mark that is minor on a hidden face and major on the show face is 2 rows, not a footnote.
- Example and counter-example. The counter-example is the one everybody skips and the one that does the work: it fixes the boundary against the class next door.
A worked list for one part family
Below is the shape for a powder-coated steel bracket with one show face. The numbers are illustrative — yours come from the drawing — but the columns are the point. Any class you cannot fill in across all four is not yet a class.
| Class | Definition | Minimum reportable | Severity / disposition |
|---|---|---|---|
| Coating void | Substrate visible through the coating | Any extent | Critical / scrap — corrosion path |
| Gouge | Mechanical mark exposing substrate | Any extent | Critical / scrap |
| Handling scratch | Mark in the coating, substrate not exposed | ≥ 5 mm on show face | Minor / use-as-is off the show face |
| Run or sag | Coating flow ridge with a raised edge | Height ≥ 0.3 mm | Major / rework and re-coat |
| Orange peel | Uniform texture, no substrate exposed | Over 25% of the show face | Minor / use-as-is unless customer-facing |
| Inclusion | Particle in or under the coating, raising the surface | ≥ 1 mm | Major / rework if raised only |
Four rules that keep the list usable a year later
- Mutually exclusive by construction. If a part can honestly belong to 2 classes, write the precedence rule into the definitions rather than leaving it to whoever holds the part.
- One disposition per class, no conditions hidden in prose. A class whose disposition depends on a second judgement is 2 classes.
- An explicit unclassified bucket, reviewed on a fixed cadence. Without it every surprise is pushed into the nearest wrong class.
- Version the list. Adding a class after a model exists invalidates the earlier grading, so the change needs a date, an owner and a note on what must be re-scored.
If two experienced inspectors cannot put the same part in the same class, no amount of training data will make a model do it. It will just be confidently inconsistent instead of visibly inconsistent.
What the list decides downstream
The disposition column reaches the floor. Classes dispositioned use-as-is should never be wired to a reject actuator, however confidently detected, because each one is a good part pulled off the belt and a minute of somebody's re-check. Severity runs the other way: it names the classes an escape is unacceptable for.
The minimum-size column sorts the technology. A class defined by a number a caliper could confirm belongs with classical tools, not a learned stage — the split is argued in the measurement your caliper agrees with. What the list does not do is set sensitivity: where the cut falls stays a governed decision, described in the anomaly score is a ranking, the threshold a business decision.
The list is also what acceptance is measured against: per-class recall targets and overkill budgets can only be written once the classes exist, the argument in an exit criterion is not a good feeling. That is why the taxonomy is a scoping deliverable in MVP and product builds, alongside visual inspection and defect detection and the wider industrial vision practice.
Frequently asked questions
Short answers to the follow-ups this page tends to raise.
How many defect classes should a first inspection system have?
Fewer than the plant currently names, and each one mapping to a distinct disposition. Most reject logs carry 20 or 30 labels accumulated over years, of which perhaps 6 to 8 drive different actions. Start with the classes that change what happens to the part, keep the rest as free-text notes, and promote a note to a class when it earns its own disposition.
What is the difference between a cosmetic and a functional defect?
A functional defect impairs how the part performs or how long it lasts; a cosmetic defect affects only appearance. The distinction decides who owns the accept boundary: functional limits come from engineering and the drawing, cosmetic limits from the customer specification, and are usually zone-dependent. Mark which each class is — the two get argued in different rooms.
Should the defect class list include classes we have never seen?
Include them only with a defined disposition and a source — a customer complaint, a known process failure mode, a supplier's history. A class with no examples cannot be trained or graded, so mark it anticipated and route it to the unclassified bucket until real instances arrive. Training on a class nobody has photographed produces a detector for an idea, not a defect.
Who signs off the defect class list?
Quality signs it, because quality owns severity and disposition. Engineering confirms each class is physically detectable as defined, and production confirms the dispositions can be executed — a rework disposition with no rework station is scrap in disguise. Where a customer specification governs appearance, the cosmetic classes need their agreement in writing.
- defect taxonomy
- quality standards
- inspection design
- severity
The work behind this page
Builds from our portfolio that this page draws on.
Open Vision PPE Monitoring
Boundary surveillance, PPE compliance monitoring, and intrusion detection via real-time video analytics. Runs fully on-premise — no cloud required.
Safety & ComplianceFactory OS
Production planning and task management for a tier-1 apparel manufacturer — replacing Excel with automated milestone planning, SOP gate enforcement, and real-time visibility.
ManufacturingRead next
- The golden sample is a decision record, not just a good partThe part in the drawer is the easy half. Its record — revision, attribute, approver, date, recheck trigger — is what an automated inspection inherits.definition
- The anomaly score is a ranking; the threshold is a business decisionThe score orders parts by how far they sit from normal. It carries no units, no probability and no severity — which is why the threshold belongs to quality, not to engineering.definition
- False reject rate: what the line feels, not the accuracy on the slideA 1% false reject rate sounds like rounding. At 1,200 parts an hour it is 96 good parts a shift and over an hour of somebody re-checking them.definition
Working on something in this space?
Tell us where you are in a sentence or two. We'll tell you honestly whether we're the right team, and what a sensible first slice of the work looks like.
Start the conversation