Subjectivity in grading is not evidence that grading is meaningless. It is the unavoidable consequence of converting a complex physical object into a simplified condition judgement. A collectible does not naturally possess a number such as 8.5, MS-65 or a phrase such as Near Mint. It possesses observable characteristics: wear, scratches, fading, centring, tears, stains, missing components, manufacturing variation, restoration and many other features. A grading system decides which of those characteristics matter, how heavily they should count and where the final condition falls within a named band.
That translation can be disciplined, repeatable and evidence-led without becoming perfectly objective. The practical challenge is therefore not to eliminate judgement, but to make it bounded: informed by the correct standard, grounded in observable evidence, transparent enough to explain, and open to review when the evidence changes.
Collector scenario
One object, two defensible grades
A seller describes a boxed figure as Excellent. The figure itself is bright, complete and lightly handled. The box presents well from the front, but one flap has compression, the window is clouded and a small accessory is a period-correct replacement rather than the original supplied with that example.
The seller treats the defects as qualifications around an otherwise strong display piece. The buyer treats the replaced accessory and window condition as grade-controlling. Both may agree on every visible fact. Their dispute begins only when those facts are weighted and translated into a single label.
The useful question is not simply, ‘Who is stricter?’ It is: which component is being graded, which standard applies, what feature controls the boundary, and whether condition, completeness, originality and restoration are being collapsed into one claim.
Grading as bounded judgement
The most useful model is not objectivity versus opinion. It is bounded subjectivity: judgement operating inside a framework of observable evidence, published definitions, specialist knowledge, comparison examples and recognised conventions.
Core principle
Subjective does not mean arbitrary
A subjective judgement is one that still requires interpretation. An arbitrary judgement is one that cannot be defended by a recognisable standard, relevant evidence or consistent reasoning.
Two competent graders may reach adjacent grades while still working responsibly. A grade becomes weak when the grader cannot identify what was observed, why it mattered, which standard was applied, what uncertainty remains or what feature controlled the boundary.
Collector test
Ask whether the judgement can show its working
A reasoned grade should be traceable backwards from the label to the object. The collector should be able to identify the significant observations, the interpretation placed on them and the rule or convention that turns them into a grade.
What was physically observed?
Was each feature classified correctly?
How severe, extensive or visually dominant is it?
How does the applicable standard weight it?
Does it impose a grade cap or merely influence position within a band?
How confident is the conclusion?
Separate observation from judgement
Many grading arguments remain circular because the parties move directly from a label to a counter-label. Separating the physical observation from its interpretation reveals whether the dispute concerns evidence, classification or weighting.
Observable fact
What is physically present
Observation describes features that can often be photographed, measured or independently checked. Examples include three colour-breaking spine ticks, a 12 mm edge tear, whitening on two corners, a missing insert, a shallow dent, a repaired split or approximately 60/40 centring.
Observation is not always easy. Faint restoration, cleaning, trimming, recolouring, pressing or concealed damage may require magnification, raking light, ultraviolet examination, measurement or comparison with known examples.
Interpretation
What the observation means
Judgement begins when the grader decides whether the feature is slight or substantial, whether it is manufacturing variation or later damage, whether it caps the grade, how it combines with other defects, or whether strong presentation places the object at the top of the permitted band.
The first layer can be highly factual. The second is interpretive. A credible system keeps those layers distinct enough that disagreement can be located rather than hidden inside the final number.
Why subjectivity cannot be completely removed
Irregular deterioration
Real objects do not match textbook examples
One trading card may have sharp corners, poor centring and a faint surface line. Another may have ideal centring, mild edge whitening and weaker print focus. Both may fall near the same grade for different reasons. No natural formula proves that one scratch equals two softened corners or that a centring problem must outweigh a small stain.
Continuous condition
Grades divide a continuum into artificial bands
There is no physical cliff between 8 and 8.5, Fine and Very Fine, or MS-64 and MS-65. Objects near a boundary may reasonably move between adjacent grades when examined by different competent graders or under different conditions.
Holistic qualities
Some qualities are real but not mechanically measurable
Eye appeal, freshness, balance, originality, sharpness and visual disturbance affect how an object presents. These qualities can be analysed through colour, gloss, focal distractions, uniformity and overall coherence, but their final weighting remains partly interpretive.
Production overlap
Manufacturing quality can resemble later damage
Factory rough cuts may resemble edge damage; a weak strike may resemble circulation wear; production creases may resemble handling creases; moulding lines may resemble scratches. Correct grading therefore depends on knowledge of how the particular object was made, not merely on noticing a mark.
Judgement language
Standards still rely on interpretive terms
Words such as slight, moderate, distracting, outstanding, acceptable and virtually perfect narrow judgement without automating it. A standard can constrain the answer, but it cannot prescribe every possible combination of defects.
Multiple components
A single label may conceal different condition profiles
A comic has covers, pages, staples and inserts. A boxed toy has the figure, paint, accessories, box, window and seals. A video game may have a case, seal, cartridge or disc, manual and inserts. The more components involved, the less informative a single composite number becomes.
Where judgement enters the grading process
Subjectivity does not enter at one single moment. It appears at several linked stages. Locating the stage matters because each kind of disagreement requires a different form of evidence or review.
01
Detection
The grader decides whether a feature is present at all. Subtle trimming, cleaning, retouching, pressing, reglossing or concealed cracks may be missed or disputed.
02
Classification
The feature is classified as wear, production variation, damage, restoration, conservation, active deterioration or an original manufacturing characteristic.
03
Severity
The grader decides whether the feature is slight, moderate, extensive, localised, widespread, structurally important or visually dominant.
04
Weighting
Unlike defects are balanced. The grader decides whether one major defect controls the result, whether several minor defects accumulate, and how presentation influences the permitted range.
05
Grade caps
Some defects create an effective ceiling: a crease, missing page, cleaning, restoration, broken seal or incomplete component may prevent the higher grade regardless of other strengths.
06
Placement within the band
Even broad agreement may leave a final choice between a low-end 8, solid 8, high-end 8, 8.5 or low-end 9. That narrow decision can have a disproportionate market effect.
Diagnose the disagreement before defending the grade
Most disputes are narrower than the final label makes them appear. Naming the dispute type prevents a conversation about condition from becoming a personal argument about generosity, strictness or expertise.
Type 1
Observation disagreement
The parties do not agree on what is physically present. One sees a manufacturing mark; another sees later damage. One sees original toning; another sees artificial treatment. This dispute requires better inspection, images, measurements, references or specialist examination before the grade can be settled.
Type 2
Classification disagreement
Both parties see the feature, but one calls it wear while the other calls it restoration, alteration or a production flaw. Classification disputes can be more consequential than an adjacent numerical difference because they may change the label category entirely.
Type 3
Weighting disagreement
The facts are accepted, but their importance is not. A replaced part, stain, missing insert, softened edge or area of retouching may be treated as minor by one participant and grade-controlling by another.
Type 4
Standard disagreement
The participants are applying different scales, companies, eras or market conventions. A specialist standard may be tighter than general marketplace language. An 8 from one company is not automatically equivalent to an 8 from another.
Type 5
Confidence disagreement
The apparent condition may look strong, but one party is less confident because hidden areas, internal components, restoration history or photographic evidence remain incomplete. The disagreement concerns how firmly the grade can be claimed.
Type 6
Value expectation disagreement
A buyer may challenge the grade because the price feels too high, while a seller may defend the grade because a lower label affects value. Grade evidence and price emotion should be separated before the condition dispute is assessed.
Technical grading, market grading and eye appeal
A major source of conflict is the hidden use of different grading philosophies. One participant may expect a condition-led technical assessment while another applies the conventions of the specialist market, including originality and eye appeal.
Technical emphasis
Condition characteristics lead the judgement
Technical grading concentrates on wear, marks, completeness, structural damage, surface preservation and alterations. It attempts to describe the physical state without allowing desirability to dominate the result.
Market emphasis
Field conventions shape how defects are weighted
Market grading also considers eye appeal, originality, colour, strike, typical production quality and the way experienced buyers place the object. This does not mean assigning a grade from price. It means the market's grading tradition has evolved around preferences that are not reducible to defect counting alone.
Dispute risk
Conflict arises when the two approaches are hidden
A collector may insist that an object has fewer scratches and therefore deserves the higher grade. Another may treat a single central stain as more visually significant than several peripheral marks. Both are assessing condition, but their weighting principles differ.
Eye appeal is not a free override
Translate overall impression into identifiable elements
Eye appeal can reflect visual balance, colour strength, gloss or lustre, centring, absence of focal distractions, uniformity of wear and quality of manufacture. Those factors are real, but a robust grade should reflect established field conventions rather than one grader's personal taste.
Strong presentation may place an object at the upper end of an allowed band. It should not quietly erase a defect that the standard treats as a cap, an alteration or a separate disclosure requirement.
Subjectivity changes by collectible type
No universal defect hierarchy works across every collecting field. Each object type has its own manufacturing history, vulnerable areas, component structure and market conventions.
Coins
Wear, strike, lustre and surfaces interact
Weak strike may resemble wear. Toning may be attractive, neutral or distracting. Cleaning can change the label category even where detail remains strong. Details-graded coins also show why technically similar problems can produce very different buyer preferences.
Trading cards
Centring, corners, edges and surfaces compete
A card may be extremely clean but off-centre, or perfectly centred with a small surface defect. Factory print quality, focus, gloss, trimming and recolouring add further classification questions.
Comics and paper
Multi-component structures produce mixed condition
Spine stress, tears, page quality, staple condition, cover attachment, staining, missing pieces and restoration may affect different parts of the object. A headline number may conceal the feature that actually controlled the grade.
Toys and figures
Production variation and completeness are central
Factory paint, stress marks, yellowing, articulation wear, packaging seals, window clarity, reproduction accessories and replacement parts must be separated from ordinary later damage.
Video games
The component set may matter as much as the headline grade
Box, seal, case, media, manual, inserts, hang tabs and production variants can each carry different condition and authenticity questions. A composite score may hide which component caused the reduction.
Autographs and memorabilia
Authentication and condition must not be merged
A genuine autograph can have weak condition because of fading, skipping, smearing or poor placement. Game-used material may also involve attribution confidence, which is an evidential judgement rather than a straightforward condition score.
Professional grading reduces variation; it does not abolish judgement
Third-party grading can improve consistency by using published standards, trained specialists, controlled examination, multiple graders, internal verification, reference sets, standardised labels and formal review routes. The outcome remains an institutional opinion rather than a direct physical measurement.
Inter-grader variation occurs when different graders weight or detect features differently. Intra-grader variation occurs when the same grader reaches another conclusion later under different lighting, with more information, after standards evolve or because the object is genuinely borderline. Repeatability should therefore be judged as a narrow defensible range, not as a promise that every examination will produce an identical point score.
Normal grading noise
Adjacent grades on a borderline object
A marginal object moves between 7 and 7.5, 8 and 8.5, or two neighbouring descriptive bands. The object lies close enough to the boundary that a narrow difference in weighting or viewing conditions changes the final label.
Material inconsistency
Widely separated outcomes without explanation
An unchanged object moves from 6 to 9, Fine to Near Mint, unaltered to restored, or complete to materially incomplete. Such a discrepancy deserves investigation rather than being dismissed as ordinary subjectivity.
Detection difference
A later grader finds a previously missed feature
Cleaning, trimming, restoration, recolouring, replaced parts, resealing or concealed damage may alter the factual basis of the assessment. This is not merely boundary noise.
Policy difference
The object is unchanged but the rules are not
Companies may treat centring, factory defects, restoration, pressing, writing, missing components or qualifiers differently. Standards may also evolve over time. The same number can therefore carry different meanings across services and eras.
Commercial pressure
Why a small grade difference becomes a large dispute
Grade 8 — $500
Grade 9 — $2,000
A one-point difference is no longer a minor descriptive disagreement when it changes the market outcome by $1,500. Price cliffs encourage repeated resubmission, selective disclosure of earlier grades, speculative buying of borderline examples and pressure to treat the exact label as objective truth.
The financial difference magnifies the dispute; it does not make the grading boundary more scientifically certain.
The slab or certificate effect
Once an object is encapsulated or certified, the label can become psychologically dominant. Collectors begin to compare labels rather than objects, even though two items with the same grade may have entirely different strengths and weaknesses.
Assuming all objects with the same grade are equally attractive.
Paying for the label without inspecting the underlying defect profile.
Ignoring grader notes, qualifiers, restoration status or completeness.
Assuming one company's number maps directly to another's.
Treating the grade as permanent despite later damage, conservation or review.
Believing certification guarantees future value rather than recording an opinion at a point in time.
A holder stabilises and communicates an opinion. It does not abolish the physical object beneath it.
Resubmission, regrading and claims of inconsistency
A higher grade after resubmission proves that a later assessment differed. It does not by itself prove why. The first grade may have been low, the second high, the object may be borderline, conservation may have occurred, a defect may have been missed, standards may have changed or unsuccessful attempts may have been omitted from the story.
Before treating an upgrade as evidence of arbitrary grading, document whether the same company and service were used, whether the object remained in-holder, whether it was conserved or altered, how many submissions occurred, and whether the complete history is available. Successful upgrades are publicised more often than failed attempts, producing survivorship bias.
Better ways to represent uncertainty
A transparent collector record can acknowledge uncertainty without making the grade useless. The aim is to preserve what is known, what is disputed and what might change the conclusion.
Grade range
Use a range where the evidence does not support a single point
An informal record such as ‘likely 7.5–8.5, with centring the main uncertainty’ is often more honest than an unsupported exact 8.
Boundary notation
State how the object sits within the band
‘Strong 8; possible low 9 under a centring-tolerant standard’ communicates both the judgement and the reason another competent grader may differ.
Confidence
Separate the apparent grade from certainty
Record high, moderate or low confidence, particularly when images are incomplete, hidden areas remain unexamined or suspected alteration cannot be resolved.
Defect-led explanation
Name the feature that controls the result
‘Sharp corners and clean surface, but a faint crease imposes a substantial cap’ is more durable than the number alone.
Separate condition axes
Record components independently where a composite score hides meaning
For a boxed figure: figure 8.5; paint 8; accessories complete; box 6.5; window 7; seal intact; restoration none observed. This preserves the actual profile behind the overall impression.
A sensible grading-dispute hierarchy
Begin with identity and evidence, not with the desired outcome. This order prevents a disagreement about one feature from expanding into a general argument about trust, price or expertise.
01
Confirm identity
Make sure the parties are discussing the same object, variant, holder and component configuration.
Is the certification number correct?
Has the object changed holders or components?
Are before-and-after images genuinely of the same example?
02
Confirm the applicable standard
Identify the company, guide, scale, service category and grading era. Do not compare numbers before confirming what each number means.
03
Identify the disputed observation
Name the precise feature: centring, a crease, restoration, cleaning, a missing insert, seal integrity, a production mark or another specific point.
04
Separate fact from interpretation
Decide whether the dispute is about the existence or classification of the feature, or only about how heavily it should count.
05
Gather relevant evidence
Use evidence that addresses the disputed point rather than overwhelming the discussion with unrelated material.
06
Assess materiality
Ask whether resolving the issue would realistically change the grade, label category, restoration designation, completeness status or value.
07
Use the proportionate remedy
Seek an explanation, grader notes, return, review, regrade, label correction, authenticity review, guarantee claim or independent specialist opinion as appropriate.
08
Record the outcome
Preserve the evidence, reasoning, prior labels and final resolution so that future owners do not have to reconstruct the dispute from memory.
Build the evidence around the disputed point
Direct condition evidence
Show the object clearly
High-resolution images, raking-light photographs, measurements, magnified details, unpacking images and pre-submission photographs can establish what was present and when.
Institutional evidence
Use the grading record
Certification data, grader notes, previous labels, review correspondence and guarantee terms help identify the standard and the reasoning available from the service.
Historical evidence
Preserve intervention and ownership history
Conservation reports, restoration invoices, auction descriptions, provenance records and shipping documentation may explain why assessments differ over time.
Comparative evidence
Choose genuinely comparable examples
The strongest comparisons use the same issue, material, grading company, standard and era, with images that reveal the relevant defect. Comparison objects are precedents, not binding authorities.
Evidence-led challenge
Weak
“This looks much better than an 8.”
Stronger
“The published 8 standard permits slight corner wear, but no corner wear is visible in the supplied images. The apparent upper-edge whitening may be holder reflection. Please review the edge and surface assessment and confirm whether another feature controlled the grade.”
A useful challenge should identify:
The disputed feature
The relevant standard
The evidence relied upon
The claimed error or inconsistency
The remedy requested
Practical rules for sellers, buyers and collection records
For sellers
Make the grading claim inspectable
Name the grading system, distinguish a personal estimate from a certified grade, disclose visible defects, describe completeness separately and identify restoration, repair or alteration. Use ranges where confidence is limited and retain pre-shipping condition records.
Avoid vague shorthand such as ‘minty’ or ‘investment grade’.
Photograph known problem areas rather than only flattering angles.
Do not imply that a self-grade guarantees a third-party result.
State what has not been professionally examined.
For buyers
Buy the object rather than the label
Read the actual standard, inspect the defect profile, request missing views and select examples whose weaknesses you can tolerate. Distinguish condition from authenticity, restoration, completeness and value.
Check certification details and grader notes.
Understand the seller's return policy before purchase.
Budget for uncertainty rather than paying solely for an upgrade possibility.
Remember that the numerically highest example may not be the most desirable to you.
For collection records
Preserve the reasoning behind the grade
A useful record stores the grading authority, scale, grade, certification number, date, qualifiers, restoration status, completeness, notes, visible defects, confidence, photographs, prior certification history and conservation or repair history.
The record should make it possible to understand how the assessment evolved and whether the object itself changed between grades.
Myth versus reality
Myth
Professional grading is objective.
Reality
Professional grading is standardised expert judgement. It aims for impartiality and consistency, but interpretation remains involved.
Myth
Subjective means random.
Reality
Subjective assessments can be tightly constrained by evidence, standards, reference material and review.
Myth
Two different grades prove corruption.
Reality
They may reflect borderline condition, different standards, missed features, changed condition or ordinary adjacent-grade variation.
Myth
The highest number is always the best purchase.
Reality
Defect profile, originality, restoration, completeness and eye appeal may matter more to an individual collector than the headline grade.
Myth
Published standards remove judgement.
Reality
Standards define what to consider and constrain the range of acceptable outcomes; they cannot prescribe every combination of features.
Myth
Automated grading will be completely objective.
Reality
Automation may improve measurement and repeatability, but people still choose the training examples, feature weightings, boundaries and override rules.
When subjectivity becomes unacceptable
Judgement is unavoidable. Unconstrained judgement is not. The system should be clear enough that another competent grader can understand how the conclusion was reached and identify the point of disagreement.
No published or explainable standard.
No clarity about what was inspected.
Dramatically different grades without review or explanation.
Inconsistent treatment of the same material defect.
Refusal to correct factual or mechanical label errors.
Undisclosed conflicts of interest or guaranteed outcomes before examination.
Alteration claims without supporting detail.
Precision claimed from inadequate photographs.
Composite grades with no defined component weighting.
Marketing language that presents an expert opinion as scientific certainty.
When specialist review is warranted
Specialist review should be proportionate to the stakes and ideally independent of the transaction. It is most useful where the factual basis is contested, the discrepancy is large or further collector-led inspection would be unreliable or damaging.
The value difference between the disputed outcomes is substantial.
Restoration, alteration, trimming, cleaning or recolouring is suspected.
Authenticity and condition questions overlap.
A rare manufacturing characteristic may have been misclassified as damage.
The object is fragile and further inspection could cause harm.
Prior assessments differ widely rather than by an adjacent grade.
Insurance, litigation, a high-value private sale or a guarantee claim depends on the result.
Images are insufficient and the disputed feature requires in-person examination.
A seller alleges that damage occurred after delivery.
Final collector principle
A trustworthy grade communicates more than a number
Mature collectors can hold several ideas at once: grades are useful; grades are not absolute; professional opinions deserve weight; professional opinions can still be wrong; adjacent-grade variation may be normal; and wide discrepancies deserve investigation.
The best grading systems do not pretend subjectivity has disappeared. They make it informed, bounded, repeatable, transparent, reviewable and independent of improper influence. The underlying object, its evidence and its documented defect profile remain more durable than any unsupported headline label.
Key takeaways
A grade is a structured opinion about condition, not an intrinsic measurement carried by the object.
Subjectivity becomes defensible when it is bounded by evidence, standards, expertise and transparent reasoning.
Resolve disputed facts and classifications before arguing about the final number.
Adjacent-grade variation may be normal; wide unexplained discrepancies require investigation.
Condition, completeness, originality, restoration and value should not be collapsed into one label.
The underlying object and its documented defect profile remain more durable than any single grade.