Expert Disagreement

Expert disagreement is not an abnormal failure at the edge of grading. It is a predictable consequence of asking knowledgeable people to convert a complex physical object into a simplified condition category, designation or number. Two competent graders may recognise substantially the same defects and still disagree about severity, weighting, production origin, alteration, eligibility or the precise side of a grade boundary on which the object sits.

The mature collector does not ask which expert supplied the most desirable answer. The better question is which conclusion best explains the physical evidence, under a clearly stated standard, with the fewest unsupported assumptions. Not every disagreement is equally reasonable: some reflect legitimate boundary judgement, while others expose incomplete examination, inconsistent standards, conflicts of interest, poor documentation or error.

Collector scenario

Two careful experts disagree by half a grade

Both experts inspect the same collectible in hand. They agree that the colour is strong, the surface is original, the centring is acceptable and one corner shows light wear. One assigns 8.0 because the corner wear is visible under normal viewing. The other assigns 8.5 because the wear is confined and offset by unusually strong surfaces.

This is not a dispute about what the object is or whether the defect exists. It is a weighting and boundary disagreement. A responsible record may therefore state that the item occupies a defensible expert range of approximately 8.0–8.5 rather than pretending that one number is a complete physical truth.

Domain boundary

Expert disagreement is not the same as every neighbouring grading problem

This chapter focuses on how collectors compare and act upon conflicting informed opinions. The existence of judgement belongs more broadly to subjectivity; uncertainty about available information belongs to evidence gaps and grade confidence; a challenge to genuineness belongs primarily to authentication; and disputes over intervention may require restoration or conservation expertise.

Where authenticity, restoration or material stability controls the dispute, resolve that underlying question before treating the issue as a simple contest between numerical grades.

What expert disagreement actually means

A grading disagreement becomes useful only after the collector identifies its type. A difference in final number may be less significant than a conflict about restoration, completeness, attribution or the adequacy of the examination itself.

Boundary judgement

Grade placement

The experts recognise broadly the same condition but place the object on different sides of a grade boundary: 8.0 versus 8.5, Very Fine versus Very Fine/Near Mint, or Good Plus versus Very Good.

What is the feature?

Defect classification

The experts disagree about what they are seeing: a scratch or print line, wear or weak manufacture, natural toning or staining, factory variation or later alteration.

How serious is it?

Severity

Both identify the same feature but disagree about its extent, visibility or consequence. A soft corner may be minor to one expert and grade-controlling to another.

How should features combine?

Weighting

The observations may be almost identical, yet one expert prioritises structural integrity while another gives greater weight to presentation, originality, colour, lustre or eye appeal.

What is the object?

Attribution or alteration

The disagreement concerns originality, completeness, restoration, trimming, replacement parts, issue state or authenticity. This can invalidate the normal grade rather than merely move it up or down.

Was the evidence adequate?

Process sufficiency

One opinion is based on an in-hand examination while another relies on photographs, an obstructive holder, poor lighting or incomplete access to the reverse, edges or interior.

Why competent experts can disagree

Grading compresses many observations into a short label. Information is inevitably lost, written standards cannot describe every object, and the boundary between adjacent grades exists in the grading system rather than as a physical line on the collectible.

Standards and comparison populations differ

A commercial third-party scale, dealer description, auction condition statement, conservation assessment and insurance record may be answering different questions. Experts may also compare the object with different populations: all surviving examples, recent submissions, a particular production period, certified examples or an internal reference set.

Small differences may therefore be understandable even when both experts are capable. Before comparing grades, confirm that the standards and reference populations are comparable.

Observation conditions change what can be known

Photographs can show centring, missing pieces, major creases and broad colour change, but they are weaker for shallow dents, gloss, texture, cleaning, subtle restoration, waviness, odour, flexibility and defects hidden by glare. An in-hand examiner usually has an evidential advantage, but direct access does not make every conclusion correct.

Expertise is specific, not universal

A general grader may miss a known factory cut, production seam, characteristic shrinkage pattern or regional component variation. Conversely, a production specialist may identify the object correctly but be less calibrated to a formal commercial grade scale. The strongest conclusion may require object identification, material knowledge, alteration detection and grading-system calibration rather than a single undifferentiated idea of “expertise.”

Find the layer where the experts part company

Condition and grade are related but not identical. A productive dispute analysis separates the physical evidence, the grading rule and the final judgement.

Layer 1

Condition evidence

What is physically present?

Examples include corner wear, a crease, gloss loss, a missing component, altered dimensions, staining, a repair or strong colour.

Layer 2

Grading rule

How does the chosen system treat it?

The system may impose a grade ceiling, allow a production tolerance, require a qualifier, or make the object ineligible for a normal numerical grade.

Layer 3

Grade judgement

How are the facts weighted into the outcome?

The expert combines severity, location, originality, presentation, comparison examples and uncertainty into the final descriptor or number.

The most useful diagnostic question

Which single issue would need to change for the experts to agree?

The answer may be whether a mark is a crease, whether an edge was trimmed, whether fading is moderate or severe, whether restoration is present, or whether a specific defect imposes a grade ceiling. Once the decisive issue is named, the next evidence or specialist can be chosen intelligently.

How serious is the disagreement?

Not all disagreements deserve the same response. The hierarchy below moves from ordinary variation in reasoning to disputes that challenge gradability or authenticity.

1

Same grade, different reasoning

The outcome matches, but the experts identify different grade-controlling features.

Collector meaning

Low immediate dispute risk, but the reasoning should still be preserved because it affects future sale descriptions and reassessment.

2

Adjacent-grade disagreement

Examples include 8.0 versus 8.5 or Fine versus Very Fine.

Collector meaning

Often a defensible boundary difference, especially where the experts agree on the physical evidence.

3

Multi-grade disagreement

The opinions differ by several grades rather than one adjacent step.

Collector meaning

Investigate whether a defect was missed, differently classified, or assessed under another standard.

4

Normal grade versus qualified grade

One expert assigns a standard grade while another identifies restoration, cleaning, trimming, incompleteness or another qualifier.

Collector meaning

High significance: the dispute concerns the nature of the object or the permitted grading route, not only degree of wear.

5

Gradable versus ungradable

One expert considers the object eligible for a normal grade; another would authenticate only, details-grade it, or refuse grading.

Collector meaning

Seek relevant specialist evidence before taking commercial or irreversible action.

6

Authentic versus non-authentic

The experts disagree about genuineness, attribution or whether the item is what it claims to be.

Collector meaning

Authentication takes priority. A condition grade is secondary until identity and genuineness are sufficiently resolved.

Legitimate variation, possible error and systemic inconsistency

The collector should test whether the second opinion falls within a reasonable expert range before alleging incompetence, misconduct or institutional failure.

Legitimate expert variation

More likely where grades are adjacent, the item sits near a boundary, both experts identify similar defects, recognised standards are used, material evidence is not overlooked and each expert can explain the reasoning.

Possible error

More concerning where an obvious major defect is missed, the object is misidentified, a factual observation is demonstrably false, alteration evidence is ignored or the result contradicts the stated standard without explanation.

Possible systemic inconsistency

Requires a pattern across comparable objects, defect types, qualifiers or submissions. One disappointing result is not enough to establish a broad institutional problem.

Compare evidence quality before comparing confidence

Ten opinions derived from the same inadequate image do not equal one reliable in-hand examination. Evidence strength is contextual, but the following progression helps collectors weigh conflicting views.

Lowest confidence

Remote impression

A casual opinion from memory or limited photographs may identify obvious defects, but it is weak for texture, gloss, shallow dents, cleaning, odour, flexibility, internal damage and subtle restoration.

Improved but incomplete

Controlled visual evidence

High-resolution front, reverse, edge and angled-light images can narrow disagreement, especially when scale, colour control and defect location are documented.

Stronger observation

In-hand examination

Direct access normally gives the expert an evidential advantage because the object can be viewed from multiple angles, outside sleeves or frames, and under appropriate lighting or magnification.

Issue-matched expertise

Relevant specialist examination

A specialist in the decisive feature—printing, paper, restoration, packaging, materials, signatures or a particular production run—may correct assumptions made by a capable general grader.

Higher confidence

Repeatable or comparative evidence

Measurements, specialist imaging, recognised reference examples, preserved chain of custody and independent convergence on the same observations strengthen the conclusion.

Institutional opinion, consensus and independent opinion

A certified grade is an institutional conclusion under that organisation’s standards at the time of examination. It is not the only possible grade, a permanent truth, a complete condition report or a guarantee of market price.

Multiple-grader and senior-review processes are designed to reduce individual outlier judgements and improve calibration. Consensus increases reliability, but it does not create scientific objectivity. A group can share the same institutional assumptions, incomplete evidence or mistaken attribution.

Independent experts may challenge institutional grades legitimately, but their opinions should be tested for direct access, category fit, familiarity with the grading system, quality of explanation, appropriate comparables and financial independence. A quiet, evidence-linked assessment can be stronger than a forceful declaration that an object is “clearly” another grade.

Technical correctness and market acceptance are different questions

Technical question

Which condition conclusion is best supported by the object, evidence and relevant standard?

Market question

Which label, service or expert opinion will buyers, insurers or registry systems recognise?

Commercial materiality: what decision actually changes?

A disagreement matters in proportion to its consequences. The same half-grade difference can be trivial in one collection and commercially decisive in another.

Often lower materiality

  • The object is retained for personal enjoyment.
  • Rarity dominates price more than small condition differences.
  • The value gap between adjacent grades is modest.
  • The disputed defect is already fully documented and disclosed.

Potentially high materiality

  • A population, registry or sale threshold changes at the higher grade.
  • Auction value jumps sharply by one grade or designation.
  • Insurance, lending or consignment depends on the certified outcome.
  • The dispute concerns authenticity, restoration or sale eligibility.

Large price discontinuities intensify confirmation bias, repeated resubmission, selective disclosure and pressure on experts. The greater the financial jump, the more important it becomes to inspect the collectible rather than purchase the label alone.

Reviews, crossovers and the resubmission paradox

The grading market recognises that prior outcomes may be reconsidered. Review, reconsideration, resubmission, crossover and independent second opinion are different processes with different risks.

Review or reconsideration

The original service re-examines its own certified item, usually because the owner believes the existing outcome is too low or something material was overlooked.

Resubmission

The object enters a new grading event. The result may change because of boundary variation, changed evidence, different calibration or error; a different number does not identify which result is the complete truth.

Crossover

One service evaluates an item already certified by another, sometimes subject to a minimum acceptable outcome before removal from the existing holder.

Independent second opinion

A relevant specialist may clarify the disputed feature without re-encapsulating the object. The opinion may be technically strong even if it carries less market authority than a recognised holder.

Cracking a holder can destroy evidence

Removing an item from a certified holder may cause accidental damage, environmental exposure, loss of certification, weaker chain-of-custody proof and difficulty showing that the prior and later submissions involve the same object.

Document certification numbers, front and back holder images, submission dates, prior grades, alteration designations and the reason for resubmission before any irreversible action. Repeatedly submitting an item and advertising only the highest outcome creates a distorted history.

A proportionate dispute-resolution ladder

The aim is to improve the evidence and resolve the decisive question—not to keep asking until someone supplies the preferred label.

1

Confirm the object and its condition

Check certification numbers, distinctive marks, component configuration, examination dates and whether the item changed between opinions.

2

Reconstruct the examination conditions

Record whether each opinion was in hand, holdered, photographic, video-based, magnified or supported by specialist testing.

3

Identify the standard

Establish the grading system, treatment of production defects, qualifiers, restoration rules and the comparison population being used.

4

Separate observation from interpretation

List what is physically present separately from what each expert believes it means and how it affects the grade.

5

Locate the decisive disagreement

Ask which single issue would need to change for the experts to agree: classification, severity, origin, grade ceiling or eligibility.

6

Choose the right next expertise

Seek a third opinion only when it targets the decisive issue rather than merely adding another general vote.

7

Use the formal review route

Where a grading service is involved, obtain notes, review rules, guarantee terms or appeal procedures before escalating publicly or legally.

8

Quantify materiality and preserve history

Decide what practical outcome changes, then retain all grades, labels, reports, images and correspondence whether or not the dispute is resolved.

When a third opinion is—and is not—warranted

A third opinion should answer a sharper question than the first two opinions did.

Seek targeted specialist input when

  • The first two opinions are materially different.
  • Authenticity, alteration or restoration is disputed.
  • The financial or legal consequence is substantial.
  • The experts used different standards or lacked issue-specific knowledge.
  • Better examination or testing can realistically narrow the difference.

Stop seeking opinions when

  • The reasonable expert range is already clear.
  • Further handling adds risk but no new evidence.
  • The cost exceeds the likely value difference.
  • The remaining disagreement is an irreducible judgement call.
  • The process has become grade shopping rather than investigation.

Conflict of interest and the weight of authority

Conflict does not automatically invalidate an opinion, but it changes how much unsupported judgement should be trusted.

Potential conflicts include a dealer lowering the grade while buying, a seller raising it while marketing, a restorer minimising prior intervention, an auctioneer benefiting from a stronger estimate, an insurer favouring a lower settlement or an owner emotionally and financially invested in an upgrade.

Test authority against relevance and reasoning

  • What does this expert know that the other expert may not know?
  • Did the expert inspect the object directly and under appropriate conditions?
  • Is the expertise relevant to the decisive feature?
  • Can the conclusion be traced back to observations, standards and comparables?
  • Who commissioned the opinion, and does payment or ownership affect the outcome?

Responsible language for disagreement

Good expert language states what is known, what is inferred and what remains uncertain. Confidence should reflect the evidence rather than the desire to sound authoritative.

High confidence

“Page 11 is missing.” “The signature is printed rather than handwritten.” Use only where the evidence is strong and directly supports the statement.

Moderate confidence

“The edge characteristics are consistent with trimming.” “The mark is more likely handling wear than a production defect.”

Limited confidence

“The photographs are insufficient to distinguish a print line from a scratch.” “Both 8.0 and 8.5 appear defensible from the available evidence.”

Myth versus reality

Myth

Experts should always agree.

Reality

Exact agreement is unrealistic where the standard requires judgement, particularly near grade boundaries or where evidence is ambiguous.

Myth

Disagreement proves grading is meaningless.

Reality

A system can be useful and broadly repeatable without producing identical outcomes in every borderline case.

Myth

The highest grade is the correct grade.

Reality

It may simply be the most generous interpretation. The best-supported conclusion is the one that explains the physical evidence with the fewest unsupported assumptions.

Myth

More opinions automatically create truth.

Reality

Several opinions based on the same inadequate photographs can reproduce the same limitation rather than overcome it.

Myth

A changed resubmission grade proves misconduct.

Reality

It proves that the outcome varied. Error, changed condition, different calibration or boundary judgement must be established separately.

Myth

A grading-company label overrides the object.

Reality

The label records an institutional opinion under a particular system and at a particular time. The physical object remains the primary evidence.

Collector risk signals

Exercise greater caution when the disagreement cannot be reconstructed or when commercial incentives are stronger than the evidence.

  • The seller or owner refuses reverse, edge or interior photographs.
  • Only the highest prior grade is disclosed despite multiple certification histories.
  • A major value jump depends on half a grade or one designation.
  • The expert will not identify the standard or grade-controlling feature.
  • Alteration is dismissed confidently without an appropriate examination.
  • A remote opinion is presented as certain despite evidence limits.
  • The expert has a strong financial interest that is not disclosed.
  • No one can explain whether the dispute concerns observation, interpretation or weighting.
  • Reputation and assertion replace reasons, measurements or comparable evidence.

Document the disagreement, not only the final label

A bare grade is difficult to challenge or understand because it conceals the route from evidence to conclusion.

Example collection record

Certified grade: 8.0

Independent assessment: approximately 8.0–8.5

Principal limiting features: light corner wear and shallow reverse indentation

Disputed point: whether the indentation is production-related

Confidence: moderate

Review history: formal review completed; grade unchanged

Evidence retained: grader notes, macro photographs and submission correspondence

Preserve the full disagreement history

Never discard a lower prior grade merely because a later submission achieved a higher one. For significant objects, the disagreement history can become part of the object’s provenance and may be essential in a later sale, insurance claim, review or authenticity dispute.

Object identity, description and distinctive features

Ownership and relevant provenance history

Certification numbers, labels and prior grades

Dates and conditions of every examination

Names, roles and relevant expertise of the assessors

Grading standards, definitions and qualifiers applied

Full-resolution front, reverse, edge and detail photographs

Measurements, imaging or test results

Grader notes, condition reports and restoration reports

Invoices, shipping records and chain-of-custody information

Review, crossover, appeal or resubmission outcomes

Any known change to the object between examinations

The five-element model for a disputed grade

Turn an emotional disagreement into an examinable claim by reconstructing the object, evidence, standard, judgement and confidence.

Given this object, observed under these conditions, using this grading standard, and weighting these features in this way, the expert assigns this grade with this degree of confidence.

1

Object

2

Evidence

3

Standard

4

Judgement

5

Confidence

Key takeaways

  • Treat the object as primary evidence and the grade as an expert interpretation.
  • First determine whether experts see different facts or interpret the same facts differently.
  • Adjacent-grade disagreement is not equivalent to disagreement about alteration, gradability or authenticity.
  • Give greater weight to relevant expertise, adequate access, stated standards and transparent reasoning than to confidence or reputation alone.
  • Use a defensible grade range when the evidence cannot justify a single precise outcome.
  • Preserve all prior grades, reports and opinions, including unfavourable ones.
  • Escalate only when the disagreement materially changes price, rights, insurability, authenticity or another practical decision.
  • Stop seeking opinions when the process becomes grade shopping rather than evidence gathering.

Continue learning

Related topics