The scratch labels look less certain than the images suggest (Fairino FR10)

RaviChan1085 · 10 May 2026, 17:04 UTC

Reply to discussion
RA
RaviChan1085
Our planned presenting reflective components for surface inspection in a cosmetic components workshop has Fairino FR10 presenting a bright-finished enclosure faceplate to a fixed camera. The example images contain bright streaks that shift with small presentation differences, yet those streaks are being labelled as scratches. Before commissioning, i'd like to establish reliable references and image conditions for threshold tuning.

21 replies

TO
TobyAbbott0057
Replying to RaviChan1085

Were those scratches confirmed on the physical samples using your agreed appearance method? Start by linking each image to a sample and its inspection result.

7 points
RA
RaviChan1085
Replying to TobyAbbott0057

@TobyAbbott0057 Some labels came from the pictures alone. We still have the samples, but those labels aren't tied to physical inspection results. That's irritatingly circular.

20 points
TO
TobyAbbott0057
Replying to RaviChan1085

Mark those labels uncertain and have the samples inspected properly. Then compare images with presentation and lighting controlled; the moving streaks are a useful reflection clue.

14 points
RO
RobinBell0686
Replying to TobyAbbott0057

That clue needs a limit: presentation can also change how a real scratch appears. Movement in the image isn't enough to reclassify the sample as acceptable.

-2 points
TO
TobyAbbott0057
Replying to RobinBell0686

Correct. I meant a reason to investigate the imaging, not a new defect label. Physical acceptance stays with the agreed inspection method.

13 points
SA
SarahBaker0502
Replying to TobyAbbott0057

So uncertain isn't the same as good? Where do those images go?

9 points
TO
TobyAbbott0057
Replying to SarahBaker0502

Keep them explicitly undecided, with the sample link. Don't use them as confirmed examples of either class until the physical review settles them.

16 points
RA
RaviChan1085
Replying to TobyAbbott0057

@TobyAbbott0057 Our two comparison folders changed exposure and presentation together. i can't tell which change made the images look better. That comparison needs a caveat too.

20 points
LE
LeoBaker0449
Replying to RaviChan1085

I had a separate inspection setup where the 'better lighting' folder used easier samples too. Its name implied a lighting conclusion that the comparison didn't support.

14 points
RA
RaviChan1085
Replying to LeoBaker0449

What did you keep beside each image? i can manage sample ID and settings, but i don't want a form nobody fills in.

18 points
LE
LeoBaker0449
Replying to RaviChan1085

@RaviChan1085 For my separate setup, sample identity, exposure, lighting arrangement and presentation condition were the useful essentials. They let us understand comparisons without relying on recollection.

17 points
TO
TobyAbbott0057
Replying to LeoBaker0449

Use those details to ask your vision specialist for controlled comparisons on representative surfaces. Keep the sample set fixed when assessing an acquisition change.

13 points
SA
SarahBaker0502
Replying to TobyAbbott0057

Would a learned model just cope with the lighting differences?

23 points
TO
TobyAbbott0057
Replying to SarahBaker0502

It still needs trustworthy labels and representative evaluation. A different model doesn't establish whether your reference scratches are real or your acquisition comparison is fair.

22 points
RO
RobinBell0686
Replying to TobyAbbott0057

@TobyAbbott0057 Keep physical sample identity in mind when forming evaluation sets. Different images of a sample used for tuning don't provide an independent sample-level evaluation.

18 points
RA
RaviChan1085
Replying to RobinBell0686

@RobinBell0686 i'll keep all images of one sample together when we split tuning and evaluation sets. Our file names alone wouldn't have caught that overlap.

12 points
TO
TobyAbbott0057
Replying to RaviChan1085

Report missed defects and false rejects separately on the held-out samples. One agreement percentage can hide which mistake your method is making.

13 points
RA
RaviChan1085
Replying to TobyAbbott0057

Should undecided physical samples count as misses in that report? i don't want them quietly disappearing just because they're awkward.

17 points
TO
TobyAbbott0057
Replying to RaviChan1085

Show them as undecided with their count, outside confirmed-class error rates. That keeps the coverage limit visible without inventing a ground-truth label.

12 points
RA
RaviChan1085
Replying to TobyAbbott0057

@TobyAbbott0057 The reference problem is clearer, but i still can't judge inspection performance from our example images. The physical labels and image comparisons need more work.

6 points

Add to the discussion

Welcome to Application Robot

Everyone can read the forum. Sign in or create an account to start a discussion, reply, or upload photos.

Forgot your password?

By creating an account, you agree to our Terms and Conditions and community guidelines. Read our Privacy Policy for how your information is handled.