The scratch labels look less certain than the images suggest

RachelChan1131 · 23 Jul 2026, 15:28 UTC

Reply to discussion
RA
RachelChan1131
We're planning presenting reflective components for surface inspection, with Universal Robots UR5e presenting a polished metal cover to a fixed camera in a robot-presented camera station. Bright streaks move with small presentation changes in our example images, but people keep calling them scratches. We haven't commissioned the cell. How do I sort the references and imaging before we tune thresholds around this mess?

20 replies

HA
HanaAllen0274
Replying to RachelChan1131

Check whether the suspected scratches were confirmed on physical samples under the agreed appearance procedure. Link each image to its sample identity and inspection result first.

22 points
RA
RachelChan1131
Replying to HanaAllen0274

Some labels came from the pictures alone. We still have the samples, but those labels aren't tied to physical inspection results. That's irritatingly circular.

16 points
HA
HanaAllen0274
Replying to RachelChan1131

Mark those labels uncertain and have the samples inspected properly. Then compare images with presentation and lighting controlled; the moving streaks are a useful reflection clue.

12 points
LO
LouisChan1094
Replying to HanaAllen0274

That clue needs a limit: presentation can also change how a real scratch appears. Movement in the image isn't enough to reclassify the sample as acceptable.

-1 points
HA
HanaAllen0274
Replying to LouisChan1094

Correct. I meant a reason to investigate the imaging, not a new defect label. Physical acceptance stays with the agreed inspection method.

13 points
BE
BethBaker0498
Replying to HanaAllen0274

So uncertain isn't the same as good? Where do those images go?

21 points
HA
HanaAllen0274
Replying to BethBaker0498

Keep them explicitly undecided, with the sample link. Don't use them as confirmed examples of either class until the physical review settles them.

11 points
RA
RachelChan1131
Replying to HanaAllen0274

I've checked the image folders and found that exposure and presentation both changed between them. Their improved appearance can't be attributed to either change individually.

15 points
FI
FionaChen1195
Replying to RachelChan1131

@RachelChan1131 On my inspection setup, a folder called 'better lighting' also contained easier samples. The folder name was doing a lot of unearned work.

17 points
RA
RachelChan1131
Replying to FionaChen1195

What did you keep beside each image? I can manage sample ID and settings, but I don't want a form nobody fills in.

6 points
FI
FionaChen1195
Replying to RachelChan1131

On my setup we kept sample ID, exposure, lighting arrangement and presentation condition. Enough to explain a comparison without reconstructing the whole session from memory.

2 points
HA
HanaAllen0274
Replying to FionaChen1195

@FionaChen1195 Use those details to ask your vision specialist for controlled comparisons on representative surfaces. Keep the sample set fixed when assessing an acquisition change.

10 points
BE
BethBaker0498
Replying to HanaAllen0274

@HanaAllen0274 Would a learned model just cope with the lighting differences?

18 points
HA
HanaAllen0274
Replying to BethBaker0498

It still needs trustworthy labels and representative evaluation. A different model doesn't establish whether your reference scratches are real or your acquisition comparison is fair.

13 points
LO
LouisChan1094
Replying to HanaAllen0274

@HanaAllen0274 Keep physical sample identity in mind when forming evaluation sets. Different images of a sample used for tuning don't provide an independent sample-level evaluation.

24 points
RA
RachelChan1131
Replying to LouisChan1094

@LouisChan1094 I'll keep all images of one sample together when we split tuning and evaluation sets. Our file names alone wouldn't have caught that overlap.

9 points
HA
HanaAllen0274
Replying to RachelChan1131

Report missed defects and false rejects separately on the held-out samples. One agreement percentage can hide which mistake your method is making.

4 points
RA
RachelChan1131
Replying to HanaAllen0274

Should undecided physical samples count as misses in that report? I don't want them quietly disappearing just because they're awkward.

4 points
HA
HanaAllen0274
Replying to RachelChan1131

Show them as undecided with their count, outside confirmed-class error rates. That keeps the coverage limit visible without inventing a ground-truth label.

10 points
FI
FionaChen1195
Replying to HanaAllen0274

That helped on my setup. People could see how much remained undecided instead of mistaking a cleaner-looking table for better inspection performance.

1 points

Add to the discussion

Welcome to Application Robot

Everyone can read the forum. Sign in or create an account to start a discussion, reply, or upload photos.

Forgot your password?

By creating an account, you agree to our Terms and Conditions and community guidelines. Read our Privacy Policy for how your information is handled.