I'm not comfortable commissioning our UR5e bezel inspection from these labels. A fixed camera sees bright streaks move with small presentation changes on the reflective aluminium, yet the examples call them scratches. What should establish the physical references and capture conditions before we tune thresholds?
What changed between those pictures: seating, orientation, lighting, or several things? The streak moving is useful evidence, but it doesn't decide whether the surface also has a scratch.
Preserve the full images and their setup notes. Cropping to the bright region may hide a seating difference. Keep uncertain labels out of the accepted tuning set without deleting the examples.
I've seen passed test results lose all their limitations in a handover, so I'd put the reference status beside each image where the next person will actually find it
The bezels are retained. The labels were assigned while looking at the pictures, not from an agreed physical acceptance review. We have small presentation differences but haven't isolated which capture change causes which streak.
Sara, would you stop all imaging comparisons until quality finishes? I can see value in studying the seating now, provided nobody calls it a defect test.
Use identified pieces and record one deliberate permitted setup change at a time for that study. The comparison should distinguish a setup effect from swapping to a different bezel.
Quality has now examined the retained pieces under an agreed view. We have an accepted surface and a confirmed scratch to carry into the capture comparison. The original picture labels weren't reliable enough to keep unchanged.
Yes, I've marked the old examples superseded and linked the physical judgements. Capture comparison is still underway. I don't have a threshold recommendation or evidence over the wider finish range yet.