Forty frames, one box set
The reason this page exists is the folder. Somebody has forty captures from the same app, at the same size, each carrying the same header or the same watermark strip, and clearing them one at a time is forty repetitions of an identical drag.
Drop the whole set. Frames are grouped by their pixel dimensions, you draw the boxes once on the first frame of the largest group, and the queue applies that set to every frame in it with a status line each. What comes back is a ZIP, written here in the tab from the files the queue produced.
What the queue will not do
It will not process a frame whose dimensions do not match its group. Those are listed and set aside, because a box at 1,200 by 90 means something quite different on a frame that is 900 pixels wide, and the failure would be silent.
It will not survive a closed tab. There is no resume in this build, so a run of two hundred frames wants a window you are not about to shut. That is a limitation of what is shipped today rather than a design position.
Flat ground is why this page is the easy one
An interface panel has no texture and, usually, no structure crossing it. Both of those are what make a fill hard. Over a flat or softly graded panel the diffusion fill — the one that downloads nothing — is effectively invisible and finishes before you have let go of the button, and there is very little the trained model can add.
Where a screenshot does get hard is a second line of smaller type immediately under the one you boxed, or a table rule running through the box. Structure crossing a border is the case where a trained fill invents a confident, wrong surface, and the box list marks those regions before you run them.
Questions about screenshots
- Why are screenshots the easy case?
- Because the ground under the lettering is usually flat. An interface panel is one colour or a gentle gradient, and both fills rebuild that perfectly — the diffusion one does it in a fifth of a second with nothing downloaded. The hard cases on this site are photographs with texture or structure behind the words, and a screenshot rarely has either.
- Should I save as PNG or JPG?
- PNG, unless something downstream insists otherwise. A screenshot is flat colour and hard edges, which is exactly what JPG handles worst: re-encode a cleared caption bar as JPG and it picks up ringing along every edge, right where you have been looking. PNG is the format this page offers first for that reason.
- How does the queue decide what goes together?
- By pixel dimensions, and by nothing else. Frames that match are one group and share one box set; frames that do not are listed and set aside. It is a deliberately dumb rule, because the alternative is guessing that two differently sized frames have their captions in the same relative place, and a wrong guess spoils a picture rather than clearing one.
- Can I stop a long run part-way?
- Yes. Pause takes effect after the frame in progress, and the finished ones are still yours — the ZIP holds whatever completed. Closing the tab ends the run for good, though, because nothing here is written to your disk without you pressing save.
Elsewhere
- Subtitle stripsOpens with a wide box along the lower third, which is where a burned-in caption almost always sits.
- Flat ground, busy groundThe single best predictor of whether you will be pleased with the result, worked through on three grounds.
- ImageTextRemoverDraw a box over the lettering, keep the boxes you meant, and get the picture back with the ground painted across them.
- Lettering somebody else put thereWhat clearing a mark does and does not change about who owns the picture underneath it.