State the difference you want to examine
Choose a narrow non-graphic question, such as the stability of a plain background during a visual transition. Write it before selecting the comparison pair. If the criteria change after you see a preferred result, the comparison becomes harder to interpret.
Use a source that the service permits and that you have appropriate rights to submit. Rules are specific to each platform; a similar-looking catalogue does not establish identical input conditions.
Keep the other conditions constant
Where the interface allows it, retain the same source, output type and relevant settings. Change only the treatment being examined. Record any option you cannot keep equal, because it limits how confidently a difference can be attributed.
Give the pair neutral labels A and B. A preferred brand label or a more attractive preview can bias a review before the complete outputs are inspected.
Use the same acceptance criteria
Review both outputs at normal speed, then inspect corresponding positions. Use the same definition of an unacceptable defect. If a background shift rejects A, do not ignore the same shift in B merely because B has a better opening frame.
Write tradeoffs separately: one treatment may have clearer framing and poorer continuity. Keeping both observations is more informative than forcing an overall winner.
Record costs and uncertainty alongside quality
A comparison includes what each attempt consumed and whether another attempt was required. If one output is rejected, leave its cost in the record. These are your observed account figures, not prices established by this guide.
A small pair establishes limited evidence. If the conditions could not be aligned, explain that rather than publishing a confident ranking.
A practical record
| Keep constant | Record if it differs |
|---|---|
| Permitted source | Composition or source-category changes |
| Output conditions | Duration, type or available quality settings |
| Acceptance rule | The same defect threshold for both results |
A controlled comparison preserves the reasons for a decision. It does not need an unsupported universal winner.