A useful test starts with a narrow question such as whether character traits remain recognizable across scene changes or whether a pose reference preserves the intended structure. The prompt set, model, dimensions, settings, references, and sample count are recorded before interpreting outputs.
When comparing models or controls, all other available inputs stay fixed. Each condition receives multiple runs because generative systems vary. Outputs are assessed against declared criteria such as identity traits, prompt adherence, composition, anatomical defects, or reference influence.
A result should include the test date, active product version, model labels shown to users, prompt set, settings, reference rights, sample handling, and known limitations. Lewdora does not convert an internal impression into a benchmark claim or present a small test as universal performance.