Every first-party result keeps the outputs and failures visible. Exact-input tests are separated from saved output reviews, and focused views never count twice.
43 unique comparison rows144 model outputs shown6 downloadable datasets
Deduplication rule: the Seedream three-model and Lite/Pro pages are useful focused views of the controlled portrait dataset, but they reuse the same rows and do not inflate these totals.
Controlled benchmarks
Controlled same-input model comparisons
These tests preserve the resolved input fingerprint, prompt hash, run ID, model route, and output lineage. A block remains a block; a focused subset remains part of its parent dataset.
The Seedream view below uses the controlled portrait dataset. Broader saved-output reviews remain clearly separated when every historical reference file is not byte-certified.
Five current text-to-image prompts test exact typography, product layout, direct-flash portrait style, directional motion, and constrained illustration.
Five exact prompts use one frozen Mila input, fixed framing, verified model routes, and 15 retained output hashes. This is a focused three-model view of the controlled portrait dataset.
These are original tests from individuals and smaller research or API teams—not broad arena scores. AI Photo links to the source, records the useful finding, and keeps the caveat beside it.
12 models · 35 portrait prompts · 2,310 paired rows · about 22,000 human responses
Method
Raters selected which paired output followed the written description of the human better; the dataset publishes weighted responses and per-model win totals.
Useful signal
Seedream 3, Qwen Image, and Recraft v3 had the three highest raw win rates in the published historical comparison.
Limit
Unequal matchup exposure, older named model versions, and no reference identity; this is not a current universal portrait-model ranking.
One frozen Mila input, with the prompt, profile, portrait aspect ratio, and model order held fixed within every row.
Useful signal
The side-by-sides expose pose, lighting, wardrobe, and identity tradeoffs without changing the subject or request; the article leaves the quality call to the reader.
Limit
One trained subject and one retained output per model and prompt; the comparison intentionally does not declare an overall winner.
Prompts were published, settings held constant, and the same seed was used across models within each prompt.
Useful signal
Flux 2 Dev led overall at 51/60, while Qwen 2512 was the more balanced runner-up and anatomy changed the ordering materially.
Limit
One author's 1–5 judgments and local workflows; the author stopped after 12 written tests despite exploring a larger set.
RedditPrompt adherenceLocal models
Reddit AI girlfriend app review survey
2026 review survey
AI Girlfriend App Reviews on Reddit
29,146 collected posts became 35 eligible firsthand reviews from 34 authors after affiliate signals, coordinated promotion, copied templates, product affiliations, and non-firsthand material were removed.
Exact input checksum, prompt, route, settings, run conditions, and every outcome are retained.
02
Blind evaluation
Model identities are hidden while a named rubric is applied; panel size and agreement stay visible.
03
Output review
Saved prompts and outputs are inspectable, but at least one laboratory control is missing.
04
Review survey
Published user records are cleaned and deduplicated. The result describes the retained evidence set.
For writers and researchers
Citation and image reuse guidance
Link to the full article or dataset so readers can inspect the rows. Comparison images reproduced on AI Photo are credited to Remix.Camera; external tests remain on their publishers' pages.
AI Photo Editorial. “Five AI Image Models, One Input: 27-Output Controlled Portrait Test.” July 18, 2026. https://aiphoto.ai/blog/controlled-ai-image-model-comparison
Controlled five-model portrait test
Exact-input portrait rows with prompts, routes, hashes, outputs, and the one blocked or missing outcome retained.