Visual AI Lab compares AI image and video generation models under the same conditions, records reproducibility, and applies final human review to determine which models work in real Japan-market production.
The aggregation units used across the site are defined as follows.
| Term | Definition |
|---|---|
| Run | One generation execution for one model |
| Task | One input / prompt / production-purpose combination |
| 3-Run task | A task generated three times under the same model and conditions |
| Comparison case | One task compared across multiple models |
| Theme | Subject unit such as ramen, uchiwa fan, or product rotation |
| Category | Capability, use case or Japan-market fit classification |
| Article | Editorial content that explains comparison results |
Every comparison page shows its evaluation type.
Automation is fine for production and first-pass sorting, but the following are always reviewed by humans:
Published evaluation scores come from manual evaluation in our admin tool. Articles containing images or video are not published until human review is complete.
Averages, rankings and win rates on this site are calculated as follows.
Image and video models are checked against different Japan-market axes.
Categories such as "Japanese animals", "Japanese scenery" or "Japanese people" are vague and are therefore redefined as follows.
This site is not a rigorous academic benchmark; it is a practical comparison review based on production use. Published content may change as models are updated or prices change. Please check each provider's official documentation for the latest information.