A model benchmark describes the tested model and task. It does not measure the whole tool, its interface, or every available plan.
Documented models with evaluation sources
These model names appear in both the product documentation and an external evaluation. Endpoint revisions, settings and the complete tool workflow have not been matched.
Artificial Analysis · Text-to-video models. What it measures: Published evaluation protocol. Read the publisher’s evaluation method, task scope, and tested model version before applying its results to a tool. No tool score is inferred here.
VBench team · Text-to-video models. What it measures: Quality and consistency dimensions. A multi-dimensional video benchmark covering visual quality, motion and consistency.