brain/
conceptai-video-generation

VBench

Notes

VBench

One-line summary: Open, multi-dimension benchmark family for general text-to-video and image-to-video quality — independent of any one model author, but not an avatar / talking-head / hand task board in the retrieved abstracts.

The insight

VBench-class suites exist and are independent. They dissect video quality into named dimensions (including subject identity inconsistency) and, later, intrinsic-faithfulness slices such as Human Fidelity. That is not the same object as a level-playing-field ranking of pose-driven or talking-head avatar models. Author-reported VBench scores on animation models are uses of a general-video metric.

Evidence

Design implications

Use VBench to interpret general video-quality claims. Do not treat a VBench number, or VBench-2.0 Human Fidelity, as an independent avatar / identity / hand leaderboard. See independent-avatar-benchmarks.

Contradictions / tensions

Open questions

  • Does any VBench release (beyond the retrieved abstracts) define a human-animation or talking-head task?
  • Is VBench-2.0 Human Fidelity a face/body identity metric or T2V anatomical correctness?

Sources

Related

Referenced by