---
格式版本: 2
标题: "[2608.25452] VGA-BenchV2: An Expanded Unified Benchmark and Multi-Model Framework for Evaluating Video Aesthetics and Generation Quality"
原文链接: "https://arxiv.org/abs/2608.25452"
发布日期: "2026-08-26"
发布时间校准状态: "found"
发布时间需复核: "否"
发布时间来源: "llm:local:strict_original_body"
发布时间证据: "[Submitted on 26 Aug 2026]"
发布时间校准原因: "arXiv论文正文明确标注提交日期为2026年8月26日，即文章发布时间。"
发布时间校准置信度: "1"
发布时间候选数量: 12
发布时间严格候选数量: 3
发布时间原页读取状态: "source template page reused from URL open"
发布时间未找到原因: ""
发布时间校准时间: "2026-08-27T23:25:27+08:00"
发布时间仲裁状态: "confirmed"
发布时间仲裁尝试次数: 3
发布时间仲裁耗时毫秒: 25079
发现时间: "2026-08-27T23:21:22+08:00"
入库时间: "2026-08-27T15:26:02.380Z"
来源平台: "arXiv 学术论文搜索"
搜索渠道: "source_template"
搜索词: "https://arxiv.org/search/?query=Scale-up&searchtype=all"
匹配关键词:
  - "Scale-up"
  - "AI"
相关厂家:
  - "Google"
相关专家:
  []
内容类型: "网页"
抓取工具: "Jina Reader"
清洗工具: "Jina Reader Markdown + Defuddle/Readability 正文提取"
原始附件:
  []
图片摘要:
  - "✗ ./assets/img-5838d115.png | ad | 与正文无关的赞助商logo"
  - "✗ ./assets/img-26707af3.png | ad | 与正文无关的赞助商logo"
  - "✗ ./assets/img-43a00bd2.png | ad | 与正文无关的赞助商logo"
AI优质: "否"
AI打分: 28
AI分档: "非优质"
AI质检状态: "不通过"
AI打分理由: "正文主线是视频生成质量与美学评价基准，新增1,016个提示、逾6万段视频、3.6万项标注及多模型评价框架；来源为可追溯的arXiv预印本摘要页。固定知识库未见该基准，但未命中不能证明首次出现。文章没有机架级AI架构、互连、供电、散热、RAS或生产部署事实，新增内容属于模型评价与优化，命中应用与模型方向的弱相关否决项，不具备超节点情报准入价值。"
AI质检模型: "gpt-5.6-sol"
AI质检时间: "2026-08-27T23:26:15+08:00"
AI主题相关性: 0
AI来源权威性: 11
AI新颖性: 10
AI技术细节: 0
AI商业部署信号: 0
AI完整性: 7
AI评分提示词版本: "v17-精简生产版"
AI评分提示词SHA256: "48fb9777f386026761b4873eaff30807694fb11e9b352d7c69bf2dfde750cc7d"
AI评分知识库版本: "knowledge_base_v1-20260819+runtime.12"
AI评分知识库SHA256: "13aade27529fda213430b6def2f11efe60b599f8676ae127cad35798cf7169c8"
AI评分知识库检索词: "[\"Scale-up\",\"Google\",\"Intel\",\"VGA-BenchV2\",\"URL\",\"arxiv.org/abs/2608.25452\",\"GMT\",\"HTML\",\"VGA-Bench\",\"PDF\",\"arxiv.org/pdf/2608.25452\",\"arxiv.org/html/2608.25452v1\"]"
AI评分知识库命中: "[{\"id\":\"historical-jan-apr-02\",\"title\":\"二、Google Cloud Next '26：AI Hypercomputer 与第八代 TPU 发布\",\"sourceType\":\"curated_item\",\"time\":\"2026-01_to_2026-04\",\"matchedTerms\":[\"Google\",\"Intel\"],\"rank\":-11.351221802971637},{\"id\":\"july-correct-0081\",\"title\":\"AAI 2026: 6th Gen AMD EPYC Server CPUs Power the Agentic Data Center\",\"sourceType\":\"labeled_article\",\"time\":\"2026-07\",\"matchedTerms\":[\"Intel\",\"URL\",\"HTML\",\"PDF\"],\"rank\":-11.214836116537597},{\"id\":\"july-correct-0111\",\"title\":\"StrataCL: Fabric-Native Communication Library for Production Supernodes\",\"sourceType\":\"labeled_article\",\"time\":\"2026-07\",\"matchedTerms\":[\"URL\",\"HTML\",\"PDF\"],\"rank\":-8.359283396509767},{\"id\":\"july-correct-0104\",\"title\":\"NVIDIA Vera Rubin 提升每瓦性能，为全球合作伙伴实现最低 Token 成本\",\"sourceType\":\"labeled_article\",\"time\":\"2026-07\",\"matchedTerms\":[\"Scale-up\",\"Google\",\"Intel\"],\"rank\":-7.386371895636619},{\"id\":\"july-correct-0024\",\"title\":\"Salience Labs Wants To Scale Up AI With Silicon Photonics Optical Switch\",\"sourceType\":\"labeled_article\",\"time\":\"2026-07\",\"matchedTerms\":[\"Scale-up\",\"Google\",\"URL\"],\"rank\":-7.032150528466252}]"
AI摘要: "研究者发布VGA-BenchV2，把视频美学与生成质量基准扩展到约6万段主流模型生成的视频，并新增3.6万条人工标注；"
AI摘要模型: "ali-deepseek-v4-flash"
AI摘要时间: "2026-08-27T18:59:34.955Z"
采集批次: "2026年8月27日21点49分30秒"
采集批次ID: "20260827-214930-609"
去重键: "https://arxiv.org/abs/2608.25452"
---

Title: [2608.25452] VGA-BenchV2: An Expanded Unified Benchmark and Multi-Model Framework for Evaluating Video Aesthetics and Generation Quality

URL Source: https://arxiv.org/abs/2608.25452

Published Time: Thu, 27 Aug 2026 00:29:23 GMT

Markdown Content:
[Skip to main content](https://arxiv.org/abs/2608.25452#content)[](https://arxiv.org/IgnoreMe)[Search](https://arxiv.org/search)[Submit](https://arxiv.org/user/create)[Donate](https://info.arxiv.org/about/donate.html)[Log in](https://arxiv.org/login)

Search arXiv 

 Press Enter to search · [Advanced search](https://arxiv.org/search/advanced)

# Computer Science > Computer Vision and Pattern Recognition

**arXiv:2608.25452** (cs) 

 [Submitted on 26 Aug 2026]

# Title:VGA-BenchV2: An Expanded Unified Benchmark and Multi-Model Framework for Evaluating Video Aesthetics and Generation Quality

Authors:[Longteng Jiang](https://arxiv.org/search/cs?searchtype=author&query=Jiang,+L), [DanDan Zheng](https://arxiv.org/search/cs?searchtype=author&query=Zheng,+D), [Qianqian Qiao](https://arxiv.org/search/cs?searchtype=author&query=Qiao,+Q), [Heng Huang](https://arxiv.org/search/cs?searchtype=author&query=Huang,+H), [Huaye Wang](https://arxiv.org/search/cs?searchtype=author&query=Wang,+H), [Yihang Bo](https://arxiv.org/search/cs?searchtype=author&query=Bo,+Y), [Bao Peng](https://arxiv.org/search/cs?searchtype=author&query=Peng,+B), [Jingdong Chen](https://arxiv.org/search/cs?searchtype=author&query=Chen,+J), [Jun Zhou](https://arxiv.org/search/cs?searchtype=author&query=Zhou,+J), [Xin Jin](https://arxiv.org/search/cs?searchtype=author&query=Jin,+X)

View a PDF of the paper titled VGA-BenchV2: An Expanded Unified Benchmark and Multi-Model Framework for Evaluating Video Aesthetics and Generation Quality, by Longteng Jiang and 9 other authors

[View PDF](https://arxiv.org/pdf/2608.25452)[HTML (experimental)](https://arxiv.org/html/2608.25452v1)
> Abstract:We introduce VGA-BenchV2, an extended human-aligned benchmark and optimization framework for jointly evaluating and improving video generation quality and aesthetic value. Built upon VGA-Bench, VGA-BenchV2 preserves the original fine-grained taxonomy with two primary dimensions-Aesthetic and Generation-and 52 sub-dimensions. Guided by this taxonomy, we curate 1,016 diverse prompts and collect over 60,000 videos generated by 12 mainstream video generation models. More importantly, VGA-BenchV2 substantially expands human-labeled supervision by adding 36,000 task-level annotations, including 16,200 for aesthetic quality, 13,200 for aesthetic tagging, and 6,600 for generation quality, corresponding to 13.46x, 11.15x, and 1.55x scale-ups over VGA-Bench, respectively. Leveraging this enlarged annotation corpus, we develop a hybrid evaluator architecture consisting of VAQA-Net for continuous aesthetic scoring and two Qwen-based Large Vision-Language Model evaluators, VTag-Net and VGQA-Net, for aesthetic tagging and generation quality assessment. Extensive experiments demonstrate strong alignment with human judgments across diverse generation models. Beyond evaluation, VGA-BenchV2 further introduces an evaluation-to-optimization pipeline, where the learned aesthetic evaluator serves as a reward model for reinforcement learning-based generator fine-tuning. This closes the loop from benchmark construction and human supervision to automated evaluation and model optimization, enabling video generators to improve not only in realism but also in aesthetic quality and human preference alignment. Resources are available at [this https URL](https://huggingface.co/datasets/BestiVictoryLab/VGA-Bench).

Comments:IJCAI 2026
Subjects:Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI)
Cite as:[arXiv:2608.25452](https://arxiv.org/abs/2608.25452) [cs.CV]
(or [arXiv:2608.25452v1](https://arxiv.org/abs/2608.25452v1) [cs.CV] for this version)
[https://doi.org/10.48550/arXiv.2608.25452](https://doi.org/10.48550/arXiv.2608.25452)

Focus to learn more

 arXiv-issued DOI via DataCite (pending registration)

## Submission history

 From: Xin Jin [[view email](https://arxiv.org/show-email/ad54cc05/2608.25452)] 

**[v1]** Wed, 26 Aug 2026 07:16:46 UTC (2,358 KB)

[](https://arxiv.org/abs/2608.25452)Full-text links:
## Access Paper:

 View a PDF of the paper titled VGA-BenchV2: An Expanded Unified Benchmark and Multi-Model Framework for Evaluating Video Aesthetics and Generation Quality, by Longteng Jiang and 9 other authors

*   [View PDF](https://arxiv.org/pdf/2608.25452)
*   [HTML (experimental)](https://arxiv.org/html/2608.25452v1)
*   [TeX Source](https://arxiv.org/src/2608.25452)

[view license](http://creativecommons.org/licenses/by/4.0/ "Rights to this article")

### Current browse context:

cs.CV

[<prev](https://arxiv.org/prevnext?id=2608.25452&function=prev&context=cs.CV "previous in cs.CV (accesskey p)") | [next>](https://arxiv.org/prevnext?id=2608.25452&function=next&context=cs.CV "next in cs.CV (accesskey n)")

[new](https://arxiv.org/list/cs.CV/new) | [recent](https://arxiv.org/list/cs.CV/recent) | [2026-08](https://arxiv.org/list/cs.CV/2026-08)

 Change to browse by: 

[cs](https://arxiv.org/abs/2608.25452?context=cs)

[cs.AI](https://arxiv.org/abs/2608.25452?context=cs.AI)

### References & Citations

*   [NASA ADS](https://ui.adsabs.harvard.edu/abs/arXiv:2608.25452)
*   [Google Scholar](https://scholar.google.com/scholar_lookup?arxiv_id=2608.25452)
*   [Semantic Scholar](https://api.semanticscholar.org/arXiv:2608.25452)

export BibTeX citation Loading...

## BibTeX formatted citation

×

Data provided by: [](https://arxiv.org/abs/2608.25452)

### Bookmark

Bibliographic Tools 

# Bibliographic and Citation Tools

- [x] Bibliographic Explorer Toggle 

Bibliographic Explorer _([What is the Explorer?](https://info.arxiv.org/labs/showcase.html#arxiv-bibliographic-explorer))_

- [x] Connected Papers Toggle 

Connected Papers _([What is Connected Papers?](https://www.connectedpapers.com/about))_

- [x] Litmaps Toggle 

Litmaps _([What is Litmaps?](https://www.litmaps.co/))_

- [x] scite.ai Toggle 

scite Smart Citations _([What are Smart Citations?](https://www.scite.ai/))_

Code, Data, Media 

# Code, Data and Media Associated with this Article

- [x] alphaXiv Toggle 

alphaXiv _([What is alphaXiv?](https://alphaxiv.org/))_

- [x] Links to Code Toggle 

CatalyzeX Code Finder for Papers _([What is CatalyzeX?](https://www.catalyzex.com/))_

- [x] DagsHub Toggle 

DagsHub _([What is DagsHub?](https://dagshub.com/))_

- [x] GotitPub Toggle 

Gotit.pub _([What is GotitPub?](http://gotit.pub/faq))_

- [x] Huggingface Toggle 

Hugging Face _([What is Huggingface?](https://huggingface.co/huggingface))_

- [x] ScienceCast Toggle 

ScienceCast _([What is ScienceCast?](https://sciencecast.org/welcome))_

Demos 

# Demos

- [x] Replicate Toggle 

Replicate _([What is Replicate?](https://replicate.com/docs/arxiv/about))_

- [x] Spaces Toggle 

Hugging Face Spaces _([What is Spaces?](https://huggingface.co/docs/hub/spaces))_

- [x] Spaces Toggle 

TXYZ.AI _([What is TXYZ.AI?](https://txyz.ai/))_

Related Papers 

# Recommenders and Search Tools

- [x] Link to Influence Flower 

Influence Flower _([What are Influence Flowers?](https://influencemap.cmlab.dev/))_

- [x] Core recommender toggle 

CORE Recommender _([What is CORE?](https://core.ac.uk/services/recommender))_

*   [Author](https://arxiv.org/abs/2608.25452)
*   [Venue](https://arxiv.org/abs/2608.25452)
*   [Institution](https://arxiv.org/abs/2608.25452)
*   [Topic](https://arxiv.org/abs/2608.25452)

 About arXivLabs  

# arXivLabs: experimental projects with community collaborators

arXivLabs is a framework that allows collaborators to develop and share new arXiv features directly on our website.

Both individuals and organizations that work with arXivLabs have embraced and accepted our values of openness, community, excellence, and user data privacy. arXiv is committed to these values and only works with partners that adhere to them.

Have an idea for a project that will add value for arXiv's community? [**Learn more about arXivLabs**](https://info.arxiv.org/labs/index.html).

[Which authors of this paper are endorsers?](https://arxiv.org/auth/show-endorsers/2608.25452) | [Disable MathJax](javascript:setMathjaxCookie()) ([What is MathJax?](https://info.arxiv.org/help/mathjax.html)) 

 We gratefully acknowledge support from our **major funders**, [**member institutions**](https://info.arxiv.org/about/ourmembers.html), , and all contributors. 

[About](https://info.arxiv.org/about)·[Help](https://info.arxiv.org/help)·[Contact](https://info.arxiv.org/help/contact.html)·[Subscribe](https://info.arxiv.org/help/subscribe)·[Copyright](https://info.arxiv.org/help/license/index.html)·[Privacy](https://info.arxiv.org/help/policies/privacy_policy.html)·[Accessibility](https://info.arxiv.org/help/web_accessibility.html)·[Operational Status (opens in new tab)](https://status.arxiv.org/)

Major funding support from
