HyperAI
Command Palette
Search for a command to run...
Vchitect T2V ビデオ生成データセット
Vchitect T2Vは、上海人工知能研究所が2025年に公開したビデオ生成データセットです。関連する論文の結果は次のとおりです。Vchitect-2.0: ビデオ拡散モデルのスケールアップのための並列トランスフォーマー「テキストとビジュアルコンテンツを変換するモデルの能力を向上させることに重点を置き、研究者や開発者が画像生成、意味理解、クロスモーダルタスクなどの進歩を促進できるようにすることを目指しています。」 このデータセットには、それぞれ詳細なテキスト キャプションが付いた 1,400 万本の高品質ビデオが含まれています。

引用
@article{fan2025vchitect,
title={Vchitect-2.0: Parallel Transformer for Scaling Up Video Diffusion Models},
author={Fan, Weichen and Si, Chenyang and Song, Junhao and Yang, Zhenyu and He, Yinan and Zhuo, Long and Huang, Ziqi and Dong, Ziyue and He, Jingwen and Pan, Dongwei and others},
journal={arXiv preprint arXiv:2501.08453},
year={2025}
}
@article{si2025RepVideo,
title={RepVideo: Rethinking Cross-Layer Representation for Video Generation},
author={Si, Chenyang and Fan, Weichen and Lv, Zhengyao and Huang, Ziqi and Qiao, Yu and Liu, Ziwei},
journal={arXiv 2501.08994},
year={2025}
}
このデータセットはコミュニティユーザーによって提供されており、教育および情報提供のみを目的としています。著作権侵害に関わるコンテンツがある場合は、[email protected]までご連絡ください。速やかに確認し、削除いたします。