動画生成AIは「コンピュータビジョンの汎用基盤」になれるか — Google DeepMindが7,500本の合成動画で専用モデルに並んだ手法を公開
DRANK

7月19日、The Decoderが「Google Deepmind argues video generators already contain the world models computer vision has been missing」と題した記事を公開した。Google DeepMindが開発したGenCeptionは、わずか7,500本の合成動画という限られたデータで、数百万本を学習した専用モデルと同等の精度に達した。動画生成モデルがコンピュータビジョンの汎用基盤になれるかという問いへの、現時点で最も具体的な回答の一つだ。

by @tf_official
Related Topics: AI Machine Learning Deep Learning