株式会社極東書店トップ > 商品一覧 > Video Grounding and Its Generalization: From I.D. and Task-specific Models to O.O.D. and Large Foundation Models.
商品詳細
Video Grounding and Its Generalization: From I.D. and Task-specific Models to O.O.D. and Large Foundation Models.
・ISBN 978-3-031-94836-7 hard EUR 149.99
¥40,091.- (税込) ※(※)価格はご注文時の参考価格となります。
納品価格につきましては書籍の入荷時点で確定となります。
版元の原価改定、外国為替の変動等により異なる場合がございますので、予めご了承下さい。
| 著者・編者 | Wang, Xin / Lan, Xiaohan / Zhu, Wenwu, |
|---|---|
| 出版社 | (Springer International Publishing AG, SZ) |
| 出版年月 | 2026 |
| ページ数 | 209 pp. |
| 言語 | ENG |
| ニュース番号 | <A04-48522> |
解説
This book consists of two parts: Part I Methodologies for Video Grounding and Part II Generalized Video Grounding and Trending Directions. To make this book self-contained and cutting edge, Part I will cover basic and advanced methodologies for Video Grounding, discussing key comparisons with several representative Vision-Language learning tasks including multimodal understanding and generation. Part II will cover our insights for Generalized Video Grounding and the development of Video Grounding in the era of large foundation models, discussing future directions such as Out-of-Distribution settings which deserve further investigations.
Discussions on Video Grounding will cover both the task of Video Grounding and other Vision-Language Task, as well as their relations. The basics and advances will touch Video Grounding from model to benchmark, from supervised learning to unsupervised pre-training, from single video grounding to video corpus grounding, and from in-distribution setting to out-of-distribution setting. As for Generalized Video Grounding, we discuss cross-modal grounding, event grounding for multi-modal tasks, various distribution shifts in out-of-distribution setting, explainable Video Grounding, and large foundation model for Video Grounding.
We deeply hope this book can benefit interested readers from both academy and industry, covering needs from junior starters in research to senior practitioners in IT companies.