Spatiotemporal Image Understanding: Action, Activity and Behavior
暫譯: 時空影像理解:行動、活動與行為

Zhang, Yu-Jin

  • 出版商: Springer
  • 出版日期: 2026-05-16
  • 售價: $8,470
  • 貴賓價: 9.5$8,046
  • 語言: 英文
  • 頁數: 246
  • 裝訂: Hardcover - also called cloth, retail trade, or trade
  • ISBN: 9819571294
  • ISBN-13: 9789819571291
  • 相關分類: 影像辨識 Image-recognition
  • 海外代購書籍(需單獨結帳)

商品描述

What does it take to move from recognizing static objects in images to truly understanding dynamic human behavior in complex scenes? This book provides the answer by introducing a groundbreaking framework that fuses image engineering with spatiotemporal behavior understanding (STBU), offering an in-depth exploration of how action, interaction, and context converge in real-world image analysis.

Positioned at the intersection of image understanding, neural networks, and behavioral modeling, this volume equips researchers and engineers with the principles, methods, and architectures needed to analyze and interpret dynamic visual data. It guides readers through each stage of the pipeline--from interest point detection and trajectory learning to action classification, activity modeling, and human-object interaction analysis--culminating in advanced topics such as abnormal event detection and graph-based neural modeling. Throughout, the book introduces deep learning strategies for action and behavior recognition, high-order modeling techniques that integrate motion, posture, and context, and transformer-based approaches for human-object interaction. It also addresses practical challenges such as differential explosion and adapting recognition models to varied scene content.

This book is essential reading for graduate students, researchers, and practitioners in computer vision, artificial intelligence, and robotics who seek a comprehensive yet accessible guide to high-level image understanding. A working knowledge of machine learning and basic computer vision concepts is recommended for full benefit. Whether you're advancing academic research or building real-world intelligent systems, this volume provides both the theoretical insight and applied techniques to push the frontier of spatiotemporal image understanding.

商品描述(中文翻譯)

要從識別靜態物體於影像中,轉變為真正理解複雜場景中的動態人類行為,需要什麼?本書提供了答案,介紹了一個開創性的框架,將影像工程與時空行為理解(spatiotemporal behavior understanding, STBU)融合在一起,深入探討行動、互動和背景如何在現實世界的影像分析中交匯。

本書位於影像理解、神經網絡和行為建模的交匯點,為研究人員和工程師提供分析和解釋動態視覺數據所需的原則、方法和架構。它引導讀者通過每個階段的流程——從興趣點檢測和軌跡學習到行動分類、活動建模和人機互動分析——最終涵蓋異常事件檢測和基於圖形的神經建模等進階主題。在整個過程中,本書介紹了用於行動和行為識別的深度學習策略、高階建模技術(整合運動、姿勢和背景)以及基於變壓器的方式來處理人機互動。它還解決了如微分爆炸和將識別模型適應於不同場景內容等實際挑戰。

本書是計算機視覺、人工智慧和機器人學研究生、研究人員和實務工作者必讀的資料,尋求一個全面而易於理解的高階影像理解指南。建議具備機器學習和基本計算機視覺概念的工作知識,以獲得最佳效益。無論您是在推進學術研究還是構建現實世界的智能系統,本書都提供了理論見解和應用技術,以推進時空影像理解的前沿。

作者簡介

Yu-Jin Zhang is a Professor of Image Engineering at Tsinghua University, Beijing, where he has been on faculty since 1993. He received his Ph.D. in Applied Science from the State University of Liège, Belgium, in 1989, followed by postdoctoral research at Delft University of Technology in the Netherlands. With over three decades of experience in image processing, Image Analysis, image understanding, and visual information retrieval, he is widely recognized as a leading scholar in the field.

Professor Zhang has authored more than 550 peer-reviewed research papers and published over 20 academic books, including Handbook of Image Engineering (Springer, 2021), A Selection of Image Processing Techniques (CRC Press, 2022), A Selection of Image Analysis Techniques (CRC Press, 2023), A Selection of Image Understanding Techniques (CRC Press, 2023), and 3-D Computer Vision: Foundations and Advanced Methodologies (Springer, 2024). His work has significantly shaped the landscape of image engineering (image processing, analysis and understanding).

In addition to his research, he has published more than 40 textbooks in Image engineering, developed and taught more than ten specialized courses at Tsinghua University and abroad, inspiring generations of students and practitioners. He is a Fellow of SPIE and the China Society of Image and Graphics (CSIG), and he served as Program Chair of ICIP 2017.

作者簡介(中文翻譯)

張宇進是北京清華大學影像工程的教授,自1993年以來一直在該校任教。他於1989年在比利時列日大學獲得應用科學博士學位,隨後在荷蘭代爾夫特理工大學進行博士後研究。擁有超過三十年的影像處理、影像分析、影像理解和視覺信息檢索的經驗,他被廣泛認為是該領域的領軍學者。

張教授已發表超過550篇經過同行評審的研究論文,並出版了超過20本學術著作,包括《影像工程手冊》(Springer, 2021)、《影像處理技術選輯》(CRC Press, 2022)、《影像分析技術選輯》(CRC Press, 2023)、《影像理解技術選輯》(CRC Press, 2023)以及《3D計算機視覺:基礎與進階方法論》(Springer, 2024)。他的研究顯著地塑造了影像工程(影像處理、分析和理解)的發展。

除了研究外,他還出版了超過40本影像工程的教科書,在清華大學及國外開發並教授了十多門專業課程,啟發了幾代學生和從業者。他是SPIE和中國圖像與圖形學會(CSIG)的會士,並曾擔任2017年國際影像處理會議(ICIP)的程序主席。