March 18, 2024, 4:45 a.m. | Hang Zhang, Wenxiao Zhang, Haoxuan Qu, Jun Liu

cs.CV updates on arXiv.org arxiv.org

arXiv:2403.10107v1 Announce Type: new
Abstract: Human-centered dynamic scene understanding plays a pivotal role in enhancing the capability of robotic and autonomous systems, in which Video-based Human-Object Interaction (V-HOI) detection is a crucial task in semantic scene understanding, aimed at comprehensively understanding HOI relationships within a video to benefit the behavioral decisions of mobile robots and autonomous driving systems. Although previous V-HOI detection models have made significant strides in accurate detection on specific datasets, they still lack the general reasoning ability …

abstract arxiv autonomous autonomous systems benefit capability cs.ai cs.cv cs.mm detection dynamic human llms multiple object pivotal reasoning relationships robotic role semantic systems type understanding via video

Data Architect

@ University of Texas at Austin | Austin, TX

Data ETL Engineer

@ University of Texas at Austin | Austin, TX

Lead GNSS Data Scientist

@ Lurra Systems | Melbourne

Senior Machine Learning Engineer (MLOps)

@ Promaton | Remote, Europe

Principal Data Engineering Manager

@ Microsoft | Redmond, Washington, United States

Machine Learning Engineer

@ Apple | San Diego, California, United States