all AI news
Enhancing Human-Centered Dynamic Scene Understanding via Multiple LLMs Collaborated Reasoning
March 18, 2024, 4:45 a.m. | Hang Zhang, Wenxiao Zhang, Haoxuan Qu, Jun Liu
cs.CV updates on arXiv.org arxiv.org
Abstract: Human-centered dynamic scene understanding plays a pivotal role in enhancing the capability of robotic and autonomous systems, in which Video-based Human-Object Interaction (V-HOI) detection is a crucial task in semantic scene understanding, aimed at comprehensively understanding HOI relationships within a video to benefit the behavioral decisions of mobile robots and autonomous driving systems. Although previous V-HOI detection models have made significant strides in accurate detection on specific datasets, they still lack the general reasoning ability …
abstract arxiv autonomous autonomous systems benefit capability cs.ai cs.cv cs.mm detection dynamic human llms multiple object pivotal reasoning relationships robotic role semantic systems type understanding via video
More from arxiv.org / cs.CV updates on arXiv.org
Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
2 days, 2 hours ago |
arxiv.org
Jobs in AI, ML, Big Data
Data Architect
@ University of Texas at Austin | Austin, TX
Data ETL Engineer
@ University of Texas at Austin | Austin, TX
Lead GNSS Data Scientist
@ Lurra Systems | Melbourne
Senior Machine Learning Engineer (MLOps)
@ Promaton | Remote, Europe
Principal Data Engineering Manager
@ Microsoft | Redmond, Washington, United States
Machine Learning Engineer
@ Apple | San Diego, California, United States