Web: http://arxiv.org/abs/2209.10918

Sept. 23, 2022, 1:14 a.m. | Zhijian Hou, Wanjun Zhong, Lei Ji, Difei Gao, Kun Yan, Wing-Kwong Chan, Chong-Wah Ngo, Zheng Shou, Nan Duan

cs.CV updates on arXiv.org arxiv.org

Video temporal grounding (VTG) targets to localize temporal moments in an
untrimmed video according to a natural language (NL) description. Since
real-world applications provide a never-ending video stream, it raises demands
for temporal grounding for long-form videos, which leads to two major
challenges: (1) the long video length makes it difficult to process the entire
video without decreasing sample rate and leads to high computational burden;
(2) the accurate multi-modal alignment is more challenging as the number of
moment candidates …

alignment arxiv fine framework temporal video

More from arxiv.org / cs.CV updates on arXiv.org

Postdoctoral Fellow: ML for autonomous materials discovery

@ Lawrence Berkeley National Lab | Berkeley, CA

Research Scientists

@ ODU Research Foundation | Norfolk, Virginia

Embedded Systems Engineer (Robotics)

@ Neo Cybernetica | Bedford, New Hampshire

2023 Luis J. Alvarez and Admiral Grace M. Hopper Postdoc Fellowship in Computing Sciences

@ Lawrence Berkeley National Lab | San Francisco, CA

Senior Manager Data Scientist

@ NAV | Remote, US

Senior AI Research Scientist

@ Earth Species Project | Remote anywhere

Research Fellow- Center for Security and Emerging Technology (Multiple Opportunities)

@ University of California Davis | Washington, DC

Staff Fellow - Data Scientist

@ U.S. FDA/Center for Devices and Radiological Health | Silver Spring, Maryland

Staff Fellow - Senior Data Engineer

@ U.S. FDA/Center for Devices and Radiological Health | Silver Spring, Maryland

Computational Linguist

@ Constant Contact | Waterloo, ON/Hybrid, ON

Senior Data Scientist

@ Advian | Espoo, Finland

Data Scientist, EU In-stock Management

@ Gopuff | London, England