Feb. 9, 2024, 5:43 a.m. | Xing Han L\`u Zden\v{e}k Kasner Siva Reddy

cs.LG updates on arXiv.org arxiv.org

We propose the problem of conversational web navigation, where a digital agent controls a web browser and follows user instructions to solve real-world tasks in a multi-turn dialogue fashion. To support this problem, we introduce WEBLINX - a large-scale benchmark of 100K interactions across 2300 expert demonstrations of conversational web navigation. Our benchmark covers a broad range of patterns on over 150 real-world websites and can be used to train and evaluate agents in diverse scenarios. Due to the magnitude …

cs.cl cs.cv cs.lg

