Feb. 13, 2024, 5:44 a.m. | Ahmed Heakl Youssef Mohamed Ahmed B. Zaky

cs.LG updates on arXiv.org arxiv.org

This study presents AraSpider, the first Arabic version of the Spider dataset, aimed at improving natural language processing (NLP) in the Arabic-speaking community. Four multilingual translation models were tested for their effectiveness in translating English to Arabic. Additionally, two models were assessed for their ability to generate SQL queries from Arabic text. The results showed that using back translation significantly improved the performance of both ChatGPT 3.5 and SQLCoder models, which are considered top performers on the Spider dataset. Notably, …

arabic community cs.ai cs.cl cs.db cs.ir cs.lg dataset english generate language language processing multilingual natural natural language natural language processing nlp processing speaking sql sql queries study text translation

