Constrained Decoding for Cross-lingual Label Projection | allainews.com

Feb. 6, 2024, 5:47 a.m. | Duong Minh Le Yang Chen Alan Ritter Wei Xu

cs.LG updates on arXiv.org arxiv.org

Zero-shot cross-lingual transfer utilizing multilingual LLMs has become a popular learning paradigm for low-resource languages with no labeled training data. However, for NLP tasks that involve fine-grained predictions on words and phrases, the performance of zero-shot cross-lingual transfer learning lags far behind supervised fine-tuning methods. Therefore, it is common to exploit translation and label projection to further improve the performance by (1) translating training data that is available in a high-resource language (e.g., English) together with the gold labels into …

become cross-lingual cs.cl cs.lg data decoding exploit fine-grained fine-tuning languages llms low multilingual nlp paradigm performance popular predictions projection supervised fine-tuning tasks training training data transfer transfer learning translation words zero-shot

More from arxiv.org / cs.LG updates on arXiv.org

TAnet: A New Temporal Attention Network for EEG-based Auditory Spatial Attention Decoding with a Short … 20 hours ago | arxiv.org

abstract arxiv attention cs.lg +14

State Derivative Normalization for Continuous-Time Deep Neural Networks 20 hours ago | arxiv.org

abstract arxiv continuous cs.lg +15

Measurement-driven neural-network training for integrated magnetic tunnel junction arrays 20 hours ago | arxiv.org

abstract applications arrays arxiv +21

Higher-Order Equivariant Neural Networks for Charge Density Prediction in Materials 20 hours ago | arxiv.org

abstract arxiv challenge cond-mat.mtrl-sci +21

Non-parametric regression for robot learning on manifolds 20 hours ago | arxiv.org

abstract applications arxiv cs.lg +17

Learning the dynamics of a one-dimensional plasma model with graph neural networks 20 hours ago | arxiv.org

abstract arxiv class cs.lg +20

RealFill: Reference-Driven Generation for Authentic Image Completion 20 hours ago | arxiv.org

arxiv authentic cs.ai cs.cv +6

Leveraging Self-Supervised Vision Transformers for Segmentation-based Transfer Function Design 20 hours ago | arxiv.org

abstract arxiv color cs.cv +19

Dilated convolutional neural network for detecting extreme-mass-ratio inspirals 20 hours ago | arxiv.org

abstract arxiv astro-ph.im binary +19

Data Engineer

@ Lemon.io | Remote: Europe, LATAM, Canada, UK, Asia, Oceania

View on ai-jobs.net

Artificial Intelligence – Bioinformatic Expert

@ University of Texas Medical Branch | Galveston, TX

View on ai-jobs.net

Lead Developer (AI)

@ Cere Network | San Francisco, US

View on ai-jobs.net

Research Engineer

@ Allora Labs | Remote

View on ai-jobs.net

Ecosystem Manager

@ Allora Labs | Remote

View on ai-jobs.net

Founding AI Engineer, Agents

@ Occam AI | New York

View on ai-jobs.net