all AI news
Show Your Work with Confidence: Confidence Bands for Tuning Curves
April 10, 2024, 4:43 a.m. | Nicholas Lourie, Kyunghyun Cho, He He
cs.LG updates on arXiv.org arxiv.org
Abstract: The choice of hyperparameters greatly impacts performance in natural language processing. Often, it is hard to tell if a method is better than another or just better tuned. Tuning curves fix this ambiguity by accounting for tuning effort. Specifically, they plot validation performance as a function of the number of hyperparameter choices tried so far. While several estimators exist for these curves, it is common to use point estimates, which we show fail silently and …
abstract accounting arxiv confidence cs.cl cs.lg impacts language language processing natural natural language natural language processing performance plot processing show stat.ml type validation work
More from arxiv.org / cs.LG updates on arXiv.org
Jobs in AI, ML, Big Data
AI Research Scientist
@ Vara | Berlin, Germany and Remote
Data Architect
@ University of Texas at Austin | Austin, TX
Data ETL Engineer
@ University of Texas at Austin | Austin, TX
Lead GNSS Data Scientist
@ Lurra Systems | Melbourne
Senior Machine Learning Engineer (MLOps)
@ Promaton | Remote, Europe
Senior Machine Learning Engineer
@ Samsara | Canada - Remote