Nlp Highlights

128 - Dynamic Benchmarking, with Douwe Kiela

Autor: Vários
Narrador: Vários
Editora: Podcast
Duração: 0:47:00
Mais informações

Adicionar à lista

Ouvir

preview

Ouvir

Sinopse

We discussed adversarial dataset construction and dynamic benchmarking in this episode with Douwe Kiela, a research scientist at Facebook AI Research who has been working on a dynamic benchmarking platform called Dynabench. Dynamic benchmarking tries to address the issue of many recent datasets getting solved with little progress being made towards solving the corresponding tasks. The idea is to involve models in the data collection loop to encourage humans to provide data points that are hard for those models, thereby continuously collecting harder datasets. We discussed the details of this approach, and some potential caveats. We also discussed dynamic leaderboards, a recent addition to Dynabench that rank systems based on their utility given specific use cases. Papers discussed in this episode: 1. Dynabench: Rethinking Benchmarking in NLP (https://www.semanticscholar.org/paper/Dynabench%3A-Rethinking-Benchmarking-in-NLP-Kiela-Bartolo/77a096d80eb4dd4ccd103d1660c5a5498f7d026b) 2. Dynaboard: An Evaluation-As

Mostrar mais

Nlp Highlights

128 - Dynamic Benchmarking, with Douwe Kiela

Sinopse

Experimente 7 dias grátis

Precisando de ajuda?

Instale o aplicativo:

Nlp Highlights

128 - Dynamic Benchmarking, with Douwe Kiela

Informações:

Sinopse

Experimente 7 dias grátis

Precisando de ajuda?

Instale o aplicativo: