Fetching the paper…
Reading the bibliography…
Multi-label classification problems with thousands of classes are hard to solve with in-context learning alone, as language models (LMs) might lack prior knowledge about the precise classes or how to assign them, and it is generally infeasible to demonstrate every class in a prompt.
The medical dictionary for regulatory activities (meddra)
Elliot G Brown, Louise Wood, and Sue Wood. 1999 · 1999
Earlier work this paper cites.
A geometric interpretation of r-precision and its correlation with average precision
Javed A Aslam, Emine Yilmaz, and Virgiliu Pavlu. 2005 · 2005
Earlier work this paper cites.
The extreme classification repository: Multi-label datasets and code
K. Bhatia, K. Dahiya, H. Jain, P. Kar, A. Mittal, Y. Prabhu, and M. Varma. 2016 · 2016
Earlier work this paper cites.
ESCO, European skills, competences, qualifications and occupations
European Commission Directorate-General for Employment, Social Affairs and Inclusion. 2017 · 2017
Earlier work this paper cites.
Sentence-BERT: Sentence embeddings using Siamese BERT-networks
Nils Reimers and Iryna Gurevych. 2019 · 2019
Earlier work this paper cites.
On the dangers of stochastic parrots: Can language models be too big?
Emily M Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell. 2021 · 2021
Earlier work this paper cites.
Scaling instruction-finetuned language models
Hyung Won Chung, Le Hou, Shayne Longpre, Barret Zoph, Yi Tay, William Fedus, Yunxuan Li, Xuezhi Wang, Mostafa Dehghani, Siddhartha Brahma, et al. 2022 · 2022
Cited alongside, same era.
Design of negative sampling strategies for distantly supervised skill extraction
Jens-Joris Decorte, Jeroen Van Hautte, Johannes Deleu, Chris Develder, and Thomas Demeester. 2022 · 2022
Cited alongside, same era.
BioLORD: Learning ontological representations from definitions for biomedical concepts and their textual descriptions
François Remy, Kris Demuynck, and Thomas Demeester. 2022 · 2022
Cited alongside, same era.
SkillSpan: Hard and soft skill extraction from English job postings
Mike Zhang, Kristian Jensen, Sif Sonniks, and Barbara Plank. 2022 · 2022
Cited alongside, same era.
Large language models as batteries-included zero-shot ESCO skills matchers
Benjamin Clavié and Guillaume Soulié. 2023 · 2023
Extreme multi-label skill extraction training using large language models
Jens-Joris Decorte, Severine Verlinden, Jeroen Van Hautte, Johannes Deleu, Chris Develder, and Thomas Demeester. 2023 · 2023
Later among the works it cites.
BioDEX: Large-scale biomedical adverse drug event extraction for real-world pharmacovigilance
Karel D’Oosterlinck, François Remy, Johannes Deleu, Thomas Demeester, Chris Develder, Klim Zaporojets, Aneiss Ghodsi, Simon Ellershaw, Jack Collins, and Christopher Potts. 2023 · 2023
Later among the works it cites.
DSPy: Compiling declarative language model calls into self-improving pipelines
Omar Khattab, Arnav Singhvi, Paridhi Maheshwari, Zhiyuan Zhang, Keshav Santhanam, Sri Vardhamanan, Saiful Haq, Ashutosh Sharma, Thomas T. Joshi, Hanna Moazam, Heather Miller, Matei Zaharia, and Christopher Potts. 2023 · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
IDAS: Intent discovery with abstractive summarization
Maarten De Raedt, Fréderic Godin, Thomas Demeester, and Chris Develder. 2023 · 2023
Cited alongside, same era.
Yaxin Zhu and Hamed Zamani. 2023 · 2023
Later among the works it cites.