arXiv · 2210.13979
Meta-learning Pathologies from Radiology Reports using Variance Aware Prototypical Networks
Abstract
Large pretrained Transformer-based language models like BERT and GPT have changed the landscape of Natural Language Processing (NLP). However, fine tuning such models still requires a large number of training examples for each target task, thus annotating multiple datasets and training these models on various downstream tasks becomes time consuming and expensive. In this work, we propose a simple extension of the Prototypical Networks for few-shot text classification. Our main idea is to replace the class prototypes by Gaussians and introduce a regularization term that encourages the examples to be clustered near the appropriate class centroids. Experimental results show that our method outperforms various strong baselines on 13 public and 4 internal datasets. Furthermore, we use the class distributions as a tool for detecting potential out-of-distribution (OOD) data points during deployment.
Explore related subjects
Keep this discovery
Arijit Sehanobish, Kawshik Kannan, Nabila Abraham, Anasuya Das, Benjamin Odry. 2022-10-22. Meta-learning Pathologies from Radiology Reports using Variance Aware Prototypical Networks. https://arxiv.org/abs/2210.13979
Cite the original work for its findings. Save a collection to share your selection of sources.