arXiv Preprint
|
2023
Menghini et al.
Abstract
The paper explores the use of pseudolabels, which are heuristic labels for unlabeled data, to enhance the performance of vision-language models like CLIP via prompt tuning. The authors investigate different learning paradigms and prompt modalities and find that iterative prompt-training strategies leveraging CLIP-based pseudolabels lead to significant improvements in CLIP’s image classification performance.
Coming Fall 2026
A one-day, invite-only summit providing a first look at the benchmarks and research that will shape the frontier. Sign up for updates.