A Method for Curation of Web-Scraped Face Image Datasets

Web-scraped, in-the-wild datasets have become the norm in face recognition research. The numbers of subjects and images acquired in web-scraped datasets are usually very large, with number of images on the millions scale. A variety of issues occur when collecting a dataset in-the-wild, including images with the wrong identity label, duplicate images, duplicate subjects and variation in quality.…

Paper

Similar papers

© 2026 NYSGPT2525 LLC