Auditing vision-language models (VLMs) for societal bias requires distinguishing direct algorithmic valuation disparities from confounders embedded within archival metadata. In this study, we audit Contrastive Language-Image Pretraining (CLIP) models using historical artwork…
Read the original source — arxiv.org
paper · Shared by tscosj
0 comments
No comments yet.