Computer Science > Computer Vision and Pattern Recognition

arXiv:2211.05809 (cs)

[Submitted on 10 Nov 2022]

Title:Casual Conversations v2: Designing a large consent-driven dataset to measure algorithmic bias and robustness

Authors:Caner Hazirbas, Yejin Bang, Tiezheng Yu, Parisa Assar, Bilal Porgali, Vítor Albiero, Stefan Hermanek, Jacqueline Pan, Emily McReynolds, Miranda Bogen, Pascale Fung, Cristian Canton Ferrer

View PDF

Abstract:Developing robust and fair AI systems require datasets with comprehensive set of labels that can help ensure the validity and legitimacy of relevant measurements. Recent efforts, therefore, focus on collecting person-related datasets that have carefully selected labels, including sensitive characteristics, and consent forms in place to use those attributes for model testing and development. Responsible data collection involves several stages, including but not limited to determining use-case scenarios, selecting categories (annotations) such that the data are fit for the purpose of measuring algorithmic bias for subgroups and most importantly ensure that the selected categories/subcategories are robust to regional diversities and inclusive of as many subgroups as possible.
Meta, in a continuation of our efforts to measure AI algorithmic bias and robustness (this https URL), is working on collecting a large consent-driven dataset with a comprehensive list of categories. This paper describes our proposed design of such categories and subcategories for Casual Conversations v2.

Subjects:	Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Computation and Language (cs.CL); Computers and Society (cs.CY)
Cite as:	arXiv:2211.05809 [cs.CV]
	(or arXiv:2211.05809v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2211.05809

Submission history

From: Caner Hazirbas [view email]
[v1] Thu, 10 Nov 2022 19:06:21 UTC (104 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Casual Conversations v2: Designing a large consent-driven dataset to measure algorithmic bias and robustness

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Casual Conversations v2: Designing a large consent-driven dataset to measure algorithmic bias and robustness

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators