Fairness benchmark for multimodal LLMs: dataset, image perturbations, and results (paper: Fairness Failure Modes of Multimodal LLMs).
AI & ML interests
None defined yet.
Recent Activity
datasets 11
MLL-Lab/MultiBBQ-results
Preview • Updated • 372
MLL-Lab/MultiBBQ-realworld
Viewer • Updated • 79 • 242
MLL-Lab/MultiBBQ-perturbations
Viewer • Updated • 8.59k • 1.15k
MLL-Lab/MultiBBQ
Viewer • Updated • 1.64k • 196
MLL-Lab/viewsuite
Updated • 225 • 1
MLL-Lab/BAGEN
Viewer • Updated • 445 • 97 • 3
MLL-Lab/MindTopo
Viewer • Updated • 8.91k • 91 • 2
MLL-Lab/Theory-of-Space
Viewer • Updated • 13.7k • 256 • 6
MLL-Lab/ENACT
Preview • Updated • 91 • 7
MLL-Lab/MindCube
Viewer • Updated • 4.28k • 1.05k • 11