AIIC AI Intelligence Centre

SOURCE-LINKED INTELLIGENCE

Face-voice Association across LAnguages and Gender (FLAG) 2027 Challenge Evaluation Plan

arXiv · AI, language, vision and robotics · article · Sep 15, 2026 · UTC

Face--voice association models may rely on language or gender cues in the voice rather than on speaker-specific voice characteristics, which can lead to a performance deterioration when the model has to identify a multilingual speaker or distinguis same-gender speakers. To investigate these issues, we introduce the Face-voice Association across LAnguages and Gender (FLAG) 2027 Challenge. The challenge formulates face--voice association as a cross-modal verification task: given a voice, identify the speaker's face from a ``gallery'' of faces consisting of the speaker's face and a set of negativ

Read original source ↗ Open in workspace

recordType
paper
region
Global

Evidence & attribution

First collected: 2026-09-20T08:20:57.646Z. This is not the publication date.