MMAU

dataset·general audio reasoning·active

A multi-domain benchmark of expert audio understanding and reasoning spanning speech, environmental sound and music.

Recorded facts

Official sitehttps://sakshi113.github.io/mmau_homepage/
GeographyUnited States
sizeskills: 27 · questions: 10000 · source corpora: 13
scopeInformation extraction and complex reasoning across speech, sound and music.
genresmultiple
licensemixed/source-specific
stewardUniversity of Maryland-led collaboration
creatorsMMAU authors
languagesprimarily English questions; source audio varies
modalitiesaudio; multiple-choice questions; answers; skill taxonomy
geographiesGlobal
access methodOfficial project links code and data.
intended usesaudio-language model evaluation; audio reasoning evaluation
consent claimsInherited from source corpora; annotator terms not verified.
related papersMMAU: A Massive Multi-Task Audio Understanding and Reasoning Benchmark
annotation methodExpert-oriented question design with answer choices and skill labels.
collection methodQuestions and audio are assembled across thirteen source corpora.
provenance claimsOfficial page documents source domains and benchmark organization.
current availabilityProject page and artifacts available.
commercial use limitsDepends on each source corpus.
us market scope basisPublicly available or materially used in US-facing music/audio-AI research.
limitations biases disputesMultiple-choice accuracy is not equivalent to open-ended auditory competence.; Music is only one of three domains.

Sources & changes

Checked 10d ago · highhow verification works
  • Status unknown → active
  • Official URL unknown → sakshi113.github.io/mmau_homepage
  • Record maintenance · 12 fields updated

Is this yours? Claim this record →·See something wrong? Report a correction →