Glossary
Speaker metadata
Structured information about who is speaking — accent, age range, gender, languages, and role — collected with consent. It lets buyers assemble demographically balanced datasets.
Typical fields: accent or dialect, age band, gender, native and other languages, role in the recording, and recording setup. It is collected by self-report at consent time, not inferred afterwards — inference is unreliable and raises privacy problems of its own. Aggregated, it becomes the demographic table in the data card.
Why it matters
Metadata is what lets a buyer assemble a balanced dataset instead of an unknown mix.