Human + source
~20 fields as a baseline
Annotations supplied with the content. Exported fields vary by content type and catalog.
Musical, descriptive and technical fields for your workflow, with human annotations and machine enrichment identified separately.
01 / Provenance
A field’s origin helps you decide how to use it. Keep source annotations and machine-derived attributes identifiable in the same schema.
Human + source
Annotations supplied with the content. Exported fields vary by content type and catalog.
Machine enrichment
Automated attributes, available or added after ingestion. Coverage varies; these are machine-derived estimates, not human ground truth.
Descriptor vocabulary: 10,000+ tags. Subgenre taxonomy: ~3,000 subgenres. These describe the vocabulary and taxonomy, not the number of tags attached to every recording.
02 / Field families
Available fields can support selection, search, classification and model conditioning. The field set and coverage are defined for each dataset.
03 / A usable schema
Use consistent labels. Demo 1 displays its source key as D minor.
Track IDs, files and associated stems connect the asset records. Technical properties remain distinct from descriptive annotations.
Master CSV and/or JSON records accompany the agreed content. Confirm required fields, provenance and missing-value handling as part of the brief.
Work with ERISV
Bring the labels, source requirements and output structure you want. We’ll confirm available coverage and any selected enrichment.