Skip to main content

VocalStyleV1

Vocal style classifier (taxonomy-based). Use this model to describe the characteristics of present vocals in more detail (useful for more specific vocal characterization beyond general vocals presence and gender).

versionVersionrequired
Constant value: VocalStyleV1
Default value: VocalStyleV1
segments objectrequired

Segment-level scores for vocal-style tags.

timestampsSecondsnumber[]required

Possible values: >= 0

values objectrequired

Map of vocal-style tag → array of scores (aligned with timestampsSeconds).

femaleACapellanumber[]required

Possible values: >= 0 and <= 1

femaleChoirnumber[]required

Possible values: >= 0 and <= 1

femaleForegroundVocalsnumber[]required

Possible values: >= 0 and <= 1

femaleBackgroundVocalsnumber[]required

Possible values: >= 0 and <= 1

instrumentalnumber[]required

Possible values: >= 0 and <= 1

maleACapellanumber[]required

Possible values: >= 0 and <= 1

maleChoirnumber[]required

Possible values: >= 0 and <= 1

maleForegroundVocalsnumber[]required

Possible values: >= 0 and <= 1

maleBackgroundVocalsnumber[]required

Possible values: >= 0 and <= 1

mixedACapellanumber[]required

Possible values: >= 0 and <= 1

mixedChoirnumber[]required

Possible values: >= 0 and <= 1

syntheticACapellanumber[]required

Possible values: >= 0 and <= 1

syntheticChoirnumber[]required

Possible values: >= 0 and <= 1

syntheticForegroundVocalsnumber[]required

Possible values: >= 0 and <= 1

syntheticBackgroundVocalsnumber[]required

Possible values: >= 0 and <= 1

scores objectrequired

Track-level scores (0.0-1.0) keyed by vocal-style tag.

femaleACapellaFemale Acapellarequired

Possible values: >= 0 and <= 1

femaleChoirFemale Choirrequired

Possible values: >= 0 and <= 1

femaleForegroundVocalsFemale Foreground Vocalsrequired

Possible values: >= 0 and <= 1

femaleBackgroundVocalsFemale Background Vocalsrequired

Possible values: >= 0 and <= 1

instrumentalInstrumentalrequired

Possible values: >= 0 and <= 1

maleACapellaMale Acapellarequired

Possible values: >= 0 and <= 1

maleChoirMale Choirrequired

Possible values: >= 0 and <= 1

maleForegroundVocalsMale Foreground Vocalsrequired

Possible values: >= 0 and <= 1

maleBackgroundVocalsMale Background Vocalsrequired

Possible values: >= 0 and <= 1

mixedACapellaMixed Acapellarequired

Possible values: >= 0 and <= 1

mixedChoirMixed Choirrequired

Possible values: >= 0 and <= 1

syntheticACapellaSynthetic Acapellarequired

Possible values: >= 0 and <= 1

syntheticChoirSynthetic Choirrequired

Possible values: >= 0 and <= 1

syntheticForegroundVocalsSynthetic Foreground Vocalsrequired

Possible values: >= 0 and <= 1

syntheticBackgroundVocalsSynthetic Background Vocalsrequired

Possible values: >= 0 and <= 1

tagsVocalStyleV1Tags[]required

Possible values: [femaleACapella, femaleChoir, femaleForegroundVocals, femaleBackgroundVocals, instrumental, maleACapella, maleChoir, maleForegroundVocals, maleBackgroundVocals, mixedACapella, mixedChoir, syntheticACapella, syntheticChoir, syntheticForegroundVocals, syntheticBackgroundVocals]

VocalStyleV1
{
"version": "VocalStyleV1",
"segments": {
"timestampsSeconds": [
0
],
"values": {
"femaleACapella": [
0
],
"femaleChoir": [
0
],
"femaleForegroundVocals": [
0
],
"femaleBackgroundVocals": [
0
],
"instrumental": [
0
],
"maleACapella": [
0
],
"maleChoir": [
0
],
"maleForegroundVocals": [
0
],
"maleBackgroundVocals": [
0
],
"mixedACapella": [
0
],
"mixedChoir": [
0
],
"syntheticACapella": [
0
],
"syntheticChoir": [
0
],
"syntheticForegroundVocals": [
0
],
"syntheticBackgroundVocals": [
0
]
}
},
"scores": {
"femaleACapella": 0,
"femaleChoir": 0,
"femaleForegroundVocals": 0,
"femaleBackgroundVocals": 0,
"instrumental": 0,
"maleACapella": 0,
"maleChoir": 0,
"maleForegroundVocals": 0,
"maleBackgroundVocals": 0,
"mixedACapella": 0,
"mixedChoir": 0,
"syntheticACapella": 0,
"syntheticChoir": 0,
"syntheticForegroundVocals": 0,
"syntheticBackgroundVocals": 0
},
"tags": [
"femaleACapella"
]
}