Correctness consistency

67.6% passed tests (25 passed / 12 failed).

Samples In Expected High Range

Proportion of samples whose predictions fall into the expected value range of >= 0.55

Threshold: 0.75

Data

anger

fear

surprise

crema-d-1.2.0-emotion.categories.test.gold_standard

0.76

0.41

danish-emotional-speech-1.1.1-emotion.test

0.42

0.21

emodb-1.2.0-emotion.categories.test.gold_standard

1.00

0.91

emovo-1.2.1-emotion.test

0.93

0.54

0.60

iemocap-2.3.0-emotion.categories.test.gold_standard

0.87

0.82

meld-1.3.1-emotion.categories.test.gold_standard

0.95

0.90

0.84

polish-emotional-speech-1.1.1-emotion.categories.test.gold_standard

0.92

0.42

ravdess-1.1.2-emotion.speech.test

0.91

0.62

0.97

mean

0.84

0.66

0.66

Samples In Expected Low Range

Proportion of samples whose predictions fall into the expected value range of <= 0.45

Threshold: 0.75

Data

boredom

sadness

crema-d-1.2.0-emotion.categories.test.gold_standard

0.90

danish-emotional-speech-1.1.1-emotion.test

1.00

emodb-1.2.0-emotion.categories.test.gold_standard

0.83

1.00

emovo-1.2.1-emotion.test

0.92

iemocap-2.3.0-emotion.categories.test.gold_standard

0.84

meld-1.3.1-emotion.categories.test.gold_standard

0.25

polish-emotional-speech-1.1.1-emotion.categories.test.gold_standard

0.90

1.00

ravdess-1.1.2-emotion.speech.test

0.81

mean

0.86

0.84

Samples In Expected Neutral Range

Proportion of samples whose predictions fall into the expected value range of [0.3, 0.6]

Threshold: 0.75

Data

neutral

crema-d-1.2.0-emotion.categories.test.gold_standard

0.44

danish-emotional-speech-1.1.1-emotion.test

0.23

emodb-1.2.0-emotion.categories.test.gold_standard

0.89

emovo-1.2.1-emotion.test

0.86

iemocap-2.3.0-emotion.categories.test.gold_standard

0.79

meld-1.3.1-emotion.categories.test.gold_standard

0.55

polish-emotional-speech-1.1.1-emotion.categories.test.gold_standard

0.90

ravdess-1.1.2-emotion.speech.test

0.50

mean

0.65

Visualization

Distribution of dimensional model predictions for samples with different categorical emotions. The expected range of model predictions is highlighted by the green brackground.

../../../_images/visualization_crema-d-1.2.0-emotion.categories.test.gold_standard6.png
../../../_images/visualization_danish-emotional-speech-1.1.1-emotion.test6.png
../../../_images/visualization_emodb-1.2.0-emotion.categories.test.gold_standard6.png
../../../_images/visualization_emovo-1.2.1-emotion.test6.png
../../../_images/visualization_iemocap-2.3.0-emotion.categories.test.gold_standard6.png
../../../_images/visualization_meld-1.3.1-emotion.categories.test.gold_standard6.png
../../../_images/visualization_polish-emotional-speech-1.1.1-emotion.categories.test.gold_standard6.png
../../../_images/visualization_ravdess-1.1.2-emotion.speech.test6.png