Files changed (1) hide show
  1. README.md +10 -4
README.md CHANGED
@@ -109,11 +109,17 @@ https://github.com/NLP-UMUTeam/umuteam-speech-emotion
109
 
110
  The model was evaluated on the Spanish held-out test set used in the `speech-emotion` toolkit.
111
 
112
- | Language | Mode | Accuracy | Weighted Precision | Weighted F1 | Macro F1 |
113
- |---|---:|---:|---:|---:|---:|
114
- | Spanish | Text | 77.0204 | 77.0449 | 76.8367 | 69.3886 |
115
 
116
- These results correspond to the text-only Spanish configuration. In the full toolkit, multimodal configurations combining audio and text obtain higher performance, showing the benefit of integrating acoustic and linguistic information.
 
 
 
 
 
 
 
 
117
 
118
  ## How to use
119
 
 
109
 
110
  The model was evaluated on the Spanish held-out test set used in the `speech-emotion` toolkit.
111
 
112
+ ### Performance comparison on Spanish emotion recognition
 
 
113
 
114
+ | Configuration | Accuracy | Weighted Precision | Weighted F1 | Macro F1 |
115
+ |---|---:|---:|---:|---:|
116
+ | Speech-only | 88.1207 | 88.3244 | 88.1357 | 84.4829 |
117
+ | Text-only | 77.0204 | 77.0449 | 76.8367 | 69.3886 |
118
+ | Multimodal (Concat) | **90.0682** | **90.2048** | **90.0642** | **87.7455** |
119
+ | Multimodal (Mean) | 88.5102 | 88.6163 | 88.5011 | 84.1653 |
120
+ | Multimodal (Multihead) | 82.6680 | 82.3820 | 82.4600 | 75.5606 |
121
+
122
+ These results show that text-only emotion recognition is effective for Spanish emotion analysis, although multimodal approaches combining acoustic and linguistic representations achieve higher overall performance.
123
 
124
  ## How to use
125