This is a leaderboard for multimodal emotion recognition on the IEMOCAP dataset. The modality abbreviations are A: Acoustic T: Text V: Visual Please include the modality in the bracket after the model name. All models must use standard five emotion categories and are evaluated in standard leave-one-session-out (LOSO). See the papers for references.