Speaker independent emotion recognition by early fusion of acoustic and linguistic features within ensembles | Synapse