| [1] |
SÖNMEZ Y Ü, VAROL A. In-Depth Investigation of Speech Emotion Recognition Studies from Past to Present -The Importance of Emotion Recognition from Speech Signal for AI-[J]. Intelligent Systems with Applications, 2024:200351.
|
| [2] |
KHAN M, GUEAIEB W, EL SADDIKA, et al. MSER:Multimodal Speech Emotion Recognition Using Cross-Attention with Deep Fusion[J]. Expert Systems with Applications, 2024:122946.
|
| [3] |
SINGH J, SAHEER L B, FAUST O. Speech Emotion Recognition Using Attention Model[J]. International Journal of Environmental Research and Public Health, 2023, 20(6):5140.
|
| [4] |
CHEN W D, XING X F, CHEN P H, et al. Vesper:A Compact and Effective Pretrained Model for Speech Emotion Recognition[J]. IEEE Transactions on Affective Computing, 2024(15):1711-1724.
|
| [5] |
ZHANG K X, WEN Q S, ZHANG C L, et al. Skip-Step Contrastive Predictive Coding for Time Series Anomaly Detection[C]// ICASSP 2024-2024 IEEE International Conference on Acoustics,Speech and Signal Processing(ICASSP). Piscataway:IEEE, 2024:7065-7069.
|
| [6] |
ZHENG C Y, SALAKHUTDINOV R, EYSENBACH B. Contrastive Difference Predictive Coding(2023)[J/OL]. [2024-02-26]. https://arxiv.org/abs/2310.20141.
|
| [7] |
LEE S, PARK T, LEE K. Soft Contrastive Learning for Time Series[C]// The Twelfth International Conference on Learning Representations(ICLR 2024). Piscataway:IEEE, 2024:1-25.
|
| [8] |
MIKOLOV T, CHEN K, CORRADO G, et al. Efficient Estimation of Word Representations in Vector Space[C]// International Conference on Learning Representations. Piscataway:IEEE, 2013:1-12.
|
| [9] |
KIROS R, ZHU Y K, SALAKHUTDINOV R R, et al. Skip-thought Vectors[C]// In Advances in Neural Information Processing Systems. San Diego: NIPS, 2015:3294-3302.
|
| [10] |
RADFORD A, JOZEFOWICZ R, SUTSKEVER I. Learning to Generate Reviews and Discovering Sentiment(2017)[J/OL]. [2024-04-06]. https://arxiv.org/abs/1704.01444.
|
| [11] |
FAN R Z, POGG M, MATTOCCIA S. Contrastive Learning for Depth Prediction[C]// Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition(CVPR) Workshops. Piscataway:IEEE, 2023:3226-3237.
|
| [12] |
ZHENG R J, WANG X Y, UN Y C, et al. TACO:Temporal Latent Action-Driven Contrastive Loss for Visual Reinforcement Learning[C]// Advances in Neural Information Processing Systems 36(NeurIPS 2023). San Diego: NIPS, 2023:48203-48225.
|
| [13] |
ZHANG R, ISOLA P, EFROS A A. Colorful Image Colorization[C]// European Conference on Computer Vision. Berlin:Springer, 2016:649-666.
|
| [14] |
OORD A V D, LI Y Z, VINYALS O. Representation Learning with Contrastive Predictive Coding(2018)[J/OL]. [2024-02-18]. https://arxiv.org/abs/1807.03748.
|
| [15] |
WISKOTT L, SEJNOWSKI T J. Slow Feature Analysis:Unsupervised Learning of Invariances[J]. Neural Computation, 2002, 14(4):715-770.
|
| [16] |
GUTMANN M, HYVÄRINEN A. NOISE-CONTRASTIVE ESTIMATION:A New Estimation Principle for Unnormalized Statistical Models[C]// Proceedings of the Thirteenth International Conference on Artificial Intelligence and Statistics. Piscataway:IEEE, 2010:297-304.
|
| [17] |
BENGIO Y, SENECAL J S. Adaptive Importance Sampling to Accelerate Training of a Neural Probabilistic Language Model[J]. IEEE Transaction Neural Networks, 2008, 19(4):713-722.
|
| [18] |
CHO K, MERRIËNBOER B V, GULCEHRE C, et al. Learning Phrase Representations Using RNN Encoder Decoder for Statistical Machine Translation (2014)[J/OL]. [2023-09-12]. https://arxiv.org/pdf/1406.1078.
|
| [19] |
OORD A V D, DIELEMAN S, ZEN H, et al. WAVENET:A Generative Model for Raw Audio (2016)[J/OL]. [2024-05-11]. https://arxiv.org/abs/1609.03499.
|
| [20] |
VASWANI A, SHAZEER N, PARMAR N, et al. Attention is All You Need[C]// Advances in Neural Information Processing Systems 30.San Diego:NIPS, 2017:5998-6008.
|
| [21] |
ZHAO Y, LIAO X, HE X, et al. Accelerated Primal-Dual Mirror Dynamics for Centralized and Distributed Constrained Convex Optimization Problems[J]. Journal of Machine Learning Research, 2023, 24(343):1-59.
|
| [22] |
ZHAO Y, HE X, ZHOU M, et al. Accelerated Primal-Dual Projection Neurodynamic Approach with Time Scaling for Linear and Set Constrained Convex Optimization Problems[J]. IEEE/CAA Journal of Automatica Sinica, 2024, 11(6):1485-1498.
|
| [23] |
BELGHAZI M I, BARATIN A, RAJESWAR S, et al. Mine:Mutual Information Neural Estimation(2018)[J/OL]. [2023-10-28]. https://arxiv.org/pdf/1801.04062.
|
| [24] |
ZHAO J F, MAO X, CHEN L J. Speech Emotion Recognition Using Deep 1D & 2D CNN LSTM Networks[J]. Biomedical Signal Processing and Control, 2019, 47:312-323.
|
| [25] |
LIM H, PARK C, MYUNG H. RONet:Real-Time Range-only Indoor Localization via Stacked Bidirectional LSTM with Residual Attention[C]// 2019 IEEE/RSJ International Conference on Intelligent Robots and Systems(IROS). Piscataway:IEEE, 2019:3241-3247.
|
| [26] |
BAUTISTA J L, LEE Y K, SHIN H S. Speech Emotion Recognition Based on Parallel CNN-Attention Networks with Multi-Fold Data Augmentation[J]. Electronics, 2022, 11(23):3935.
|