Latest Research in Speech and Multimodal Analytics
844 research papers · 0.0 average citations · 2026 median publication year
Top Research Topics in Speech and Multimodal Analytics
- Sound — 154 papers
- Intelligence, Security, War Strategy — 132 papers
- Audio and Speech Processing — 126 papers
- Computer Vision and Pattern Recognition — 120 papers
- Computation and Language — 104 papers
- Emotion and Mood Recognition — 33 papers
- Speech Recognition and Synthesis — 26 papers
- Artificial Intelligence — 24 papers
- Machine Learning — 16 papers
- Music and Audio Processing — 13 papers
Highest-Cited Papers
- DOTA-ME-CS: daily oriented text audio-Mandarin English-Code switching dataset (1 citations)
- A dual-domain guided emotion-specialized swin transformer for enhancing speech emotion recognition
- EuroMillions Breakthrough Mining — 4317 Sources, 38847 Leads Extracted — E8 Intelligence Research
- Music Genre Classification Algorithm Based on Multi-Scale Fusion
- Towards inclusive voice biometrics: Dysarthria-discriminative embeddings for ASV system
- EuroMillions Breakthrough Mining — 4348 Sources, 39103 Leads Extracted — E8 Intelligence Research
- EuroMillions Breakthrough Mining — 4348 Sources, 39103 Leads Extracted — E8 Intelligence Research
- EuroMillions Breakthrough Mining — 4317 Sources, 38847 Leads Extracted — E8 Intelligence Research
- Event-Grounded Football News Generation from Match Videos with Parameter-Efficient Large Language Models
- Speech intelligibility comparison of standalone and two-stage deep learning architectures for behind-the-ear-to-binaural enhancement
- EuroMillions Breakthrough Mining — 4271 Sources, 38419 Leads Extracted — E8 Intelligence Research
- EuroMillions Breakthrough Mining — 4268 Sources, 38386 Leads Extracted — E8 Intelligence Research
- Reading User Distress From the Inside: Do a Language Model's Hidden States Track a Person's Distress Better Than the Model's Own Reply Does?
- A2SSC: An Agent-based Adaptive Semantic Speech Communication System
- Adaptive learning behavior recognition using multimodal transformer based dynamic modality regulation
- EuroMillions Breakthrough Mining — 4288 Sources, 38586 Leads Extracted — E8 Intelligence Research
- EuroMillions Breakthrough Mining — 4269 Sources, 38397 Leads Extracted — E8 Intelligence Research
- Dual-stream audio–visual Transformer fusion for harmful video content detection
- DESED and DataSED precomputed caches for Cover First, Disagree Softly: Rethinking Mismatch-First Active Learning for Frame-Level Audio Classification
- EuroMillions Breakthrough Mining — 4271 Sources, 38419 Leads Extracted — E8 Intelligence Research
Sub-Regions
- Conversational Speech Processing — 360 papers
- Strategic Intelligence Mining — 134 papers
- Multimodal Video Perception — 118 papers
- Computational Music Audio — 108 papers
- Affective Multimodal Recognition — 96 papers