Application of synchrosqueezed spectrograms to speaker verification via formant identification by visual observation

Speaker identification and verification via the measurement of formant frequencies from a sound spectrogram are often based on rules of thumb. In this Letter, we investigate the effectiveness of synchrosqueezed spectrograms, which can make formants clearer, for speaker verification using an experiment involving human participants. In the experiment, 88 participants were presented with pairs of synchrosqueezed spectrograms and asked to determine whether each pair was produced by the same speaker or different speakers. It was found that synchrosqueezing improved the performance of speaker discrimination by significantly reducing mistakes for different speakers and is thus useful for speaker verification using spectrograms.

Authors

Institutions

Publication Details

Journal
JASA Express Letters
Published
2026-10-01
DOI
https://doi.org/10.1121/10.0046805
Primary Topic
Speech and Audio Processing
Type
article
Field-Weighted Citation Impact
0.00
Controls
|||
ALL TIME
JAN
FEB
MAR
APR
MAY
JUN
JUL
AUG
SEP
article

Application of synchrosqueezed spectrograms to speaker verification via formant identification by visual observation

Masahiro Okada
JASA Express Letters
Speech and Audio Processing
article

Application of synchrosqueezed spectrograms to speaker verification via formant identification by visual observation

Masahiro Okada
article en

Abstract

Speaker identification and verification via the measurement of formant frequencies from a sound spectrogram are often based on rules of thumb. In this Letter, we investigate the effectiveness of synchrosqueezed spectrograms, which can make formants clearer, for speaker verification using an experiment involving human participants. In the experiment, 88 participants were presented with pairs of synchrosqueezed spectrograms and asked to determine whether each pair was produced by the same speaker or different speakers. It was found that synchrosqueezing improved the performance of speaker discrimination by significantly reducing mistakes for different speakers and is thus useful for speaker verification using spectrograms.

JASA Express LettersVol. 6(10)
National Research Institute of Police Science (JP)
Reduced inequalities
Openalex Percentile: Top 10%
Speech and Audio Processing
AI Navigator

Ask Laika to Summarize, Analyze, and Connect papers live on the map.

Summarize Papers & Methodologies

Extract key findings, datasets, and comparative methods across publications.

Benchmark Rankings & Visual Analytics

Rank top research institutions, authors, funders, topics, and journals by Field-Weighted Citation Impact (FWCI) and paper volume with instant charts.

Connect Distant Disciplines

Bridge topological clusters on the map to find hidden collaborative intersections.