In our new paper in IEEE Access, we use Log-Area Ratios (LARs) + a novel Conditional Speaker Normalization (CSN) conditioned on speaker proxies (e.g. height) to detect synthetic speech reliably on ASVspoof & FoR.

In our new paper in IEEE Access, we use Log-Area Ratios (LARs) + a novel Conditional Speaker Normalization (CSN) conditioned on speaker proxies (e.g. height) to detect synthetic speech reliably on ASVspoof & FoR.
🎙️ Does knowing a speaker’s gender actually boost speaker recognition?
In our latest paper in Soft Computing (Springer), we explore gender effects via bio-inspired filterbanks (Gammatone, Cascade, etc.)
)
🎙️ Are synthetic speech artifacts evenly distributed across phonemes? Not quite!
In our latest paper in MTAP, we propose a phonetic-driven framework using PPGs & phoneme pruning, boosting anti-spoofing performance on ASVspoof 2019 LA.