Grilled Cheese

ExploreLog inSign up
Terms of UsePrivacy PolicyCommunity StandardsHelpGet the app

Grilled Cheese is a product of Village Compute

Version devBuilt at: 2026-10-10 01:38:52 EDT

Explore

PostsPeople
LatestRanked
@shekofteh.bsky.socialOct 6, 2026, 2:37 AM

🎙️ Can phase space dynamics boost ASR?

Most speech systems ignore signal phase. In our paper in SIVP, we use Recurrence Plots + 2D Adaptive Wavelets to capture non-linear dynamics.

📄 Read more: link.springer.com/article/10.1...

#SpeechAI #ASR #SignalProcessing #MachineLearning

@detoxima2025.bsky.socialOct 2, 2026, 1:00 PM

Suno AI 음성 생성 기능 3가지 활용법 공개

https://bit.ly/4j1gwPs

#SunoAI #AI음성 #인공지능 #SpeechAI #AI기술 #ArtificialIntelligence #VoiceGeneration

@shekofteh.bsky.socialSep 25, 2026, 8:15 AM

Reached ~87% test accuracy and ~93% validation accuracy on spontaneous speech with our proposed CNN framework.

Demonstrates that clinical AI needs both acoustic timbre and phonetic context.

Check out the full research here: link.springer.com/article/10.1...

#SpeechAI #Acoustics #DeepLearning

@shekofteh.bsky.socialSep 25, 2026, 8:13 AM

Low-level acoustic features (MFCCs) only tell half the story in voice pathology detection.

What happens when we introduce phonetic-based features (PPPs) into deep learning models?

🧵 Our paper in IJST (Springer) explores this:
link.springer.com/article/10.1...

#SpeechAI #Acoustics #Phonetics

@shekofteh.bsky.socialSep 25, 2026, 8:09 AM

Key Numbers:

• ~85% accuracy on challenging test sets

• ~92% accuracy on evaluation sets

Proving that spontaneous speech isn’t just “noise”—it’s rich in diagnostic

Check out the paper here: link.springer.com/article/10.1...

#SpeechAI #AudioProcessing #MachineLearning

@shekofteh.bsky.socialSep 20, 2026, 4:03 AM

In our new paper in IEEE Access, we use Log-Area Ratios (LARs) + a novel Conditional Speaker Normalization (CSN) conditioned on speaker proxies (e.g. height) to detect synthetic speech reliably on ASVspoof & FoR.

ieeexplore.ieee.org/abstract/doc...

#DeepfakeDetection #SpeechAI #AudioDeepfake

@shekofteh.bsky.socialSep 18, 2026, 6:21 AM

)
🎙️ Are synthetic speech artifacts evenly distributed across phonemes? Not quite!

In our latest paper in MTAP, we propose a phonetic-driven framework using PPGs & phoneme pruning, boosting anti-spoofing performance on ASVspoof 2019 LA.

🔗 link.springer.com/article/10.1...

#SpeechAI #AudioDeepfake

@9cv9recruitment.bsky.socialSep 16, 2026, 5:50 PM

Discover Meta Muse Voice Transcribe 1.0, how its real-time speech AI works, key features, performance, pricing and use cases.

blog.9cv9.com/what-is-meta...

#Meta, #MuseVoiceTranscribe, #MuseVoiceTranscribe1, #MetaAI, #MetaAIResearch, #VoiceAI, #SpeechAI,