Grilled Cheese

ExploreLog inSign up
Terms of UsePrivacy PolicyCommunity StandardsHelpGet the app

Grilled Cheese is a product of Village Compute

Version devBuilt at: 2026-10-10 20:13:24 EDT

Explore

PostsPeople
LatestRanked
@gradientbrief.bsky.socialOct 9, 2026, 2:00 PM

A new arXiv study finds that pretrained self-supervised speech models like Wav2Vec2 and HuBERT, despite training data skewed toward high-resource languages, can recognize click consonants from Khoisan languages…

#AI #SpeechRecognition #NLP #LowResourceLanguages
https://arxiv.org/abs/2606.11542

@freegardener.bsky.socialOct 9, 2026, 10:25 AM

Streaming ASR looks fluent but secretly gambles on every character. A new single-pass model decides when to commit vs. what to commit—no revisions, no global #SpeechRecognition #ASR #MachineLearning #AI

https://freegardner.com/synapse/machine-that-says-it-is-listening.html

R@realfeedapp.comOct 9, 2026, 7:20 AM

The model transcribes audio, provides word timestamps, and generates speech embeddings, operating within a single 16.9 MB file on the CPU.

#SpeechRecognition #LLM #ONDevice

@ai-news.at.thenote.appOct 3, 2026, 11:55 PM

Testing Android’s On-Device Speech Recognition With Hindi-English Code-Switching

On-device speech recognition turned "dizzy" into डीसी for my Hindi-speaking mother. The one-flag fix, a crashing OnePlus, and a reject button.

Telegram AI Digest
#ai #speechrecognition #testing

@ai-ru.at.thenote.appOct 3, 2026, 11:45 PM

Тестирование распознавания речи на устройстве Android с переключением между хинди и английским языками

Распознавание речи на устройстве превратило «головокружение» в डीसी для моей мамы, говорящей на хинди. Исправление с одним флагом, сбой OnePlus…

Telegram ИИ Дайджест
#ai #speechrecognition #testing

@ossradarai.bsky.socialSep 29, 2026, 12:01 AM

Berkeley's HuPER models phonetic perception as adaptive inference, hitting state-of-the-art English error rates with just 100 hours of training and zero-shot transfer to 95 languages. Code and models are open-sourced on GitHub.

#OpenSourceAI #NLP #SpeechRecognition
https://arxiv.org/abs/2602.01634

@martincid.comSep 25, 2026, 3:15 PM

Gemini 3.8 Live removes the blank face from AI customer service #CustomerService #SpeechRecognition #ArtificialIntelligence

@achronicvoice.comSep 22, 2026, 8:30 PM

“In a room with normal background #noise ..my #SpeechRecognition without #HearingAids is roughly 40–50%. That means I miss about half of what’s being said and have to guess the rest from context.”: buff.ly/hiTT9jz

by @donnietownstudio.bsky.social
#conversation #deaf #disability #disabled #spoonie

@lexacom.bsky.socialSep 21, 2026, 2:33 PM

Simplify your work with smart workflows, ambient voice technology, and medical speech recognition on the go.

🔎 Search Lexacom on the App Store and Google Play

#nhs #digitaltransformation #workflows #productivity #ambientAI #speechrecognition

@lexacom.bsky.socialSep 14, 2026, 9:31 AM

Simplified, intuitive, and packed with new features, download our new app.

Intelligent workflows, ambient voice technology, and speech recognition on the go.

🔎 Search Lexacom on the App Store and Google Play.

#nhs #digitaltransformation #workflows #productivity #ambientAI #speechrecognition

@ai-news.at.thenote.appSep 12, 2026, 1:25 AM

Why Speech Recognition Misses Human Context: Dr. Sunday David Ubur’s Affective Architecture

Dr. Sunday Ubur's research explores emotion-aware AI captions that preserve tone, urgency, and context for deaf and hard-of-hearing users.

Telegram AI Digest
#ai #news #speechrecognition

@ai-ru.at.thenote.appSep 12, 2026, 1:15 AM

Почему распознавание речи упускает человеческий контекст: Аффективная архитектура доктора Сандея Дэвида Убура

Telegram ИИ Дайджест
#ai #news #speechrecognition

@cryptonews-poster.bsky.socialSep 11, 2026, 6:23 AM

Xiaomi open-sources Xiaomi-CocktailASR-1, an industrial-grade speaker diarization LLM! Solves the "cocktail party problem," precisely identifying and transcribing target speakers in multi-speaker audio using a reference voiceprint. Big leap for speech tech! #AI #OpenSource #SpeechRecognition

@hackernoon.comSep 11, 2026, 3:52 AM

Dr. Sunday Ubur's research explores emotion-aware AI captions that preserve tone, urgency, and context for deaf and hard-of-hearing users. #speechrecognition

@everycivicapp.bsky.socialSep 8, 2026, 4:52 PM

aeneas aeneas is a Python/C library and a set of tools to automagically synchronize audio and text (aka forced alignment) https://github.com/readbeyond/aeneas #GeneralPurposeOpenSourceTools #SpeechRecognition #CivicTech

@lexacom.bsky.socialSep 7, 2026, 3:57 PM

What do 152,416 words in a week mean to a GP?

For Dr Rafay, it means using Lexacom across all consultations, referrals and admin - saving an hour a day.

➡️ Read the full case study and start your free trial at lexacom.co.uk/faster

#nhs #generalpractice #speechrecognition #healthtech #productivity

@flo7up.bsky.socialSep 7, 2026, 8:12 AM

Heidi fine-tuned and deployed medical speech recognition for clinical documentation, using synthetic multilingual data, distributed GPU training, and scalable AWS serving. The platform supports 2.4M+ consultations weekly. #HealthcareAI #SpeechRecognition #ClinicalDocumentation

@aidailypost.comSep 3, 2026, 2:16 PM

Microsoft’s new MAI‑Transcribe‑2 slashes cost and cuts transcription time in half—outpacing OpenAI’s frontier models on call‑center audio and multilingual support. Curious how fast and cheap AI can get? Dive in. #MAITranscribe2 #SpeechRecognition #MicrosoftAI

🔗 aidailypost.com/news/microso...