Whissle AI Blog

Explore the latest in AI, machine learning, and voice technology.

platformWhissle Is a Whistle: Raw Streams In, Live Structured Intelligence Out

Whissle Is a Whistle: Raw Streams In, Live Structured Intell

A whistle turns a stream of breath into structured sound — instantly, while you're still exhaling. That's exactly what W

By Karan Singla

Jul 03 2026

deception detectionCatching Lies Without Sending the Video: Privacy-Preserving Multimodal Deception Detection

Catching Lies Without Sending the Video: Privacy-Preserving

A widely-cited 75% on this courtroom dataset is inflated by speaker leakage — honest, speaker-independent evaluation is

By Nikita Sharma, Pranav Saran, Karan Singla

Jun 20 2026

ASRGujlish META-ASR: A Bilingual English-Gujarati Speech Model with Built-in Speaker Profiling

Gujlish META-ASR: A Bilingual English-Gujarati Speech Model

We open-source a bilingual English-Gujarati speech recognition model that transcribes speech while extracting speaker ag

By Whissle Research Team

May 16 2026

ASRBuilding a Hindi Speech Model That Understands Context

Building a Hindi Speech Model That Understands Context

We benchmark our META-ASR Hindi model against Deepgram Nova-2 and Gemini 2.5 Flash across two test sets, measuring not j

By Whissle Research Team

May 03 2026

browserIntroducing Whissle Browser — The First Browser That Reads the Room

Introducing Whissle Browser — The First Browser That Reads t

Today we're launching Whissle Browser, a next-generation browser with Lulu — an ambient AI companion that listens, under

By Karan Jakhar

May 01 2026

ASRMandarin ASR Beyond Words: Transcription, Demographics, and Named Entities in a Single Pass

Mandarin ASR Beyond Words: Transcription, Demographics, and

We benchmarked Whissle's 157M-parameter Chinese model against Deepgram Nova-3 and Gemini 2.5 Flash on 5,200 samples acro

By Whissle Research Team

Apr 27 2026

BenchmarkDoes Your ASR's Metadata Actually Make AI Responses Better? We Benchmarked 613 Conversations to Find Out.

Does Your ASR's Metadata Actually Make AI Responses Better?

613 conversations across three public datasets. Six evaluation dimensions. The metadata-aware pipeline won 86% on conver

By Whissle Research Team

Apr 20 2026

ASRBeyond Transcription: How a Meta-Aware ASR Model Delivers Words, Emotion, and Intent in 200ms

Beyond Transcription: How a Meta-Aware ASR Model Delivers Wo

A single CTC model that outputs transcription and metadata action tokens — emotion, intent, speech rate, demographics —

By Whissle Research Team

Apr 16 2026

ASRWe Benchmarked 3 Streaming ASR Providers Across 17 Hours of Audio. Here's What We Found.

We Benchmarked 3 Streaming ASR Providers Across 17 Hours of

4,915 samples. Four datasets. Clean speech, Indian-accented tech interviews, noisy soccer broadcasts, and Indian Supreme

By Whissle Research Team

Apr 10 2026

Using Visuals for Better Sound Awareness

Using Visuals for Better Sound Awareness

Humans naturally use visual cues to understand speech in noisy places. This article explores how we're teaching AI to do

By Whissle Research Team

Apr 18 2025

Meta-aware Voice Action Model

Meta-aware Voice Action Model

To create AI companions that feel genuinely interactive, speech recognition must go beyond raw transcription — it must u

By Whissle Research Team

Mar 16 2025

Blog — Whissle · Whissle