LipConfirm logo LipConfirm
Book a Full Demo How it works Integration Patent & IP FAQ Contact

A whisper too quiet for dictation can still fix your typing.

LipConfirm pairs a barely audible whisper, or a silent lip movement, with the letters you type to improve autocorrect.

Alone, a whisper can be ambiguous. Combined with your typing, it can help distinguish “live” from “love” or recover a garbled word. Typing stays primary. Fix a word with a brief cue instead of backspacing, while keeping most of your message unspoken.

Book a Full Demo

All 30 claims allowed · U.S. Patent No. 12,737,540 scheduled to issue September 15, 2026.

LipConfirm hero illustration

Try the interactive demo

Typing stays primary in this simulation. Whisper and lip cues guide correction rather than provide a standalone dictation or lipreading result. A single articulation can supply both cues in a combined implementation; combined capture is not simulated here.

Why another cue helps

Even LLM-informed autocorrect can struggle with garbled input or a typo that produces another valid word. A quiet cue adds evidence of what you intended, beyond what spelling and context alone can establish.

A faint whisper, backed by typing

Typing can make a faint whisper more useful than that whisper alone. Partial clues, such as syllable rhythm or a distinguishing vowel, can help select among words supported by your typing. LipConfirm uses that evidence without requiring a reliable standalone transcript. Whisper assistance needs no camera.

Familiar words can be mistyped, too

A slip can turn love into live or this into thus. A correctly spelled word can still be the wrong word. Both “I love my life here” and “I live my life here” are grammatical, so context may not settle your intent. The demo includes love/live. See 20 everyday examples.

A cue can come before or after the mistake.

Quietly articulate a word as you type it, or add a cue after you notice an error. LipConfirm can reconsider that word using the original typing. Optional silent lip assistance can use a brief, requested capture.

Why not just whisper into dictation?

Keep the phone where you type

You may want to keep most of your message unspoken, stay quiet around other people, or simply keep using the keyboard. LipConfirm lets a brief cue assist the text you are already typing.

A very faint whisper may fail to register or may produce the wrong words at normal typing distance. Speaking louder or moving the phone toward your mouth can help, but it changes the flow of typing.

In my preliminary desktop work, combining faint whisper cues with typed input helped recover the intended word, even when the typing was garbled.

Try a faint-whisper check on your phone

  1. Open your usual keyboard’s dictation mode and hold the phone at your normal typing distance. Start by whispering “I love my life here”.
  2. Repeat the same sentence in progressively fainter whispers without moving the phone. Look for a range where you can still faintly hear yourself, but dictation starts substituting words, missing them, or returning no text.
  3. Try that faint whisper again with the phone closer to your mouth. Notice whether recognition improves and how moving the phone changes your typing position.

A faint cue can still carry useful information. LipConfirm is designed to use that uncertain signal together with your typing. The point at which dictation becomes unreliable varies by device and setting.

How faint cues reach correction

A faint whisper often retains useful clues to syllable count and rhythm, critical vowel or consonant sounds (phonemes), and sound timing. LipConfirm evaluates the available acoustic evidence together with the original typing, including garbled text. The whisper need not identify the word independently: a partial cue can help separate otherwise plausible candidates.

The capture path must preserve that weak signal. Depending on the system, speech-detection gates or audio filtering can exclude faint utterances before they reach recognition. A user-enabled typing session or requested capture can provide a path for evaluating the retained acoustic evidence with the text. A blank dictation result alone does not identify which stage failed.

Phonetic background: research on whispered speech documents vowel resonances and sound durations despite the absence of normal vocal-fold vibration. This helps explain why incomplete acoustic evidence can remain informative; it does not establish a minimum usable volume for LipConfirm.

20 examples where an extra cue can help

A small keyboard slip can produce a perfectly valid word with the wrong meaning. A brief whisper or silent lip cue can add evidence of the word you intended. See what the cue contributes in selected examples below.

U I O sit side by side on a QWERTY keyboard. That makes pairs such as love/live, this/thus, if/of, and shirt/short easy to mistype. These tiny slips can be frustrating because a real word may pass a spelling check, and context does not always rescue the mistake.

10 isolated word examples

#You meantTyped by mistake
01lovelive
02thisthus
03ifof
04shirtshort
05pickpuck
06sicksock
07hithot
08tiptop
09slipslop
10fillfull

10 short sentence examples

#You meantTyped by mistake
11I love my life here.I live my life here.
12I left it in the car.I left it on the car.
13I like the shirt dress.I like the short dress.
14We have three cats.We have three cars.
15Please bring the coat.Please bring the boat.
16The team is ready.The tea is ready.
17Please renew my subscription.Please review my subscription.
18I’ll go with the latte.I’ll go with the latter.
19It’s a casual relationship.It’s a causal relationship.
20I left it in the bag.I left it in the bar.

In the sentence examples, both versions are grammatical but mean different things. A language model can favor one, yet still miss your intent.

What the cue contributes

A cue need not reveal the whole word to help resolve it. Even faint whispers often carry clues to syllable count, rhythm, critical vowel or consonant sounds (phonemes), and timing. Optional lip capture can add visible closure, rounding, and movement. LipConfirm combines the available cues with the typing to help choose between plausible words; audio and lip evidence can also reinforce one another.

Whisper · A distinguishing vowel

love / live, this / thus

In “I love my life here,” a faint “uh” vowel cue can favor love over live, pronounced “liv.” The same vowel contrast helps separate thus from this. Within each pair, the consonants and syllable count match. The typing narrows the choice; the vowel adds evidence context may lack.

Whisper · The sound behind a one-key slip

shirt / short

“I like the shirt dress” and “I like the short dress” both make sense. In common American pronunciation, the “er” sound in shirt differs from the “or” sound in short. A captured trace of that vowel can help LipConfirm reconsider the neighboring I/O key slip without transcribing the whole sentence.

Silent lips · An ending the text missed

team / tea

If you type “The tea is ready” but articulate team, the lips close for the final “m.” A captured closure timed to that word ending can support team over tea. LipConfirm can weigh that visible cue alongside the shared typed letters; a resting mouth closure alone would not establish the word.

Whisper or lips · A different beginning

coat / boat

A captured initial “k” sound can favor coat over boat. In the other direction, boat begins with a “b” made by closing and releasing both lips, while “k” is formed farther back in the mouth. A captured lip closure and release at the word’s start can support boat among the typed candidates.

Whisper · Syllable pattern and timing

casual / causal

When casual is articulated with three syllables, its rhythm differs from the usual two syllables of causal. The vowel and middle consonant sounds also differ. LipConfirm can weigh those clues against the nearly identical typed letters. Pronunciation varies, so syllable count is supporting evidence rather than a fixed rule.

Whisper · Why the available channel matters

if / of

These words differ by neighboring I/O keys. A captured vowel cue can help favor if over of, whose vowel varies with pronunciation and emphasis. Lips alone offer less help: “f” and “v” use similar lower-lip-to-upper-teeth contact. LipConfirm can weigh the audio with the typing instead of relying on that shared mouth shape.

These walkthroughs illustrate how cues can support correction. Pronunciation and capture quality determine which evidence is available. When the combined evidence is insufficient, LipConfirm can retain the typed word or offer alternatives.

Phonetic background for these examples

Whispered vowels can retain acoustic resonances even though a true whisper lacks normal vocal-fold vibration. Lip cues provide a different kind of evidence: “b” and “m” involve both lips, while “f” and “v” share lower-lip-to-upper-teeth contact. A visible gesture therefore supports some distinctions without uniquely identifying every sound.

Sources: Houle and Levi, acoustic differences between voiced and whispered speech; Essentials of Linguistics, consonant articulation; Cambridge Dictionary, pronunciation of “casual”. The application to these typed pairs is illustrative.

How it works

The videos below show lip-assisted correction in a desktop prototype.

Without LipConfirm
With LipConfirm

Capture useful cues

In a user-enabled typing session or requested capture, the microphone can collect a brief cue without requiring a successful dictation transcript. The capture path must preserve usable faint speech. Optional camera capture can be limited to requested lip assistance, with local processing and disposal of raw data supported by the architecture.

Use evidence inside correction

The original typing and the acoustic or lip evidence are evaluated together. Weak cues can help distinguish between words supported by the typing. The system can rescore candidates, help recover missing candidates, or combine inputs in a learned multimodal model without first requiring a reliable standalone transcript.

Suggest, confirm, or replace

The combined evidence can update a suggestion, confirm a word, or correct a valid word that was typed unintentionally. Confidence controls matter for familiar words as well as garbled input: when the evidence is insufficient, the system can retain the text or offer alternatives.

What the demos show

  • Garbled input and everyday mix-ups. The interactive simulation includes both malformed spellings and valid-word errors such as love/live. The videos show selected lip-assisted correction scenarios.
  • Typing stays primary. Lip cues in the videos guide a typed correction. Whisper assistance applies the same principle using quiet audio.
  • No need to repeat everything. An optional cue can accompany any word you want help with, including a familiar word typed incorrectly. You do not need to whisper or mouth the rest of the message.
  • Note: The camera panel is a behind-the-scenes visual aid. An integrated keyboard need not display the feed, and requested visual correction need not keep the camera running throughout typing.

The videos below show lip-assisted correction in a desktop prototype.

Without LipConfirm
With LipConfirm

What the demos show

  • Garbled input and everyday mix-ups. The interactive simulation includes both malformed spellings and valid-word errors such as love/live. The videos show selected lip-assisted correction scenarios.
  • Typing stays primary. Lip cues in the videos guide a typed correction. Whisper assistance applies the same principle using quiet audio.
  • No need to repeat everything. An optional cue can accompany any word you want help with, including a familiar word typed incorrectly. You do not need to whisper or mouth the rest of the message.
  • Note: The camera panel is a behind-the-scenes visual aid. An integrated keyboard need not display the feed, and requested visual correction need not keep the camera running throughout typing.

Built around ordinary typing

Keep typing normally and add an occasional cue when you want help with a word, whether it is garbled, ambiguous, or an everyday word typed incorrectly. LipConfirm is designed to bring the combined evidence into existing keyboard, autocorrect, spell-checking, and language-model infrastructure.

Technical demonstrations and IP discussions available. Desktop prototypes cover both whisper-assisted and lip-assisted correction.

Patent & IP

All 30 claims allowed. U.S. Patent No. 12,737,540 is scheduled to issue September 15, 2026. The allowed claim set addresses lip-assisted correction, shared multimodal operation, and a camera-free whisper-assisted pathway that generates textual proxies for an existing spell-correction engine.

Two U.S. continuation applications are pending, seeking additional protection for lip- and whisper-assisted autocorrect. One has been granted Track One prioritized examination.

PCT/US2026/035963 is pending. The USPTO, acting as the International Searching Authority, issued a written opinion with positive novelty, inventive-step, and industrial-applicability findings for all 30 claims. The international search report contains no category X or Y citations.

FIG. 1 — Typed input fused with viseme (lip) and optional whisper signals to select the intended word.
FIG. 1 — Typed input fused with viseme (lip) and optional whisper signals to select the intended word.

FAQ

Why not just whisper into dictation as it improves?

You may prefer typing to keep most of a message unspoken or avoid disturbing others. The whisper you are comfortable making at normal typing distance may also be too faint for reliable dictation. LipConfirm adds the evidence in your actual typed input, so the same uncertain whisper can help resolve a word without having to identify it independently. This can be useful even as dictation improves.

Can typing plus a faint whisper be more accurate than that whisper alone?

In my preliminary desktop work, combining faint whisper cues with typed input helped identify the intended word, including when the typing was garbled. These are early findings from selected demonstrations. I have not yet established a representative accuracy percentage across users, devices, noise levels, and normal phone typing distances. Broader evaluation needs to count both recovered words and unwanted corrections.

Does this apply only to hard-to-spell words?

No. Familiar words can be mistyped into other valid words, including love/live and this/thus. “I love my life here” and “I live my life here” are both grammatical, so even a strong language model may not settle the user’s intent from context alone. A whisper or lip cue can supply additional evidence; confidence checks determine whether to change the word.

What if my phone does not pick up a faint whisper?

Try the progressively fainter whisper check at your usual typing distance. You may find a range where the whisper is still faintly audible to you but dictation becomes unreliable. An implementation must preserve those cues in its capture path; a blank dictation result alone does not reveal whether the failure was in detection, processing, or recognition.

Does whisper assistance keep a message private?

You can keep most of the message unspoken and add only an occasional cue. A whisper may still be audible nearby; optional lip assistance provides a silent path. For data privacy, the architecture supports local processing and discarding raw audio or mouth-region frames. The host platform determines the implemented data handling.

What is whisper-level audio?

Here, whisper-level audio means quiet acoustic evidence, including faint, low-amplitude or breathy speech. The aim is to use a comfortable brief cue while keeping the phone where you normally type. LipConfirm evaluates that evidence with the typed input; a complete, correct standalone transcript is not required. Usable volume depends on capture settings, microphone distance, background noise, and the model.

Does LipConfirm require a camera?

No. Whisper-assisted correction combines microphone input with typing and does not need a camera. Optional lip assistance can provide a silent cue or reinforce a quiet utterance.

Would the camera run throughout typing?

Continuous camera use is not required. In the on-demand approach, the camera stays off until lip assistance is requested, captures a short clip of the word, then stops. Longer visual sessions remain an optional, user-enabled configuration.

Can the microphone stay available while I type?

Yes, in an opted-in typing session. The microphone can remain available so an occasional quiet cue accompanies a word as you type it, including a familiar word that is easy to mistype. Alternatively, it can activate only when assistance is requested. There is no requirement to whisper every word.

Can I correct a word after it appears?

The architecture supports associating a new cue with the word just typed and reconsidering the correction using the retained typing evidence. Optional lip assistance can use a brief capture requested for that word.

Can whisper and lip cues work together?

Yes. One quiet articulation can supply audio and lip cues. The architecture combines both with typed evidence and adjusts their influence according to reliability. An unavailable or unreliable channel can contribute less or be omitted.

What is demonstrated today?

I can provide desktop demonstrations of whisper-assisted and lip-assisted correction. The interactive website demo is an illustrative simulation with preset examples and scores; it does not access your microphone or camera or measure live recognition accuracy.

Does this require the internet or cloud processing?

LipConfirm is designed to support on-device correction without sending raw mouth-region frames or audio to the cloud. Raw sensor data can be discarded after feature extraction or scoring. Deployment choices depend on the host platform and model.

Will this drain battery or slow down typing?

The design supports brief camera use and session-based or requested microphone capture. Actual battery use and latency depend on sensor settings, recognition models, and the device.

How does the extra evidence enter autocorrect?

It can rescore or validate existing candidates, help generate or recover candidates, or enter a learned model together with text. A proxy-string pathway can turn speech-related clues into spelling-like inputs for an existing correction engine. The interactive demo illustrates late and early fusion as two distinct strategies.

Do I need to perfectly mouth the words?

Text-assisted correction can use partial cues, such as syllable patterns and mouth shapes, rather than independently recognizing every word from lips alone. Useful visual evidence still depends on articulation and capture quality. Optional personalization can adapt scoring to a user.

Is LipConfirm only for mobile phones?

The architecture supports keyboard-based correction on phones, tablets, desktop systems, wearables, and extended-reality devices. Each platform can use the available, user-permitted combination of typed, acoustic, and visual evidence.

Will I still see suggestions?

A high-confidence result can be inserted or confirmed automatically. When the evidence is uncertain, the system can offer alternatives, retain the current text, or wait for another cue. The user does not need to see the internal candidate scores or camera feed.

Is this patented?

All 30 claims of the U.S. parent application have been allowed, and U.S. Patent No. 12,737,540 is scheduled to issue September 15, 2026. Two U.S. continuation applications and a PCT international application are pending. One continuation has been granted Track One prioritized examination. See Patent & IP for details.

Contact

Email: hello@lipconfirm.com

LipConfirm was conceived and developed by Andre Persidsky, a multidisciplinary engineer and inventor with experience building, patenting, and licensing technology. Andre has previously licensed technology to commercial partners including Spotify, and has led hardware/software health-technology projects in collaboration with external research institutions.