Yan

Speak. Words appear.

Yan brings real-time voice input to the app you're already using. Just speak and watch your words appear.

On-Device Real-Time Text One-Time Purchase
Work Notes
Listening
00:08
Current Text Field

Start with a Shortcut

Start voice input with a shortcut without leaving the app you are using, so your train of thought stays uninterrupted.

See Text as You Speak

Watch the text take shape as you speak, keeping your words in step with your thoughts.

Works Across Your Desktop Apps

Use voice input naturally across writing, editing, and communication apps.

Start Voice Input in Three Steps

Set your shortcut once, then start voice input whenever you need it.

Yan shortcut settings
01 · Set a Shortcut

Choose a convenient shortcut that does not conflict with your other apps.

Yan voice input HUD
02 · Start Voice Input

Press the shortcut where you want to write to begin voice input.

Text entered with Yan
03 · Speak Naturally

Focus on what you want to say. Yan turns your voice into words where you are writing.

Real‑Time Voice Input Let Your Words Keep Up with Your Thoughts

Traditional Transcription Text appears after the recording ends
Yan Real‑Time Voice Input See your words as you speak

VOICE INPUT, RE-ENGINEERED

Qwen understands expression Yanbrings it on-device

Next-generation speech models such as Qwen3-ASR understand expression; Yan's own inference engine makes them responsive and reliable on your computer

Yan supports multiple local speech models, so you can choose based on language, quality, and device performance

01 / THE MODEL

Beyond hearing words,
understand what you mean

For Chinese dialects, mixed Chinese and English, varied accents, and specialized terms, next-generation speech models offer stronger language understanding; Yan lets you choose from multiple local models for different situations

02 / THE ENGINE

A powerful model
should truly run on-device

Real-time streaming, hardware backend selection, and resource scheduling determine whether a large model can become an everyday input tool; Yan's inference engine turns model capability into a responsive desktop voice-input experience

CURRENT ASR LIMITS

Hearing the words
is not the same as understanding expression

QWEN + YAN ENGINE · MODEL + ON-DEVICE ENGINE

The model understands
the engine makes it work in real time

A MODEL ALONE

Common Words Pass, Proper Nouns Fail

General-purpose ASR often miswrites people, brands, and domain terms

MODEL + YAN

Language Understanding + Terminology Correction

Qwen3-ASR jointly decodes audio and language information; dictionary rules correct specialized terms

A MODEL ALONE

Low Latency and Stable Text Pull Apart

Partial results are revised as more speech arrives

MODEL + YAN

Draft First, Refine Continuously

An early decode lowers time to first text; the refiner corrects and merges stable text

A MODEL ALONE

Mixed Languages Often Mean Manual Switching

Language-specific pipelines can switch incorrectly when one recording spans multiple languages or songs

MODEL + YAN

One Timeline, Multiple Languages

A full-film test continuously recognized English, Spanish, German, and English-language songs

A MODEL ALONE

On-Device Inference Is Often Tied to CUDA

Traditional runtimes prioritize NVIDIA, leaving AMD and Intel GPUs underused

MODEL + YAN

Beyond CUDA, Discrete and Integrated GPUs Accelerate

Metal runs 1.7B F16 smoothly on M3, while optimized Vulkan runs 1.7B Q8 smoothly on Intel Xe integrated graphics

VOICE INTELLIGENCE, ON-DEVICE

It is more than switching models
It brings next-generation speech intelligence to your desktop

From models and inference to the desktop input experience, Yan connects every step from voice to text

Transcription quality and speed depend on the selected model, language, accent, recording environment, and device performance

Your Voice Stays on Your Device

Yan's on-device speech engine handles recognition locally, so your voice and transcripts do not need to pass through the cloud.

Recognition on Your Device

Speech recognition runs on your device, so your audio does not need to be uploaded to the cloud.

Content Is Not Used for Training

Yan does not use your voice or transcripts to train models.

History Stays on Your Device

Your transcript history stays on your device. Yan does not store a copy on its servers.

Local by design, clear by default

From speech models to transcription history, the essentials stay visible and under your control.

Voice Input for Everyday Work

From writing and communication to everyday tasks, Yan turns natural speech into clear text

Writing & Drafting

Draft directly in documents and editors, so your words keep up with your thoughts.

Messages & Email

Draft messages and emails quickly without repetitive typing.

Documents & Collaboration

Add content to documents, spreadsheets, and collaboration tools more efficiently.

Support & Operations

Draft replies, explanations, and internal updates while staying focused on the conversation.

Notes & Ideas

Capture thoughts as they come and keep them as text you can refine later.

Works in Common Apps with Standard Text Input

VS Code
Zoom
Slack
Teams
Notion
Google Meet
Word
Premiere Pro
Google Drive
Figma
Excel
OBS Studio
PowerPoint

The icons show common use cases. Compatibility depends on each app and text field supporting system text input.

ADDITIONAL CAPABILITY

Transcribe Existing Audio and Video

Alongside real-time voice input, Yan can batch-transcribe common audio and video files on your device, ready to review, copy, or export.

Buy Once. Keep Using Yan.

Try every feature free for 7 days. When you're ready, buy once and keep using Yan with no subscription.

One-Time Purchase
Perpetual License
14-Day Refund

Try Every Feature for 7 Days

No payment required. The trial includes everything in the full license.

Download and Start Trial