Audrix captures live audio and on-screen visuals, transcribes with noise suppression and crosstalk detection, and runs OCR and Vision locally. Transcription runs 100% offline. Live AI analysis can run locally or via your preferred cloud provider, with 15 providers including OpenAI, Anthropic, StepFun, Groq, OpenRouter, and local Ollama and LM Studio.
Audrix Engine: real-time transcription from microphone or system audio, with local OCR and Vision for instant screen analysis.
Three steps to real-time local intelligence.
Use WhisperLive (11 models), Lemonade NPU acceleration, or any external transcription endpoint. Switch at runtime without reinstalling.
Grab microphone, system audio, or both simultaneously. Audrix handles noise suppression and crosstalk detection automatically.
See text appear in real time. Select any screen region for OCR and Vision. Everything runs locally on your machine.
Reviewed and recommended by publications across the software space.
"Audrix is a tool that lets you capture sound from your microphone, PC speakers or both at the same time without dealing with complicated configuration. The transcription engine is designed to reduce duplicate words and improve accuracy, even if someone is speaking continuously or multiple people speak at the same time."
Read full review โSponsored
Audrix transforms live audio and on-screen visuals into immediate, actionable insights. Transcription runs 100% offline. Live AI analysis can run locally or via your preferred cloud provider.
Audrix is a real-time local intelligence engine for Windows. It captures microphone or system output for crisp speech-to-text with advanced noise suppression and crosstalk detection, while simultaneously running live AI analysis on the audio stream as it happens. Pair this with local OCR and Vision capabilities to instantly read and analyze any section of your screen.
Built for low-latency transcription with configurable noise suppression, automatic crosstalk detection, and 17+ selectable models. Transcription runs 100% offline. Your audio never leaves your machine.
Runs from a USB stick. No install, no account, no cloud sync.
Sponsored
Common frustrations and how Audrix solves them.
Most STT tools only capture the microphone. You cannot transcribe YouTube videos, Zoom calls, or any audio playing through your speakers.
Cloud STT services send your audio to remote servers. You pay per minute, lose privacy, and depend on internet connectivity.
In meetings or calls with multiple speakers, transcription becomes messy. Background noise and overlapping speech corrupt results.
You can search audio but not what you see. Important details in videos, demos, or documents remain inaccessible to text-based tools.
Most STT tools lock you into a single model or service. Changing transcription providers requires reinstalling or reconfiguring everything.
Everything you need for real-time transcription, screen analysis, and flexible backend support.
Audrix transcribes audio in real-time as it is captured. See text appear the moment words are spoken, perfect for live captions, meeting notes, or voice commands. Smart merging prevents duplicate text, and partial results give you faster feedback before finalization.
Capture your microphone, any app's audio output, or both simultaneously with automatic speaker separation. Transcribe your voice, YouTube videos, Zoom calls, or any system audio.
Configurable noise gate and auto-gain keep background hum and silence out of your transcripts. Dual-stream mode distinguishes the local user from remote speakers on calls.
17+ transcription models via WhisperLive (11 Whisper variants), Lemonade NPU acceleration, or any external WebSocket or HTTP transcription endpoint. Switch at runtime without reinstalling.
Extract text from any screen region in real time with local optical character recognition. AI-powered visual understanding of on-screen content, diagrams, and UI elements.
Transcription runs entirely on your machine. Your audio never leaves your device. Live AI analysis can use local providers or your preferred cloud service - your choice.
One-click capability test for transcription backends. Know if a model handles accents, noisy audio, or long-form speech before you rely on it. The only reliable way to use a model at full capacity, because provider standards differ.
Swap the AI's behavior for live transcript analysis, vision, and OCR by switching prompt profiles. Interviewer mode surfaces questions and answers. Researcher mode extracts topics and references. Sales mode highlights objections and next steps. Build your own roles or import presets. The same model adapts to the job.
Capture microphone and system audio simultaneously. Built-in crosstalk suppression distinguishes the local user from remote participants on calls.
Adjust sample rates, buffer sizes, commit intervals, and silence thresholds. Fine-tune every parameter for your hardware and use case.
From transcription to screen analysis, Audrix handles real-time audio and visual intelligence.
Sponsored
The local advantage for speech-to-text and screen intelligence.
| What you get | Audrix | Cloud STT | Generic Offline STT |
|---|---|---|---|
| Runs fully offline & private | Transcription is 100% local. Analysis runs on your machine or via your chosen cloud provider. | Cloud only. Your data on their servers. | Local processing, limited features |
| System audio capture | Loopback + mic, dual-stream mode | Microphone only | Rare or unsupported |
| Crosstalk suppression | Dual-stream distinguishes speakers | Basic noise removal only | Unsupported |
| Local OCR & Vision | Built-in screen analysis | Cloud-only vision APIs | Unsupported |
| Privacy | Audio stays local. Transcript analysis only sent if you choose a cloud provider. | Audio sent to remote servers | Local but no ecosystem |
| Cost | $0, free community build | $0.006 to $0.03 per minute | Free but limited |
| NPU acceleration | AMD Ryzen AI via Lemonade | Server-side GPUs only | CPU-bound |
| Portable | Runs from USB, no install | App install or web only | App install required |
Sponsored
Researchers, coders, journalists, meeting hosts, and anyone who values privacy and offline capability.
Transcribe interviews, lectures, and meetings with perfect accuracy and speaker separation.
Dictate articles, transcribe interviews, and capture audio notes with instant text output.
Run live captions for Zoom, Teams, or Discord calls. Dual-stream mode separates local and remote speakers.
Transcribe coding sessions, capture system audio from tutorials, and use OCR to read documentation or error messages.
100% offline transcription with optional local or cloud AI analysis. No accounts, no cloud sync, no tracking unless you choose a cloud provider.
Sponsored
Supported providers
Transcription Engines
WhisperLive (11 models), Lemonade NPU (dynamic discovery), External WebSocket
Live Analysis Providers
Cloud
Local / On-premise
Fast local transcription with AMD Ryzen NPU acceleration, plus external endpoint support.
Low-latency audio capture with configurable pipeline for transcription and analysis.
Audio is captured and processed with minimal delay. Reliable streaming delivers stable audio to the transcription engine.
To correctly capture system audio without losing your loudspeaker output, Windows must be configured to pass the virtual stream back to your hardware. This requires VB-Audio Virtual Cable, a free virtual audio driver.
Fine-tune the audio pipeline for your hardware and environment.
Read and analyze any section of your screen with local OCR and Vision.
Beyond audio, Audrix includes local OCR and Vision capabilities. Select a region, capture it, and get text extraction or visual analysis immediately. No cloud uploads, no API keys, no waiting.
Download, extract, and run. No installation required.
Choose from 17+ Whisper models via WhisperLive, Lemonade NPU, or external endpoints. Add live AI analysis with 15 providers, local or cloud.
Grab the latest portable executable from GitHub. Extract it anywhere, even a USB drive.
Use Whisper, Lemonade NPU acceleration, or any external transcription endpoint. Switch at runtime.
Select your audio device, configure noise thresholds, and watch real-time transcripts appear.
Join the community, leave a review, or support the project.
Join the Discord server to ask questions, share workflows, and get help from other Audrix users.
Join DiscordHave a question, suggestion, or bug report? Open a discussion or issue on GitHub.
GitHub FeedbackLeave feedback on Softpedia to let other users know what you think.
Softpedia FeedbackLeave feedback on AlternativeTo to help others discover Audrix.
AlternativeTo FeedbackAudrix is fully free today. A planned premium option is coming for users who want extended use and support.
Free community build with a 30-minute session cap. Bring your own models or use built-in backends. No account, no trial clock.
Optional paid tiers for extended session duration, priority support, and commercial licensing. All current features remain available in the free tier.
Not yet available, join Discord for updates.
Join Discord for updatesSponsored
Other tools from Tetramatrix for productivity, gaming, and AI-powered workflows.
Your personal AI knowledge workspace. A second brain that remembers your projects, learns your style, and acts on your files. Runs fully offline on Windows.
AI Intelligent Desktop Icon Autopilot. Automatically organizes your cluttered Windows desktop using AI. Group icons intelligently, arrange them neatly.
AI Spatial Tab Manager & Research Workspace. Maps browser tabs onto a 2D canvas, AI groups them by content, chat with any page or the live internet.
Powerful tool for managing AMD Ryzen processor power settings on Windows. Adjust CPU performance, power limits, and thermal configurations.
Real-time speech-to-text engine. Capture microphone or system audio, transcribe instantly with noise suppression and crosstalk detection. Runs entirely offline on Windows.
Quick answers before you download.
Audrix is free to download and use. A free community build is available with a 30-minute session cap. Studio and Enterprise editions are available for extended use.
Transcription audio stays 100% local. Transcript text is only sent to the cloud if you choose a cloud LLM provider for live analysis. You can run entirely offline with local providers like Ollama, LM Studio, or Lemonade.
Yes. Audrix is a portable executable. Extract it anywhere and run it. No installation, no registry changes, no admin rights required.
Audrix supports WhisperLive (11 Whisper models: tiny through large-v3, plus distil-small.en and Parakeet TDT 0.6B), Lemonade NPU acceleration with dynamic model discovery, and any external WebSocket or HTTP transcription endpoint. Backends are hot-swappable at runtime.
Yes. Audrix captures microphone, system audio (loopback), or both simultaneously. Dual-stream mode distinguishes the local user from remote speakers on calls.
Sponsored
Download Audrix and get instant local speech-to-text, system audio capture, noise suppression, and OCR. Free, private, and built for Windows.