voxtral
Here are 54 public repositories matching this topic...
Turn PDFs and EPUBs into audiobooks; subtitles or videos into dubbed videos (including translation), and more. For free. Pandrator uses local models, including voice-cloning (instant, RVC-enhanced, XTTS fine-tuning) and LLM processing. It aspires to be a user-friendly app with a GUI, an installer and all-in-one packages.
-
Updated
Sep 9, 2026 - Python
C++ ggml runtime hub for multilingual ASR and TTS models: Cohere Transcribe, Parakeet TDT, Voxtral, Canary 1B v2, etc, plus universal forced alignment, and more
-
Updated
Sep 7, 2026 - C++
Super STT let's you speak, and your words are typed straight into whatever app is focused
-
Updated
Sep 9, 2026 - Rust
Voxtral is a state-of-the-art model developed to handle both speech transcription and audio understanding with remarkable accuracy and efficiency. This demo interface lets you run the Voxtral model on powerful GPUs to evaluate its performance and see how it can be used for transcription and deeper analysis.
-
Updated
Jul 26, 2025 - Python
Offline Speech-to-Text (STT) service using Mistral's Voxtral model with Wyoming protocol compatibility for Home Assistant Assist integration.
-
Updated
Jun 17, 2026 - Python
speech to text gui for different (e.g. Whisper, Voxtral) models and backends, including whisper.cpp, crispasar, mlx-whisper, faster-whisper, ctranslate2; applies pyannote for diarization
-
Updated
Aug 16, 2026 - Python
Effortless Push-to-Talk Transcription, Anywhere.
-
Updated
Sep 2, 2026 - Python
A Web UI for easy subtitle using various models including voxtral
-
Updated
Jul 22, 2025 - Python
Professional local-first AI production pipeline for long-form narration. Clone voices and generate studio-grade audiobooks (M4B/MP3) using Coqui XTTS-v2 and support for Voxtral (cloud)
-
Updated
Sep 3, 2026 - Python
Voxtral Codec : Combining Semantic VQ and Acoustic FSQ for Ultra-Low Bitrate Speech Generation (Voxtral TTS Backbone)
-
Updated
Mar 27, 2026 - Python
Experimentation with Voxtral-Mini-4B-Realtime-2602 and DeepL API for live translation
-
Updated
Mar 23, 2026 - Astro
Talk. Ink. Push-to-talk dictation for macOS, 100% on-device. Pick your model: Qwen3-ASR, NVIDIA Nemotron or Voxtral, all via Apple MLX.
-
Updated
Jun 21, 2026 - Swift
github mirror for radioshaq - ham radio full time quarterback and part-time lobster
-
Updated
Mar 15, 2026 - Python
Open-source, local-first, system-wide voice dictation for Windows
-
Updated
Sep 6, 2026 - Rust
Local-first Plaud Note Pro audio vault. Direct Bluetooth imports, Mistral/Voxtral transcription, Markdown documents, AI chat, REST API and MCP. Tauri macOS app; experimental Android.
-
Updated
Sep 8, 2026 - TypeScript
Real-time face-to-face translation app (React Native)
-
Updated
Aug 16, 2026 - TypeScript
Real-time phone scam detection powered by Mistral's Voxtral Mini - analyzes live audio and transcripts to identify fraud patterns
-
Updated
Mar 2, 2026 - Python
Enterprise-grade speech-to-text toolkit with pluggable backends (Whisper, Voxtral). Features speaker diarization, 80%+ test coverage, CI/CD quality gates, and fully offline operation.
-
Updated
Feb 18, 2026 - Python
Local implementation for voxtral
-
Updated
Dec 20, 2025 - C++
Add this topic to your repo
To associate your repository with the voxtral topic, visit your repo's landing page and select "manage topics."