Vocalinux 0.14 beta adds keyboard shortcuts and Wayland fixes

Vocalinux, an open source speech-to-text tool for Linux, released its 0.14 beta version on July 16. The update improves usability with customizable shortcuts and better support for Wayland sessions. It also expands options for remote transcription models.

The new release allows users to configure keyboard shortcuts through the Settings menu. Combinations now support Ctrl, Alt, Shift, and Super keys paired with letters or numbers.

On GNOME Wayland, text injection works again when using a bare XKB engine. The whisper.cpp engine also stops defaulting to all CPU cores on hybrid Intel and AMD laptops.

Remote API users can now access FunASR and SenseVoice models via OpenAI-compatible endpoints. The app remains fully local under the GPL-3.0 license, with voice data staying on the user's machine.

Vocalinux supports multiple engines including whisper.cpp as default, plus options for PyTorch, NVIDIA, VOSK, and remote servers. It runs from the system tray to dictate into any text field.

相关文章

Illustration of a woman using ChatGPT voice features on a tablet in a living room setting.
AI 生成的图像

OpenAI rolls out new GPT-Live voice models for ChatGPT

由 AI 报道 AI 生成的图像

OpenAI has begun releasing two updated voice models for ChatGPT that allow simultaneous listening and speaking. The changes aim to create more natural conversations.

Canonical has introduced Myna, an early-stage AI tool for voice dictation on Ubuntu that runs entirely on local hardware. The feature is planned for Ubuntu 26.10, scheduled for release in October. It uses a push-to-talk system and focuses on privacy by keeping all processing on the device.

由 AI 报道

Anthropic has upgraded its Claude voice mode to run on more powerful models and support connections to third-party apps.

此网站使用 cookie

我们使用 cookie 进行分析以改进我们的网站。阅读我们的 隐私政策 以获取更多信息。
拒绝