Tagged "audio-processing"
7 articles tagged audio-processing, 18 March 2026 to 28 September 2026. Newest first.
-
Qualcomm Snapdragon Sound Elite Gen 2 Brings On-Device AI To Audio Wearables
Qualcomm's latest Snapdragon Sound Elite Gen 2 platform integrates on-device AI capabilities specifically optimized for audio processing in wearables and IoT devices. This hardware advancement enables real-time audio AI inference without cloud connectivity.
-
Running Qwen3-Omni With Audio and Vision in llama.cpp
One mmproj carries both encoders, --image and --audio are the same flag, and speech output does not work at all. The verified commands, real file sizes and open bugs for the only open-weights omni model.
-
Tiny microphone on my balcony to listen for any birds passing by
A practical demonstration of edge AI inference using miniature audio hardware and local ML models for real-time bird species identification without cloud connectivity.
-
Open Source Local Audio Stem Separation Tool Released
A new free, open-source tool for local audio stem separation has been released on GitHub, enabling on-device audio processing without cloud dependencies. This project demonstrates practical local ML inference for audio workloads.
-
Qwen3 Audio and Vision Support Now Available in llama.cpp
Qwen3-Omni and Qwen3-ASR models now run natively in llama.cpp with full audio and vision input support. This enables truly multimodal local inference with Alibaba's frontier-competitive model architecture.
-
Audio Processing Support Lands in llama.cpp with Gemma-4
llama.cpp now supports speech-to-text functionality with Gemma-4 E2A and E4A models, enabling local multimodal inference on consumer hardware. This expansion brings audio capabilities to the most widely-used local LLM inference engine.
-
Browser-Based Transcription Tools
Browser-based transcription solutions leverage local inference to enable audio processing entirely within the user's device, eliminating cloud dependency for speech-to-text tasks. This trend reflects growing adoption of WebAssembly and on-device AI models for privacy-preserving audio applications.