Tagged "apple"
-
Ollama v0.33.1 Adds Qwen3.8-Flash-Next Support via MLX Backend
-
Ollama v0.33.1 Adds Qwen3.8 Flash Next Support and Claude Desktop Integration
-
vLLM-iOS Achieves 88% Faster Multi-Agent Inference on Mobile Devices
-
Llama.cpp Build 10620: Continued Optimization for Local Inference
-
Qwen 3.6 Now Easier to Run Locally on Mac with JetBrains Integration
-
llama.cpp Build b10581 Adds DSpark Support for Faster Local Inference
-
Llama.cpp Release b10485: GGML Sync with Platform-Specific Optimizations
-
DeepSeek V4 Flash Shrunk to 57GB for Local macOS Inference with Compiler Generation
-
The Qwen MLX Challenge
-
Llama-macOS – Agentic and MCP Native macOS Front End for Llama.cpp
-
Ollama Adds Qwen 3.8 27B with Optimised Apple Silicon Support
-
Ollama Adds Qwen 3.8 27B with Apple Silicon Optimizations
-
DeepX's DX-M1 On-Device AI Chip Achieves $13M in Orders
-
Local Model Performance Benchmarks on MacBook Pro M5 Max: Real-World Inference Metrics
-
AMD Launches Gorgon Halo and ROCm.AI for Local AI Inference with Workstation Hardware
-
Apple's On-Device AI Strategy Focuses on Privacy and Latency, Not ChatGPT Competition
-
Ollama Releases NVIDIA Nemotron 3.5 Lightning for Local Agent Deployment
-
Meta's Muse Glimmer Now Available Across All Platforms in Ollama
-
Meta's Muse Glimmer Now Available Across All Platforms in Ollama
-
Meta's Muse Glimmer Now Available Across All Platforms via Ollama
-
MacPaw and Liquid AI: Complete On-Device AI Stack for macOS
-
Meta Releases Muse Glimmer: 30B Open-Source LLM for Local Deployment
-
Muse Glimmer Now Available on Ollama – Meta's Open Multimodal Agent Model
-
ShoutFlow Launches Pay-Once, On-Device AI Dictation App for the Mac
-
Llama.cpp Fixes Metal NORM Operations for Apple Silicon
-
MacPaw Partners With Liquid AI to Deploy On-Device AI Across Mac Ecosystem
-
Ollama v0.32.6: Faster Apple GPU Inference with Speculative Decoding
-
PrismML's Bonsai 27B Brings On-Device AI to Apple iPhone 17 Pro
-
Apple's Hardware Is Ready for On-Device AI and PrismML Just Delivered a Real Breakthrough
-
Tim Cook Called Apple's On-Device AI a 'Competitive Weapon' in Final Earnings Call as CEO
-
How Much Does a Local LLM Actually Cost to Run? Energy Costs Measured on Apple Silicon
-
CPU vs GPU vs NPU: Which Semiconductor Does What?
-
Odysseus - PewDiePie's Self-Hosted AI Finally Runs Fast on Mac
-
Arm China Unveils "Tianxuan" CPU and Xingchen 300 Platform, Targeting Ubiquitous AIoT with On-Device AI Portfolio
-
Qualcomm's Next Budget Chip Could Bring On-Device AI To The Phones Most People Actually Buy
-
On-Device AI vs Cloud AI: Which One Should Power Your Next Phone?
-
Sunday Reboot: Shrinking Models and an On-Device AI Future
-
Apple in Early Talks With PrismML on AI Compression Tech
-
Apple in Talks with PrismML to Shrink AI Models 15x for iPhone Deployment
-
Apple Boosts On-Device AI, Partners With PrismML to Enable Running Large Models Locally on iPhone
-
Show HN: AITerm – a macOS Terminal with an AI Command Loop and a Safety Gate
-
Apple's M6, M7, and M8 Chip Roadmap Shifts Focus Toward AI
-
Apple's Failed Self-Driving Car Program Left a Legacy of Powerful AI Chips
-
Running Local AI on Mac With Home Assistant Integration
-
Apple Explores Running Larger AI Models on iPhone with On-Device Compression
-
Edge AI Smartwatch Shipments Jump 70% as Apple Leads Health-Focused Boom
-
Apple's MacBook Lineup Overhaul Features M7 Chip for Enhanced Local AI
-
Ollama's New MLX Engine Delivers Significant Performance Gains on Mac
-
Using a local iPhone MCP server to plan Apple Watch workouts with Codex
-
Apple Updates Creator Studio with AI Video Editing, Image Generation, and Logic Pro Enhancements
-
Asahi Linux 7.1 Progress Report
-
You Can Now Run Max AI Models on Apple Silicon
-
Liquid AI Ships LFM2.5-230M with Broad Framework Support for On-Device Inference
-
Apple's M7 Chip Delivers 56% Memory Bandwidth Increase for On-Device AI
-
The Mac Mini is the Best On-Device AI Computer You Can Buy: Here's Why
-
Mac Mini Emerges as Top Choice for Local On-Device AI Deployment
-
2026 On-Device AI Market Intensifies: Apple, Google, and Samsung Compete for Local AI Dominance
-
Mac Mini Positioned as Premier On-Device AI Computer for Local LLM Inference
-
Apple unveils Core AI for on-device generative models
-
South Korea Launches K-On-Device AI Chip Project With 511.1B Won Funding
-
Samsung's Exynos 2600 Doubles On-Device AI Performance in MLPerf Benchmarks
-
Most People Use Ollama or llama.cpp for Local LLMs, but These Are the Tools I Switch to When It Gets Serious
-
Show HN: 11 Model Families Ported to Apple's CoreAI On-Device Framework
-
Qualcomm Launches Dragonwing MBM Silicon with Advanced On-Device AI Capabilities
-
Apple Unveils AFM 3 Core Advanced with 20 Billion Parameters for On-Device AI
-
Apple Rebuilt Its On-Device AI Stack at WWDC 2026
-
Ask HN: Thoughts on Siri AI?
-
Due to DMA, Siri AI Delayed in EU for iOS 27 and iPadOS 27
-
Apple Enhances Siri With On-Device AI for Faster, Private Voice Responses
-
Google AI Edge Gallery Launches on macOS With Offline Gemini Models
-
Apple iPad Air with M4 Chip Drops to $1349; Powerful On-Device LLM Inference Now More Accessible
-
Google Launches AI Edge Gallery on macOS for Running Gemini Models Locally
-
Google Launches AI Edge Gallery on macOS for Running Gemini Models Locally
-
Apple's Overhauled Siri Will Reportedly Run on Nvidia's Blackwell Chips
-
What Apple Knows About AI That Silicon Valley Won't Admit
-
Zoho-Backed Netrasemi Launches 12nm AI Chip, Mass Production Begins This Year
-
Apple Doubles Down on On-Device AI at WWDC 2026, Setting Privacy-First Strategy
-
Samsung's Exynos 2800 Brings HBM Memory to Mobile AI, Enabling Faster Local Model Inference
-
Apple's 2026 AI Strategy Prioritizes On-Device Model Deployment
-
Why AI Hardware Is a Chip Layer Problem
-
AMD Unveils Ryzen AI Halo Developer Platform for On-Device AI Workloads
-
M5 Max MacBook Runs Local Large Language Models Efficiently
-
Auditing Apple's DifferentialPrivacy.framework: Bugs, Misconfig, Practical Risks
-
Samsung's Exynos 2800 Could Be the First Mobile Chip to Use HBM for Powerful On-Device AI
-
AMD's Lemonade SDK Advances macOS Support for Local AI Inference with ROCm 7.13
-
Apple's M5 MacBook Air Advances On-Device AI with Redesigned Hardware
-
Running AI Models Locally on M4 Processors with 24GB Memory
-
Lucebox Brings Faster Local AI Inference to AMD Strix Halo
-
Mlx-serve: Run LLMs Natively on Your Mac
-
On-Device AI Market Poised for Explosive Growth as Major Tech Companies Invest Heavily
-
Major Smartphone Brands Introduce Advanced On-Device AI Features
-
Google's Gemma 4 Brings Powerful AI Capabilities to Phones and Laptops
-
Llama 4 Scout on MLX: The Complete Apple Silicon Guide (2026)
-
Running Gemma 4 on an iPhone 13 Pro
-
DFlash Doubles Token Generation Speed of Qwen3.5 27B on Mac M5 Max
-
oMLX Framework Implements DFlash Attention for Optimized Inference
-
MiniMax M2.7 Achieves SOTA Performance Under 64GB on Mac with TQ Quantization
-
DFlash Speculative Decoding Achieves 3.3x Speedup on Apple Silicon
-
Parakeet Streaming ASR on Apple Silicon via CoreML
-
AIYO Wisper: Local Voice-to-Text for macOS Using WhisperKit
-
On-Device Apple Intelligence Vulnerable to Prompt Injection Attacks
-
Running a 1.7B Parameters LLM on an Apple Watch
-
Comprehensive Benchmark: 37 LLMs Tested on MacBook Air M5 With Open-Source Tool
-
Apple Brings Enhanced On-Device AI Features to iPhone
-
Real-time Multimodal AI on Apple Silicon: Gemma E2B Demo Shows Practical Edge Deployment
-
Apple Research Shows Self-Distillation Significantly Improves Local Code Generation
-
Ollama Gets Blazing Fast on Macs with Full MLX Support and 2× Speedups
-
Mixed Precision Quantization on MLX with TurboQuant Implementation
-
Kokoro TTS Achieves 20× Realtime Speed on CPU-Only On-Device Inference
-
Gemma 4 26B A4B Outperforms Qwen 3.5 35B on Apple Silicon
-
Google Gemma 4 Released with GGUF Quantizations
-
OpenUMA – Apple-Style Unified Memory for x86 AI Inference
-
TinyGPU Adds Mac Support for External Nvidia GPU Acceleration
-
Apple Silicon Macs Run Local AI Faster with Ollama's New MLX Support
-
Is Anyone Working on an AI Operating System?
-
Ollama Adopts Apple's MLX Framework for Faster Local AI on Mac
-
Select the Right Hardware for Your Local LLM Deployment with This Online Guide
-
M5 Max Delivers 1.7x Faster Inference Than M3 Max on Qwen 3.5 Models
-
TurboQuant KV Cache Compression Achieves 22.8% Faster Decoding at 32K Context
-
Apple Gets Full Gemini Access and Uses Distillation to Build Lightweight On-Device AI
-
mlx-Code: Run Claude Code Locally with MLX-LM
-
Apple Plans Slimmed-Down Gemini Models for Local iPhone AI Features
-
Running an Open-Weight LLM Locally on an Apple Watch
-
Open-Source Tool Helps Determine Which Local LLMs Run on Your PC
-
Ditching Paid AI Services: Building Self-Hosted LLM Solutions as ChatGPT, Claude, and Gemini Alternatives
-
Multi-Token Prediction support coming to MLX-LM for Qwen 3.5
-
DeepSeek R1 RTX 4090 vs Apple M3 Max: Benchmark & Performance Guide
-
Apple M5 Max 128GB real-world performance benchmarks for local inference
-
Apple's On-Device AI Raises Privacy Alarms Across British Parliament
-
Startup Transforms Mac Mini Into Full-Powered AI Inference System With External GPU
-
AMD Launches Agent System Optimized for Local AI Inference With Ryzen and Radeon
-
Local LLMs on Apple Silicon Mac 2026: M1 M2 M3 Guide
-
Apple M5 Max 128GB Benchmark Results for Local LLM Inference
-
M5 Max and M5 Ultra Chipsets Demonstrate Significant Bandwidth Improvements for Local LLM Inference
-
Apple Launches MacBook Neo with A18 Pro Chip for Affordable Local AI Inference
-
Windows 11 Notepad Gets On-Device AI Text Generation Without Subscription
-
Real-World Qwen 3.5 9B Agent Performance on M1 Pro Validates Edge Deployment
-
Apple Unveils MacBook Pro with M5 Pro and M5 Max Featuring On-Device AI
-
Apple Unveils MacBook Pro With M5 Pro and M5 Max for On-Device AI
-
Apple M5 Pro and M5 Max: 4× Faster LLM Processing
-
AMD Launches Copilot+ Desktop Chips to Compete in On-Device AI Market
-
Qualcomm Snapdragon Wear Elite: 2B Parameter NPU for Personal AI Wearables
-
Apple M4 iPad Air Targets AI Users with Double M1 Speed Performance
-
Alibaba's Qwen 3.5 Small Model Runs Directly on iPhone 17
-
Qualcomm Launches Snapdragon Wear Elite for On-Device AI on Wearables
-
AMD Expands Ryzen AI 400 Series Portfolio for Consumer and Enterprise AI PC Options
-
Running Local AI Models on Mac Studio 128GB: 4B, 20B & 120B Tested
-
Apple Neural Engine Reverse-Engineered for Local Model Training on Mac Mini M4
-
Apple Intelligence, Galaxy AI, Gemini: Why Your AI-Powered Phone Is Worth Repairing
-
Apple: Python bindings for access to the on-device Apple Intelligence model
-
Apple Accelerates U.S. Manufacturing with Mac Mini Production
-
Qwen3-Code-Next Proves Practical for Local Development: Real-World Coding Tasks on Mac Studio
-
AI-Powered Reverse-Engineering of Rosetta 2 for Linux
-
Apple Researchers Develop On-Device AI Agent That Interacts With Apps for You
-
GPT4All Replaces Ollama On Mac After Quick Trial
-
Sourdine: Open-Source macOS App for 100% Local AI Transcription