Tagged "self-hosted-llm"
- Open-Weight AI on Kubernetes: Comparing vLLM and KubeAI for Local Deployment
- Netflix Details Its In-House LLM Serving Platform with Triton and vLLM
- Multiverse Computing's CompactifAI Models Now Fully Compatible with Intel Xeon 6 Processors
- Open-Source AI on OCI: Serving LLMs on Kubernetes with vLLM, Qdrant, and Terraform
- Nvidia Boosts Token Throughput 5x With Software Optimizations, Reshaping AI Inference Economics
- If You Can Write Acceptance Criteria, You Can Write an AI Routing Policy
- Using mirrord to Verify AI-SRE Fixes Against Staging Clusters
- DeepSWE Benchmark Updated with GLM 5.2 and Expanded Model Comparisons
- Show HN: SpadeBox – Sandboxed tools and JavaScript runtime for AI agents
- Due to DMA, Siri AI Delayed in EU for iOS 27 and iPadOS 27
- Running Espressif's OpenClaw-Inspired AI Agent on ESP32 with Self-Hosted LLM Works in Practice
- Running DeepSeek R1 Locally: Your Complete Setup Guide
- I Built a Local AI Stack with 5 Docker Containers, and Now I'll Never Pay for ChatGPT Again
- Self-Hosted LLM Elevates Personal Knowledge Management Systems to New Levels
- LoKI – Local AI Assistant for Linux and WSL
- Search and Analyze Documents from the DOJ Epstein Files Release with Local LLM