Tagged "reasoning-models"
4 articles tagged reasoning-models, 12 April 2026 to 23 September 2026. Newest first.
-
Ollama v0.34.4 Adds Structured Outputs for Reasoning Models
The latest Ollama release includes structured output support for thinking models and fixes intermittent model loading errors, improving reliability for local LLM deployments.
-
llama.cpp Adds DeepSeek V4 Flash Chat Template Support
llama.cpp now includes updated chat templates for DeepSeek V4 Flash models, enabling proper local inference with thinking token handling for the latest reasoning model.
-
AMD's vLLM-ATOM Plugin Supercharges DeepSeek-R1 and Kimi-K2 Inference on MI350/MI400
AMD has released a vLLM-ATOM plugin optimizing inference for DeepSeek-R1, Kimi-K2, and gpt-oss-120B models on Instinct MI350 and MI400 accelerators, delivering significant performance gains for local deployment.
-
MiniMax M2.7 Is Now Open Source
MiniMax releases M2.7, an agentic model now available as open source, expanding options for local deployment of capable reasoning models without cloud dependencies.