2.5x faster inference with Qwen 3.6 27B using MTP - Finally a viable option for local agentic coding - 262k context on 48GB - Fixed chat template - Drop-in OpenAI and Anthropic API endpoints
A r/LocalLLaMA discussion (406 comments) that references Anthropic API in the course of a broader conversation.
