Briefing

Unsloth Releases Qwen3.6 27B and 35B A3B Models with Preserved MTP Layer

ai-dev
by /u/Altruistic_Heat_9531 · Llama

Clone the Unsloth HF repos and apply the llama‑cpp MTP PR to enable MTP support for Qwen3.6 models.

What to do now

Clone the Unsloth HF repos and apply the llama‑cpp MTP PR to enable MTP support for Qwen3.6 models.

Summary

Unsloth has released Qwen3.6-27B-GGUF-MTP and Qwen3.6-35B-A3B-GGUF-MTP on Hugging Face, preserving the MTP layer for better performance. The models require checking out and building a llama‑cpp PR that adds MTP support. Unsloth provides instructions in the model card on how to use MTP with the models. The release allows developers to run Qwen3.6 models with MTP enabled, which can improve inference speed and memory efficiency.

This is a new model release that expands the available Qwen3.6 options for developers seeking MTP support.

The information is relevant for those who want to leverage MTP in their local LLM deployments.

Key changes

  • Unsloth provides Qwen3.6-27B-GGUF-MTP and Qwen3.6-35B-A3B-GGUF-MTP
  • models have preserved MTP layer
  • need to checkout and build llama‑cpp PR about MTP
  • Unsloth gives instructions in model card
  • use HF links for download

Affects

internal

Customer impact

Analyzing matches…

Ask about this story

Impact on an agency? Which customers? Compare historically Risks of waiting