Briefing

Qwen3.6‑35B‑A3B Uncensored Heretic Model with Preserved MTP Tensors

ai-dev
by /u/LLMFan46 ·

Publish the Qwen3.6‑35B‑A3B model with preserved MTP tensors and run the provided benchmark to confirm MTP preservation.

What to do now

Publish the model to your deployment pipeline and run the provided benchmark to confirm MTP preservation.

Summary

A new Qwen3.6‑35B‑A3B‑uncensored‑heretic‑Native‑MTP‑Preserved model has been released by llmfan46 on HuggingFace. The model retains the full MTP tensor count across all supported formats, including safetensors, GGUF, and GPTQ‑Int4. In safetensors format the MTP tensors appear as 19 entries because the gate_up_proj is stored as one fused tensor. In GGUF format the fused tensor is split into separate gate/up expert tensors, resulting in 20 entries. All releases have been verified to preserve the MTP components. The model is available in multiple variants: native, GGUF, NVFP4‑Experts‑Only, and NVFP4‑Experts‑Only‑GGUF. Benchmarks are provided alongside the releases to demonstrate performance. The model is hosted under the llmfan46 collection on HuggingFace.

Key changes

  • Qwen3.6‑35B‑A3B‑uncensored‑heretic‑Native‑MTP‑Preserved released
  • Full MTP tensor count preserved across safetensors, GGUF, GPTQ‑Int4 formats
  • Safetensors shows 19 entries due to fused gate_up_proj
  • GGUF splits fused tensor into 20 entries
  • All releases verified to preserve MTP components
  • Available in native, GGUF, NVFP4‑Experts‑Only, NVFP4‑Experts‑Only‑GGUF variants
  • Benchmarks included
  • Hosted under llmfan46 on HuggingFace

Affects

internal

Customer impact

Analyzing matches…

Ask about this story

Impact on an agency? Which customers? Compare historically Risks of waiting