Qwen Flash Next MTP 작업 재개
Qwen Flash Next MTP work restarted
핵심 요약
Qwen Flash Next의 MTP 구현 작업이 llama.cpp에서 다시 진행되고 있습니다.
- MTP 구현 — llama.cpp에서 Qwen Flash Next를 위한 MTP 작업이 재개됨
- GGUF 양자화 — Hugging Face를 통해 관련 GGUF 파일 제공
- 작업 상태 — 현재 해당 PR은 아직 개발 중인 상태임
- 파일 다운로드 — 전체를 받을 필요 없이 mtp-* 파일만 선택적으로 다운로드 가능
If you're using MTP with Qwen Flash Next and llama.cpp, you can switch to:
quants: https://huggingface.co/ggml-org/Qwen3.8-Flash-Next-GGUF
PR: https://github.com/ggml-org/llama.cpp/pull/29761
Please note that this is still wip

