These are quantizations of the model Jackrong / Qwopus3.5-9B-Coder
I've added the MTP layer on it.
My personal speed improvement on my 7900XTX with the vulkan backend has been from ~80 tps to around ~120 tps.
An imatrix has been calulated for coding tasks, as such it is specialized for coding.
Quick Start
- Download the latest release of llama.cpp.
- Download your preferred model variant from below.
- Downloads last month
- 599
Hardware compatibility
Log In to add your hardware
Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐ Ask for provider support
Model tree for noctrex/Qwopus3.5-9B-Coder-MTP
Base model
Qwen/Qwen3.5-9B-Base Finetuned
Qwen/Qwen3.5-9B Finetuned
unsloth/Qwen3.5-9B Finetuned
Jackrong/Qwopus3.5-9B-v3.5 Adapter
Jackrong/Qwopus3.5-9B-Coder