INCModel2/Qwen3.6-35B-A3B-MXFP4-Mixed-CT-AutoRound Image-Text-to-Text • 36B • Updated about 9 hours ago
INCModel2/Qwen3.6-35B-A3B-MXFP4-Mixed-CT-AutoRound Image-Text-to-Text • 36B • Updated about 9 hours ago
INCModel2/Hy3-MXFP4-Mixed-CT-AutoRound-Preview Text Generation • 163B • Updated 27 days ago • 1.48k • 1
INCModel2/Hy3-MXFP4-Mixed-CT-AutoRound-Preview Text Generation • 163B • Updated 27 days ago • 1.48k • 1
INCModel2/DeepSeek-V4-Pro-DSpark-MXFP4-Mixed-CT-AutoRound Text Generation • 889B • Updated Jun 30 • 79 • 2
INCModel2/DeepSeek-V4-Pro-DSpark-MXFP4-Mixed-CT-AutoRound Text Generation • 889B • Updated Jun 30 • 79 • 2
TEQ: Trainable Equivalent Transformation for Quantization of LLMs Paper • 2310.10944 • Published Oct 17, 2023 • 10
Optimize Weight Rounding via Signed Gradient Descent for the Quantization of LLMs Paper • 2309.05516 • Published Sep 11, 2023 • 14
Optimize Weight Rounding via Signed Gradient Descent for the Quantization of LLMs Paper • 2309.05516 • Published Sep 11, 2023 • 14