Saad Safi
saadsafi
AI & ML interests
None yet
Recent Activity
liked a model about 15 hours ago
wasifb/ThinkingCap-Qwen3.8-27B-AutoRound-W4A16 liked a model 2 days ago
ultimaterex/Swift-Qwen3.8-27B-Uncensored-W4A16-AutoRound new activity 2 days ago
unsloth/Qwen-Image-2.1-GGUF:backend 'cuda0' was not foundOrganizations
None yet
backend 'cuda0' was not found
1
#1 opened 3 days ago
by
saadsafi
error loading with vllm 0.30.0
6
#1 opened 3 days ago
by
saadsafi
IQ4_XS is broken
3
#1 opened 7 days ago
by
saadsafi
UD-Q2_K_XL-E216-W1280/DeepSeek-V4-Flash-0731-UD-Q2_K_XL-00001-of-00003.gguf
#1 opened 29 days ago
by
saadsafi
Duplication?
#1 opened 30 days ago
by
saadsafi
0xSero/DeepSeek-V4-Flash-0731-REAP
1
#2743 opened about 2 months ago
by
saadsafi
vectionlabs/Salience-1.5-Pro
1
#2706 opened 2 months ago
by
saadsafi
VLLM NotImplementedError in vllm/model_executor/layers/quantization/base_config.py
#1 opened 3 months ago
by
saadsafi
probably an issue with vllm 0.23.0
#2 opened 3 months ago
by
saadsafi
MTP support
14
#3 opened 4 months ago
by
Throghar
torch RuntimeError: Shape mismatch: a.size(1) = 4096, size_k = 8192
2
#1 opened 4 months ago
by
saadsafi
QWEN35_MTP requires nextn_predict_layers > 0.
➕👍 5
14
#2 opened 5 months ago
by
jidaigeist
Low quality code generated with latest llama.cpp
2
#1 opened 5 months ago
by
saadsafi
gemma-4-26b-a4b-it-q5_k_m
1
#1 opened 5 months ago
by
saadsafi
running "MIXED" gguf with latest llama.cpp gave this error:
1
#1 opened 5 months ago
by
saadsafi
tawkeed-sa/tawkeed-40b
2
#2069 opened 6 months ago
by
saadsafi
Intel/Qwen3.5-122B-A10B-int4-AutoRound
1
#1919 opened 7 months ago
by
saadsafi
https://huggingface.co/inceptionai/Jais-2-70B-Chat
2
#1640 opened 9 months ago
by
saadsafi
https://huggingface.co/inceptionai/Jais-2-8B-Chat
2
#1641 opened 9 months ago
by
saadsafi