roshniramesh 's Collections int8 llm
updated
meta-llama/Llama-Guard-3-8B-INT8
Text Generation
• 8B • Updated • 9.23k
• 38
google/gemma-7b-quant-pytorch
Text Generation
• Updated • 57
• 2
INC4AI/gpt-j-6B-int8-dynamic-inc
Text Generation
• Updated • 17
• 16
Intel/t5-small-xsum-int8-dynamic-inc
Updated • 1.33k
• 1
INC4AI/bert-base-uncased-mrpc-int8-static-inc
Text Classification
• Updated • 7
Intel/bert-large-uncased-cola-int8-inc
Text Classification
• Updated • 8
INC4AI/vit-base-patch16-224-int8-static-inc
Image Classification
• Updated • 12
• 1
INC4AI/albert-base-v2-sst2-int8-static-inc
Text Classification
• Updated • 39
Intel/roberta-base-mrpc-int8-dynamic-inc
Text Classification
• Updated • 8
INC4AI/roberta-base-mrpc-int8-static-inc
Text Classification
• Updated • 16
Intel/dynamic-minilmv2-L6-H384-squad1.1-int8-static
Question Answering
• 30.1M • Updated • 12
Intel/MiniLM-L12-H384-uncased-mrpc-int8-dynamic-inc
Text Classification
• Updated • 7
INC4AI/gpt-j-6B-int8-static-inc
Text Generation
• Updated • 19
• 9
INC4AI/gpt-j-6B-pytorch-int8-static-inc
Text Generation
• Updated • 14
Intel/bert-base-uncased-CoLA-int8-inc
Text Classification
• Updated • 13
Intel/bert-base-uncased-STS-B-int8-inc
Text Classification
• Updated • 10
INC4AI/bert-base-uncased-mrpc-int8-qat-inc
Text Classification
• Updated • 12
• 1
Intel/bert-large-uncased-rte-int8-dynamic-inc
Text Classification
• Updated • 12
Intel/bert-large-uncased-rte-int8-static-inc
Text Classification
• Updated • 12
Intel/distilbert-base-uncased-distilled-squad-int8-static-inc
Question Answering
• Updated • 536
• 5
Intel/distilbert-base-uncased-MRPC-int8-dynamic-inc
Text Classification
• Updated • 12
• 1
Intel/distilbert-base-uncased-MRPC-int8-static-inc
Text Classification
• Updated • 17
INC4AI/albert-base-v2-sst2-int8-dynamic-inc
Text Classification
• Updated • 13
Intel/albert-base-v2-MRPC-int8-inc
Text Classification
• Updated • 7
Intel/bge-small-en-v1.5-rag-int8-static
Feature Extraction
• Updated • 34
• 2
Intel/bge-base-en-v1.5-rag-int8-static
Feature Extraction
• Updated • 9
INC4AI/falcon-7b-sq-int8-inc
Text Generation
• Updated • 21
amd/Llama-3.1-8B-Instruct-w-int8-a-int8-sym-test
8B • Updated • 9.01k
RedHatAI/Llama-3.2-1B-Instruct-quantized.w8a8
Text Generation
• 1B • Updated • 25.2k
• 8
FriendliAI/Meta-Llama-3-8B-int8
Text Generation
• 8B • Updated • 8
• 1
google/gemma-7b-it-quant-pytorch
Text Generation
• Updated • 60
• 11
OpenVINO/mistral-7b-instruct-v0.1-int8-ov
Text Generation
• Updated • 15
• 1
FriendliAI/Meta-Llama-3.1-8B-Instruct-int8
Text Generation
• 8B • Updated • 15.3k
• 1
Text Generation
• 14B • Updated • 153
• 7
Text Generation
• 8B • Updated • 218
• 9
Text Generation
• 2B • Updated • 194
• 5
Qwen/Qwen1.5-1.8B-Chat-GPTQ-Int8
Text Generation
• 2B • Updated • 139
• 2
Qwen/Qwen1.5-14B-Chat-GPTQ-Int8
Text Generation
• 15B • Updated • 142
• 11
Qwen/Qwen1.5-4B-Chat-GPTQ-Int8
Text Generation
• 4B • Updated • 124
• 6
Qwen/Qwen1.5-72B-Chat-GPTQ-Int8
Text Generation
• 72B • Updated • 134
• 7
Qwen/Qwen1.5-4B-Chat-GGUF
Text Generation
• 4B • Updated • 841
• 16
Qwen/Qwen1.5-0.5B-Chat-GGUF
Text Generation
• 0.6B • Updated • 11.5k
• 35
Qwen/Qwen1.5-7B-Chat-GGUF
Text Generation
• 8B • Updated • 857
• 71
Qwen/CodeQwen1.5-7B-Chat-GGUF
Text Generation
• 7B • Updated • 1.27k
• 111
Qwen/Qwen2.5-1.5B-Instruct-GPTQ-Int8
Text Generation
• 2B • Updated • 514
• 6
Qwen/Qwen2.5-0.5B-Instruct-GPTQ-Int8
Text Generation
• 0.5B • Updated • 656
• 10
Qwen/Qwen2.5-0.5B-Instruct-GGUF
Text Generation
• 0.6B • Updated • 150k
• 123
Qwen/Qwen2-1.5B-Instruct-GGUF
Text Generation
• 2B • Updated • 16.7k
• 31
Qwen/Qwen2-0.5B-Instruct-GGUF
Text Generation
• 0.5B • Updated • 8.33k
• 76
Qwen/Qwen2-7B-Instruct-GGUF
Text Generation
• 8B • Updated • 5.89k
• 180
Qwen/Qwen2-0.5B-Instruct-GPTQ-Int8
Text Generation
• 0.6B • Updated • 197
• 4
Qwen/Qwen2-1.5B-Instruct-GPTQ-Int8
Text Generation
• 2B • Updated • 139
• 4
Qwen/Qwen2-7B-Instruct-GPTQ-Int8
Text Generation
• 8B • Updated • 336
• 17
Qwen/Qwen2-72B-Instruct-GPTQ-Int8
Text Generation
• 73B • Updated • 528
• 15