Inference Providers
Active filters: autoawq
hugging-quants/Meta-Llama-3.1-8B-Instruct-AWQ-INT4
Text Generation
• 8B • Updated • 187k
• 92
mconcat/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled-v2-AWQ-4bit
Text Generation
• 29B • Updated • 711
• 2
JunHowie/Qwythos-9B-Claude-Mythos-5-1M-AWQ
Text Generation
• 9B • Updated • 2.11k
• 1
Text Generation
• 7B • Updated • 87
• 2
Text Generation
• 6B • Updated • 10
kaitchup/Llama-3-8b-awq-4bit
Text Generation
• 8B • Updated • 5
XavierSpycy/Meta-Llama-3-8B-Instruct-zh-10k
Text Generation
• 8B • Updated • 13
• XavierSpycy/Meta-Llama-3-8B-Instruct-zh-10k-GGUF
Text Generation
• 8B • Updated • 42
XavierSpycy/Meta-Llama-3-8B-Instruct-zh-10k-GPTQ
Text Generation
• 8B • Updated • 4
XavierSpycy/Meta-Llama-3-8B-Instruct-zh-10k-AWQ
Text Generation
• 8B • Updated • 5
hugging-quants/Meta-Llama-3.1-405B-Instruct-AWQ-INT4
Text Generation
• 410B • Updated • 1.68k
• 36
hugging-quants/Meta-Llama-3.1-70B-Instruct-AWQ-INT4
Text Generation
• 71B • Updated • 258k
• 110
jburmeister/Meta-Llama-3.1-70B-Instruct-AWQ-INT4
Text Generation
• 71B • Updated jburmeister/Meta-Llama-3.1-405B-Instruct-AWQ-INT4
Text Generation
• 410B • Updated • 1
Kalei/Meta-Llama-3.1-70B-Instruct-AWQ-INT4-Custom
Text Generation
• 71B • Updated • 1
UCLA-EMC/Meta-Llama-3.1-8B-AWQ-INT4
Text Generation
• 8B • Updated • 5
UCLA-EMC/Meta-Llama-3.1-8B-Instruct-AWQ-INT4-32-2.17B
Text Generation
• 8B • Updated • 1.77k
• 1
reach-vb/Meta-Llama-3.1-8B-Instruct-AWQ-INT4-fix
Text Generation
• 8B • Updated • 4
jburmeister/Meta-Llama-3.1-8B-Instruct-AWQ-INT4
Text Generation
• 8B • Updated • 8.32k
awilliamson/Meta-Llama-3.1-70B-Instruct-AWQ
Text Generation
• 71B • Updated • 2
flowaicom/Flow-Judge-v0.1-AWQ
Text Generation
• 4B • Updated • 2.61k
• 6
hugging-quants/Mixtral-8x7B-Instruct-v0.1-AWQ-INT4
Text Generation
• 47B • Updated • 15.4k
hugging-quants/gemma-2-9b-it-AWQ-INT4
Text Generation
• 9B • Updated • 228k
• 9
ibnzterrell/Nvidia-Llama-3.1-Nemotron-70B-Instruct-HF-AWQ-INT4
Text Generation
• 71B • Updated • 458
• 6
NeuML/Llama-3.1_OpenScholar-8B-AWQ
Text Generation
• 8B • Updated • 4
• 3
fbaldassarri/TinyLlama_TinyLlama_v1.1-autoawq-int4-gs128-asym
Text Generation
• 1B • Updated • 3
fbaldassarri/TinyLlama_TinyLlama_v1.1-autoawq-int4-gs128-sym
Text Generation
• 1B • Updated • 3
fbaldassarri/EleutherAI_pythia-14m-autoawq-int4-gs128-asym
Text Generation
• 14.1M • Updated • 8
fbaldassarri/EleutherAI_pythia-14m-autoawq-int4-gs128-sym
Text Generation
• 14.1M • Updated • 3
fbaldassarri/EleutherAI_pythia-31m-autoawq-int4-gs128-asym
Text Generation
• 30.5M • Updated • 2