Inference Providers
Active filters: awq
prism-ml/Ternary-Bonsai-27B-AWQ-4bit
Image-Text-to-Text
• 27B • Updated • 3.07k
• • 15
prism-ml/Bonsai-27B-AWQ-4bit
Image-Text-to-Text
• 27B • Updated • 846
• 7
QuantTrio/Qwen3.6-27B-AWQ
Image-Text-to-Text
• 28B • Updated • 1.27M
• 22
Qwen/Qwen2.5-VL-32B-Instruct-AWQ
Image-Text-to-Text
• 33B • Updated • 300k
• 65
cyankiwi/Qwen3.6-35B-A3B-AWQ-4bit
Image-Text-to-Text
• 36B • Updated • 1.75M
• 87
pearsonkyle/gemma4-31b-imatrix-mtp-GGUF
Image-Text-to-Text
• 31B • Updated • 14.1k
• 3
casperhansen/llama-3-8b-instruct-awq
Text Generation
• 8B • Updated • 28.2k
• 31
Text Generation
• 8B • Updated • 586k
• 52
Text Generation
• 4B • Updated • 550k
• 31
openbmb/MiniCPM-o-4_5-awq
Any-to-Any
• 9B • Updated • 36.9k
• 22
QuantTrio/Qwen3.5-122B-A10B-AWQ
Image-Text-to-Text
• 125B • Updated • 28.2k
• 29
alonsoko/gemma-4-31b-it-abliterated-heretic-AWQ-W4A16
Image-Text-to-Text
• 32B • Updated • 8.01k
• 14
prism-ml/Bonsai-8B-AWQ-4-bit
Text Generation
• 8B • Updated • 99
• 5
mattbucci/Qwen3.6-27B-AWQ
27B • Updated • 43.2k
• 5
shawnw3i/Huihui-Qwen3.6-27B-abliterated-AWQ-MTP
Image-Text-to-Text
• 6B • Updated • 32.4k
• 11
spectator2026/MiniMax-M3-AWQ-int4
Image-Text-to-Text
• 69B • Updated • 1.86k
• 3
Avesed/Qwen3.6-27B-INT4-W4A16
Text Generation
• 28B • Updated • 1.8k
• 1
sahilchachra/Qwythos-9B-Claude-Mythos-5-1M-AWQ
Text Generation
• 9B • Updated • 6.51k
• 3
philbert440/Qwen3.6-27B-Uncensored-Cyber-W4A16-AWQ
28B • Updated • 357
• 1
LostGentoo/Qwen3-Embedding-8B-AWQ
Feature Extraction
• 8B • Updated • 159
• 1
8B • Updated • 18
• 1
casperhansen/mpt-7b-8k-chat-awq
Text Generation
• Updated • 16
• 3
casperhansen/falcon-7b-awq
Text Generation
• Updated • 16
• 1
casperhansen/vicuna-7b-v1.5-awq
Text Generation
• Updated • 7
• 3
casperhansen/vicuna-7b-v1.5-awq-gemv
Text Generation
• Updated • 6
• 1
casperhansen/mpt-7b-8k-chat-awq-gemv
Text Generation
• Updated • 9
casperhansen/opt-125m-awq
Text Generation
• 0.2B • Updated • 64
• 2
casperhansen/tinyllama-1b-awq
Text Generation
• Updated • 27
• 1
Bomml/Llama-2-70B-chat-w4-g128-awq
Text Generation
• Updated TheBloke/Llama-2-7B-Chat-AWQ
Text Generation
• 7B • Updated • 2.08k
• 24