Inference Providers
Active filters: 4-bit
stefanprodan/Apodex-1.1-mini-oQ4e-mtp
Text Generation
• 36B • Updated • 295
• 3
albucino/Qwen3.8-Flash-Next-W4A16-FP8PLE
Text Generation
• 73B • Updated • 310
• 4
ddark-il/Qwen3.8-Flash-Next-Uncensored-Mixed-omlx
Image-Text-to-Text
• 180B • Updated • 1.51k
• 2
lovesenko/GLM-5.3-Flash-tr3-4bpw-Abliterated
Image-Text-to-Text
• 88B • Updated • 111
• 2
ibm-granite/granite-4.2-3b-q4-mlx
4B • Updated • 119
• 2
btbtyler09/Qwen3.8-Flash-Next-GPTQ-4bit
Image-Text-to-Text
• 180B • Updated • 262
• 2
leoncca/Qwen3.8-Flash-Next-AWQ-g32
Image-Text-to-Text
• 181B • Updated • 123
• 2
abenzerps/K2-Horizon-7B-MLX-4bit
Text Generation
• 9B • Updated • 580
• 3
JC1DA/Qwopus3.8-27B-Flash-INT4-W4A16
Image-Text-to-Text
• 6B • Updated • 89
• 2
Solstice-AI/GLM-5.3-Flash-UNCENSORED-mlx-oQ4e-DSpark
Image-Text-to-Text
• 321B • Updated • 2
TheBloke/llama2_7b_chat_uncensored-GPTQ
Text Generation
• 7B • Updated • 98
• 75
AlignmentLab-AI/sentinelv2
Updated • 10
• 2
unsloth/llama-3-8b-Instruct-bnb-4bit
Text Generation
• 8B • Updated • 60.9k
• 135
MaziyarPanahi/Phi-3-Context-Obedient-RAG-GGUF
Text Generation
• 4B • Updated • 1.1k
• 7
FallenMerick/Space-Whale-Lite-13B-GGUF
Text Generation
• 13B • Updated • 25
• 1
MaziyarPanahi/Llama-3-Groq-8B-Tool-Use-GGUF
Text Generation
• 8B • Updated • 1.1k
• 9
unsloth/Meta-Llama-3.1-8B-Instruct-bnb-4bit
Text Generation
• 8B • Updated • 101k
• 104
unsloth/Llama-3.2-11B-Vision-Instruct-bnb-4bit
Image-Text-to-Text
• 11B • Updated • 6.06k
• 87
mlx-community/Llama-3.3-70B-Instruct-4bit
Text Generation
• 71B • Updated • 5.94k
• 36
Text Generation
• 15B • Updated • 1.51k
• 5
mlx-community/DeepSeek-R1-Distill-Llama-70B-4bit
Text Generation
• 71B • Updated • 1.41k
• 12
mlx-community/Llama-3.1-8B-Instruct-4bit
Text Generation
• 8B • Updated • 804k
• 4
elenapop/Llama-3.2-11B-Vision-ocr-working_model
Image-Text-to-Text
• Updated • 26
• 1
MaziyarPanahi/gemma-3-1b-it-GGUF
Text Generation
• 1.0B • Updated • 151k
• 13
Qwen/Qwen2.5-VL-32B-Instruct-AWQ
Image-Text-to-Text
• 33B • Updated • 1.69M
• 65
unsloth/Qwen3-30B-A3B-bnb-4bit
31B • Updated • 3.16k
• 22
mlx-community/Qwen3-0.6B-4bit
Text Generation
• 0.6B • Updated • 50.8k
• 17
mlx-community/Qwen3-1.7B-4bit
Text Generation
• 2B • Updated • 14.7k
• 7
mlx-community/Qwen3-4B-4bit
Text Generation
• 4B • Updated • 17.6k
• 16
mlx-community/DeepSeek-R1-0528-Qwen3-8B-4bit
Text Generation
• 8B • Updated • 1.94k
• 8