Inference Providers
Active filters: 3-bit
mlx-community/Laguna-S-2.1-oQ3e
Text Generation
• 13B • Updated • 502
• 3
MaziyarPanahi/Llama-3-8B-Instruct-64k-GGUF
Text Generation
• 8B • Updated • 57.1k
• 14
mlx-community/GLM-4.5-Air-3bit
Text Generation
• 107B • Updated • 603
• 32
mlx-community/moonshotai_Kimi-K2-Instruct-mlx-3bit
Text Generation
• 1T • Updated • 292
• 2
mlx-community/Kimi-K2-Instruct-0905-mlx-3bit
Text Generation
• 1T • Updated • 113
• 2
MaziyarPanahi/NVIDIA-Nemotron-Nano-12B-v2-GGUF
Text Generation
• 12B • Updated • 29.3k
• 4
MaziyarPanahi/VulnLLM-R-7B-GGUF
Text Generation
• 8B • Updated • 144
• 1
MoringLabs/Nemotron-3-Super-120B-A12B-MLX-3.6bit
Text Generation
• 121B • Updated • 504
• 7
MoringLabs/Qwen3.5-122B-A10B-MLX-3.7bit-VL-v2
Image-Text-to-Text
• 17B • Updated • 493
• 1
unsloth/Qwen3.6-35B-A3B-UD-MLX-3bit
Image-Text-to-Text
• 5B • Updated • 2.84k
• 13
manjunathshiva/Nemotron-3-Super-120B-A12B-tq3
Text Generation
• 129B • Updated • 529
• 2
marcinrogalski/Qwen3.5-122B-A10B-oQ3
122B • Updated • 112
• 2
deucebucket/Qwen3.6-35B-A3B-Heretic-Cerebellum-GGUF
Image-Text-to-Text
• 35B • Updated • 2.71k
• 5
rayray916/Qwen3.6-35B-A3B-uncensored-heretic-Native-MTP-Preserved-oQ3
Text Generation
• 5B • Updated • 1.52k
• 1
NapYang/Ornith-1.0-35B-MLX-oQ3.5-fp16
35B • Updated • 745
• 1
kernelpool/LongCat-2.0-3bit
Text Generation
• 1.6T • Updated • 628
• 2
mvid/Leanstral-1.5-119B-A6B-MLX-3bit
Image-Text-to-Text
• 16B • Updated • 294
• 1
avlp12/Inkling-975B-Alis-MLX-Dynamic-3.7bpw
Image-Text-to-Text
• 947B • Updated • 415
• 1
kaitchup/Llama-2-7b-gptq-3bit
Text Generation
• Updated • 4
clibrain/Llama-2-7b-ft-instruct-es-gptq-3bit
Text Generation
• Updated • 2
• 3
clibrain/Llama-2-13b-ft-instruct-es-gptq-3bit
Text Generation
• Updated • 6
• 3
MiNeves-tops/opt-125m-gptq-3bit
Text Generation
• Updated • 4
Text Generation
• Updated • 4
LoneStriker/Yi-6B-200K-3.0bpw-h6-exl2
Text Generation
• Updated • 2
danny0122/Llama-2-7b-hf-gptq-3bits
Text Generation
• 6B • Updated • 2
danny0122/Llama-2-7b-hf-gptq-3bitssafe
Text Generation
• 6B • Updated • 3
danny0122/stablelm-base-alpha-3b-gptq-3bits
Text Generation
• 3B • Updated • 2
danny0122/stablelm-base-alpha-3b-gptq-3bitssafe
Text Generation
• 3B • Updated • 2
SicariusSicariiStuff/Tenebra_PreAlpha_128g_3BIT
Text Generation
• 31B • Updated • 4
mahihossain666/llama-2-70b-hf-quantized-3bits-GPTQ
Text Generation
• 65B • Updated • 2