Inference Providers
Active filters: mlx-lm
froggeric/Qwen3.6-35B-A3B-Uncensored-Heretic-MLX-4bit
Image-Text-to-Text
• 35B • Updated • 8.96k
• 54
TensorFold/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-MLX-4bit
Text Generation
• 32B • Updated • 955
• 2
TensorFold/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-oQ4
Text Generation
• 32B • Updated • 439
• 1
guruswami-ai/deepseek-v4-pro-0813-mlx-tp4-recipe
Text Generation
• Updated • 2
barozp/Qwen3.8-Whittle-MoE-27B-MLX-4bit
Text Generation
• 27B • Updated • 2.01k
• 6
guruswami-ai/glm-5.3-mlx-mixed-4_8bit-mtp-recipe
Text Generation
• Updated • 2
abenzerps/K2-Horizon-7B-MLX-4bit
Text Generation
• 9B • Updated • 2.43k
• 4
kacperbb/phi-3.5-mlx-finetuned
Updated • 16
fourbic/disarm-ew-llama3-finetuned
Text Generation
• 8B • Updated • 10
alexander-model/gpt_model_safe
Text Generation
• 21B • Updated • 30
halley-ai/gpt-oss-120b-MLX-8bit-gs32
Text Generation
• 117B • Updated • 76
• 1
halley-ai/gpt-oss-120b-MLX-bf16
Text Generation
• 117B • Updated • 271
• 3
halley-ai/gpt-oss-120b-MLX-6bit-gs64
Text Generation
• 117B • Updated • 81
• 1
LibraxisAI/Qwen3-Next-80B-A3B-Instruct-MLX-MXFP4
Text Generation
• 80B • Updated • 180
• 2
halley-ai/Qwen3-Next-80B-A3B-Instruct-MLX-4bit-gs64
Text Generation
• 80B • Updated • 45
• 1
halley-ai/Qwen3-Next-80B-A3B-Instruct-MLX-5bit-gs32
Text Generation
• 80B • Updated • 30
• 1
halley-ai/Qwen3-Next-80B-A3B-Instruct-MLX-6bit-gs64
Text Generation
• 80B • Updated • 20
• 1
Miemczyk/CharityPurposeAnalyser
Text Generation
• Updated ProbioticFarmer/toucan-qwen3-8b-lora
Text Generation
• Updated Text Generation
• Updated pherber3/Qwen3-Omni-30B-A3B-Instruct-4bit-mlx
31B • Updated • 240
• 8
Daizee/Gemma3-Callous-Calla-4B-mlx
Text Generation
• Updated • 23
Daizee/Dirty-Calla-4B-mlx
Text Generation
• Updated • 73
meetmerchant/tech-tweet-generator-llama3
Updated
codewithdark/Llama-3.2-3B-4bit-mlx
Text Generation
• 3B • Updated • 17
QuantLLM/Llama-3.2-3B-4bit-mlx
Text Generation
• 3B • Updated • 20
QuantLLM/Llama-3.2-3B-2bit-mlx
Text Generation
• 3B • Updated • 22
QuantLLM/Llama-3.2-3B-8bit-mlx
Text Generation
• 3B • Updated • 38
QuantLLM/Llama-3.2-3B-5bit-mlx
Text Generation
• 3B • Updated • 24
QuantLLM/functiongemma-270m-it-4bit-mlx
Text Generation
• 0.3B • Updated • 18