Inference Providers
Active filters: metal
prism-ml/Ternary-Bonsai-27B-gguf
Text Generation
• 4B • Updated • 698k
• • 1.12k
Text Generation
• 4B • Updated • 2.46M
• 704
prism-ml/Bonsai-27B-mlx-1bit
Text Generation
• 2B • Updated • 43.5k
• 193
badtheorylabs/BTL-3-Compact
Text Generation
• 8B • Updated • 2.3k
• 36
prism-ml/Ternary-Bonsai-27B-mlx-2bit
Text Generation
• 3B • Updated • 34.6k
• 158
Text Generation
• 20B • Updated • 683k
• 349
huihui-ai/Huihui-DeepSeek-V4-Flash-abliterated-ds4-GGUF
284B • Updated • 513k
• 113
Text Generation
• 8B • Updated • 62.8k
• 761
prism-ml/Bonsai-1.7B-gguf
Text Generation
• 2B • Updated • 38.1k
• 83
andreaborio/DeepSeek-V4-Flash-Hebrus-GGUF
Text Generation
• 85B • Updated • 623
• 2
Text Generation
• 4B • Updated • 15.8k
• 53
mlboydaisuke/Ornith-1.0-9B-CoreAI
Text Generation
• Updated • 495
• 2
andreaborio/Qwen3.6-35B-A3B-Hebrus-GGUF
Text Generation
• Updated • 1.85k
• 1
andreaborio/GLM-5.2-Hebrus-GGUF
Text Generation
• 260B • Updated • 380
• 1
allemanfredi/inferno-glm-5.2-q2-expertpack
pyrodog/CyberNeurova-DeepSeek-V4-Flash-abliterated-v2-IQ3_XXS-AS-GGUF
Text Generation
• 284B • Updated • 272
• 1
jacklarmer/laguna-s-2.1-mlx-optimized
Text Generation
• Updated • 1
SwinliQ-AIs/Bonsai-27B-gguf
Text Generation
• 4B • Updated • 1
Text-to-Image
• Updated • 301
• • 15
RalFinger/chrome-style-sdxl-lora
Text-to-Image
• Updated • 20
• • 1
e-n-v-y/envy-metallic-xl-01
Text-to-Image
• Updated • 3
• • 2
UnionStreet/vision-1-mini
Text Classification
• 8B • Updated • 22
• 1
halley-ai/gpt-oss-20b-MLX-4bit-gs32
Text Generation
• 21B • Updated • 130
• 3
halley-ai/gpt-oss-20b-MLX-6bit-gs32
Text Generation
• 21B • Updated • 45
• 1
halley-ai/gpt-oss-20b-MLX-5bit-gs32
Text Generation
• 21B • Updated • 46
• 1
halley-ai/gpt-oss-120b-MLX-8bit-gs32
Text Generation
• 117B • Updated • 56
• 1
halley-ai/gpt-oss-120b-MLX-bf16
Text Generation
• 117B • Updated • 179
• 3
halley-ai/gpt-oss-120b-MLX-6bit-gs64
Text Generation
• 117B • Updated • 51
• 1
halley-ai/Qwen3-Next-80B-A3B-Instruct-MLX-4bit-gs64
Text Generation
• 80B • Updated • 24
• 1
halley-ai/Qwen3-Next-80B-A3B-Instruct-MLX-5bit-gs32
Text Generation
• 80B • Updated • 18
• 1