Decision models GGUF decision models for the /v1/systemone API in llama.cpp ggml-org/Clef-GGUF Image-Text-to-Text • 27B • Updated about 7 hours ago • 17.2k • 23 ggml-org/Clef-Flash-GGUF Image-Text-to-Text • 9B • Updated about 7 hours ago • 40.8k • 33 ggml-org/OpenJev-GGUF Image-Text-to-Text • 27B • Updated 10 days ago • 8.18k • 19 ggml-org/Laya-GGUF Text Classification • 0.4B • Updated 7 days ago • 13.5k • 30
NVIDIA Nemotron 3 Nano Collection for Nemotron-Nano-3-30B-A3B models ggml-org/NVIDIA-Nemotron-3-Nano-30B-A3B-GGUF Text Generation • 32B • Updated Aug 30 • 5.24k • 4
ggml-org/NVIDIA-Nemotron-3-Nano-30B-A3B-GGUF Text Generation • 32B • Updated Aug 30 • 5.24k • 4
NVIDIA Nemotron 3 Super Collection for Nemotron-3-Super-120B models ggml-org/Nemotron-3-Super-120B-GGUF 121B • Updated Mar 16 • 1.99k • 13
Gemma 4 ggml-org/gemma-4-E2B-it-GGUF Any-to-Any • 5B • Updated Jul 26 • 59.7k • 90 ggml-org/gemma-4-E4B-it-GGUF Any-to-Any • 8B • Updated Jul 26 • 478k • 95 ggml-org/gemma-4-12B-it-GGUF Any-to-Any • 12B • Updated Aug 23 • 52.5k • 93 ggml-org/gemma-4-26B-A4B-it-GGUF Image-Text-to-Text • 25B • Updated Jul 26 • 324k • 92
Gemma 3n ggml-org/gemma-3n-E2B-it-GGUF 4B • Updated Aug 22, 2025 • 3.06k • 25 ggml-org/gemma-3n-E4B-it-GGUF 7B • Updated Jun 26, 2025 • 1.91k • 20
Gemma 1.1 GGUFs ggml-org/gemma-1.1-2b-it-Q8_0-GGUF 3B • Updated Apr 5, 2024 • 101 • 1 ggml-org/gemma-1.1-7b-it-Q8_0-GGUF 9B • Updated Apr 5, 2024 • 181 ggml-org/gemma-1.1-7b-it-Q4_K_M-GGUF 9B • Updated Apr 5, 2024 • 342 • 4
Devstral 2 Collection for Devstral-Small-2-24B-Instruct-2512 models ggml-org/Devstral-Small-2-24B-Instruct-2512-GGUF 24B • Updated Dec 18, 2025 • 1.92k • 7 ggml-org/Devstral-2-123B-Instruct-2512-GGUF 125B • Updated Dec 19, 2025 • 13.1k • 3
GLM-V ggml-org/GLM-4.6V-Flash-GGUF 9B • Updated Jan 15 • 13.3k • 24 ggml-org/GLM-4.6V-GGUF 107B • Updated Jan 15 • 8.23k • 7 ggml-org/AutoGLM-Phone-9B-GGUF 9B • Updated Dec 17, 2025 • 908 • 3
GPT OSS ggml-org/gpt-oss-120b-GGUF Text Generation • 117B • Updated Jul 28 • 146k • 79 ggml-org/gpt-oss-20b-GGUF Text Generation • 21B • Updated Jul 28 • 163k • 182
InternVL 3 and InternVL 2.5 ggml-org/InternVL3-1B-Instruct-GGUF 0.6B • Updated May 10, 2025 • 1.88k • 5 ggml-org/InternVL3-2B-Instruct-GGUF 2B • Updated May 10, 2025 • 2.15k • 5 ggml-org/InternVL3-8B-Instruct-GGUF 8B • Updated May 10, 2025 • 1.76k • 6 ggml-org/InternVL3-14B-Instruct-GGUF 15B • Updated May 10, 2025 • 881 • 4
Qwen 3 ggml-org/Qwen3-0.6B-GGUF Text Generation • 0.8B • Updated Jul 16 • 40.4k • 18 ggml-org/Qwen3-1.7B-GGUF 2B • Updated Apr 28, 2025 • 86.9k • 14 ggml-org/Qwen3-4B-GGUF 4B • Updated Apr 28, 2025 • 12.8k • 8 ggml-org/Qwen3-8B-GGUF Text Generation • 8B • Updated Jul 28 • 2.7k • 8
GGUF LoRA adapters Adapters extracted from fine tuned models, using mergekit-extract-lora ggml-org/LoRA-Llama-3-Instruct-abliteration-8B-F16-GGUF 88.1M • Updated Nov 1, 2024 • 74 • 1 ggml-org/LoRA-Qwen2.5-1.5B-Instruct-abliterated-F16-GGUF 93.6M • Updated Jan 23, 2025 • 130 • 5 ggml-org/LoRA-Qwen2.5-3B-Instruct-abliterated-F16-GGUF 0.1B • Updated Jan 9, 2025 • 90 • 1 ggml-org/LoRA-Qwen2.5-7B-Instruct-abliterated-v3-F16-GGUF 90.9M • Updated Jan 8, 2025 • 55 • 4
ggml-org/LoRA-Qwen2.5-1.5B-Instruct-abliterated-F16-GGUF 93.6M • Updated Jan 23, 2025 • 130 • 5
llama.cpp presets Models that are used for presets in llama.cpp. ggml-org/gte-small-Q8_0-GGUF Sentence Similarity • 33.2M • Updated Feb 6, 2025 • 129 • 3 ggml-org/bge-small-en-v1.5-Q8_0-GGUF Feature Extraction • 33.2M • Updated Feb 6, 2025 • 1.81k • 6 ggml-org/e5-small-v2-Q8_0-GGUF Sentence Similarity • 33.2M • Updated Feb 6, 2025 • 140 • 1
ggml-org/bge-small-en-v1.5-Q8_0-GGUF Feature Extraction • 33.2M • Updated Feb 6, 2025 • 1.81k • 6
NVIDIA Nemotron 3.5 Lightning ggml-org/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-GGUF Text Generation • 32B • Updated 28 days ago • 177k • 35
ggml-org/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-GGUF Text Generation • 32B • Updated 28 days ago • 177k • 35
NVIDIA-Nemotron-3-Nano-Omni ggml-org/NVIDIA-Nemotron-3-Nano-Omni-30B-A3B-GGUF Any-to-Any • 32B • Updated Aug 30 • 2.01k • 2
ggml-org/NVIDIA-Nemotron-3-Nano-Omni-30B-A3B-GGUF Any-to-Any • 32B • Updated Aug 30 • 2.01k • 2
OCR models ggml-org/GLM-OCR-GGUF 0.9B • Updated Mar 10 • 32k • 111 ggml-org/DeepSeek-OCR-GGUF 3B • Updated Mar 25 • 5.15k • 12 ggml-org/dots.ocr-GGUF 2B • Updated Apr 5 • 3.79k • 5 ggml-org/Qianfan-OCR-GGUF 4B • Updated Apr 10 • 729 • 6
Gemma 3 ggml-org/gemma-3-270m-it-GGUF 0.3B • Updated Aug 15, 2025 • 3.95k • 22 ggml-org/gemma-3-1b-it-GGUF 1.0B • Updated Mar 12, 2025 • 207k • 52 ggml-org/gemma-3-4b-it-GGUF Image-Text-to-Text • 4B • Updated May 21, 2025 • 60.2k • 62 ggml-org/gemma-3-12b-it-GGUF Image-Text-to-Text • 12B • Updated May 21, 2025 • 10.4k • 32
Gemma 3-270m Collection of models for Gemma 3-270m ggml-org/gemma-3-270m-GGUF 0.3B • Updated Aug 14, 2025 • 1.24k • 22 ggml-org/gemma-3-270m-it-GGUF 0.3B • Updated Aug 15, 2025 • 3.95k • 22 ggml-org/gemma-3-270m-qat-GGUF 0.3B • Updated Aug 14, 2025 • 11.4k • 9 ggml-org/gemma-3-270m-it-qat-GGUF 0.3B • Updated Aug 15, 2025 • 3.63k • 15
EmbeddingGemma 300M ggml-org/embeddinggemma-300M-GGUF 0.3B • Updated Apr 29 • 335k • 52 ggml-org/embeddinggemma-300M-qat-q4_0-GGUF Feature Extraction • 0.3B • Updated Sep 15, 2025 • 3.62k • 6 ggml-org/embeddinggemma-300m-qat-q8_0-GGUF Feature Extraction • 0.3B • Updated Sep 15, 2025 • 34.1k • 21
ggml-org/embeddinggemma-300M-qat-q4_0-GGUF Feature Extraction • 0.3B • Updated Sep 15, 2025 • 3.62k • 6
ggml-org/embeddinggemma-300m-qat-q8_0-GGUF Feature Extraction • 0.3B • Updated Sep 15, 2025 • 34.1k • 21
Multimodal GGUFs Vision and audio models compatible with llama-server and llama-mtmd-cli GLM-V Collection 3 items • Updated Aug 25 • 14 Ministral 3 Collection 6 items • Updated Jul 30 • 4 Gemma 3 Collection 10 items • Updated Jul 30 • 24 Kimi-VL Collection 1 item • Updated Jul 30 • 2
Ministral 3 ggml-org/Ministral-3-3B-Reasoning-2512-GGUF Image-Text-to-Text • 3B • Updated Dec 2, 2025 • 905 • 2 ggml-org/Ministral-3-8B-Reasoning-2512-GGUF Image-Text-to-Text • 8B • Updated Dec 2, 2025 • 1.26k • 2 ggml-org/Ministral-3-14B-Reasoning-2512-GGUF Image-Text-to-Text • 14B • Updated Dec 2, 2025 • 1.09k • 3 ggml-org/Ministral-3-3B-Instruct-2512-GGUF Image-Text-to-Text • 3B • Updated Dec 2, 2025 • 1.08k • 4
ggml-org/Ministral-3-3B-Reasoning-2512-GGUF Image-Text-to-Text • 3B • Updated Dec 2, 2025 • 905 • 2
ggml-org/Ministral-3-8B-Reasoning-2512-GGUF Image-Text-to-Text • 8B • Updated Dec 2, 2025 • 1.26k • 2
ggml-org/Ministral-3-14B-Reasoning-2512-GGUF Image-Text-to-Text • 14B • Updated Dec 2, 2025 • 1.09k • 3
ggml-org/Ministral-3-3B-Instruct-2512-GGUF Image-Text-to-Text • 3B • Updated Dec 2, 2025 • 1.08k • 4
Qwen 2 VL and Qwen 2.5 VL ggml-org/Qwen2.5-VL-3B-Instruct-GGUF 3B • Updated Apr 30, 2025 • 79.1k • 17 ggml-org/Qwen2.5-VL-7B-Instruct-GGUF 8B • Updated Apr 30, 2025 • 54.4k • 18 ggml-org/Qwen2.5-VL-32B-Instruct-GGUF 33B • Updated May 15, 2025 • 18.8k • 5 ggml-org/Qwen2-VL-2B-Instruct-GGUF 2B • Updated Apr 30, 2025 • 11.9k • 4
SmolVLM GGUF ggml-org/SmolVLM2-2.2B-Instruct-GGUF 2B • Updated Apr 30, 2025 • 37.1k • 42 ggml-org/SmolVLM2-500M-Video-Instruct-GGUF 0.4B • Updated Apr 30, 2025 • 14k • 20 ggml-org/SmolVLM2-256M-Video-Instruct-GGUF Image-Text-to-Text • 0.2B • Updated Aug 23 • 5.6k • 13 ggml-org/SmolVLM-Instruct-GGUF 2B • Updated Apr 30, 2025 • 3.34k • 9
ggml-org/SmolVLM2-256M-Video-Instruct-GGUF Image-Text-to-Text • 0.2B • Updated Aug 23 • 5.6k • 13
llama.vim Recommended models for the llama.vim and llama.vscode plugins ggml-org/Qwen2.5-Coder-0.5B-Q8_0-GGUF Text Generation • 0.5B • Updated Jan 31, 2025 • 797 • 11 ggml-org/Qwen2.5-Coder-1.5B-Q8_0-GGUF Text Generation • 2B • Updated Oct 28, 2024 • 5.76k • 23 ggml-org/Qwen2.5-Coder-3B-Q8_0-GGUF Text Generation • 3B • Updated Nov 26, 2024 • 2.03k • 17 ggml-org/Qwen2.5-Coder-7B-Q8_0-GGUF Text Generation • 8B • Updated Oct 28, 2024 • 2.41k • 13
ggml-org/Qwen2.5-Coder-0.5B-Q8_0-GGUF Text Generation • 0.5B • Updated Jan 31, 2025 • 797 • 11
ggml-org/Qwen2.5-Coder-1.5B-Q8_0-GGUF Text Generation • 2B • Updated Oct 28, 2024 • 5.76k • 23
VAD Voice Activity Detection (VAD) models for whisper.cpp. ggml-org/whisper-vad Updated Nov 17, 2025 • 26
Decision models GGUF decision models for the /v1/systemone API in llama.cpp ggml-org/Clef-GGUF Image-Text-to-Text • 27B • Updated about 7 hours ago • 17.2k • 23 ggml-org/Clef-Flash-GGUF Image-Text-to-Text • 9B • Updated about 7 hours ago • 40.8k • 33 ggml-org/OpenJev-GGUF Image-Text-to-Text • 27B • Updated 10 days ago • 8.18k • 19 ggml-org/Laya-GGUF Text Classification • 0.4B • Updated 7 days ago • 13.5k • 30
NVIDIA Nemotron 3.5 Lightning ggml-org/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-GGUF Text Generation • 32B • Updated 28 days ago • 177k • 35
ggml-org/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-GGUF Text Generation • 32B • Updated 28 days ago • 177k • 35
NVIDIA Nemotron 3 Nano Collection for Nemotron-Nano-3-30B-A3B models ggml-org/NVIDIA-Nemotron-3-Nano-30B-A3B-GGUF Text Generation • 32B • Updated Aug 30 • 5.24k • 4
ggml-org/NVIDIA-Nemotron-3-Nano-30B-A3B-GGUF Text Generation • 32B • Updated Aug 30 • 5.24k • 4
NVIDIA-Nemotron-3-Nano-Omni ggml-org/NVIDIA-Nemotron-3-Nano-Omni-30B-A3B-GGUF Any-to-Any • 32B • Updated Aug 30 • 2.01k • 2
ggml-org/NVIDIA-Nemotron-3-Nano-Omni-30B-A3B-GGUF Any-to-Any • 32B • Updated Aug 30 • 2.01k • 2
NVIDIA Nemotron 3 Super Collection for Nemotron-3-Super-120B models ggml-org/Nemotron-3-Super-120B-GGUF 121B • Updated Mar 16 • 1.99k • 13
OCR models ggml-org/GLM-OCR-GGUF 0.9B • Updated Mar 10 • 32k • 111 ggml-org/DeepSeek-OCR-GGUF 3B • Updated Mar 25 • 5.15k • 12 ggml-org/dots.ocr-GGUF 2B • Updated Apr 5 • 3.79k • 5 ggml-org/Qianfan-OCR-GGUF 4B • Updated Apr 10 • 729 • 6
Gemma 4 ggml-org/gemma-4-E2B-it-GGUF Any-to-Any • 5B • Updated Jul 26 • 59.7k • 90 ggml-org/gemma-4-E4B-it-GGUF Any-to-Any • 8B • Updated Jul 26 • 478k • 95 ggml-org/gemma-4-12B-it-GGUF Any-to-Any • 12B • Updated Aug 23 • 52.5k • 93 ggml-org/gemma-4-26B-A4B-it-GGUF Image-Text-to-Text • 25B • Updated Jul 26 • 324k • 92
Gemma 3 ggml-org/gemma-3-270m-it-GGUF 0.3B • Updated Aug 15, 2025 • 3.95k • 22 ggml-org/gemma-3-1b-it-GGUF 1.0B • Updated Mar 12, 2025 • 207k • 52 ggml-org/gemma-3-4b-it-GGUF Image-Text-to-Text • 4B • Updated May 21, 2025 • 60.2k • 62 ggml-org/gemma-3-12b-it-GGUF Image-Text-to-Text • 12B • Updated May 21, 2025 • 10.4k • 32
Gemma 3n ggml-org/gemma-3n-E2B-it-GGUF 4B • Updated Aug 22, 2025 • 3.06k • 25 ggml-org/gemma-3n-E4B-it-GGUF 7B • Updated Jun 26, 2025 • 1.91k • 20
Gemma 3-270m Collection of models for Gemma 3-270m ggml-org/gemma-3-270m-GGUF 0.3B • Updated Aug 14, 2025 • 1.24k • 22 ggml-org/gemma-3-270m-it-GGUF 0.3B • Updated Aug 15, 2025 • 3.95k • 22 ggml-org/gemma-3-270m-qat-GGUF 0.3B • Updated Aug 14, 2025 • 11.4k • 9 ggml-org/gemma-3-270m-it-qat-GGUF 0.3B • Updated Aug 15, 2025 • 3.63k • 15
Gemma 1.1 GGUFs ggml-org/gemma-1.1-2b-it-Q8_0-GGUF 3B • Updated Apr 5, 2024 • 101 • 1 ggml-org/gemma-1.1-7b-it-Q8_0-GGUF 9B • Updated Apr 5, 2024 • 181 ggml-org/gemma-1.1-7b-it-Q4_K_M-GGUF 9B • Updated Apr 5, 2024 • 342 • 4
EmbeddingGemma 300M ggml-org/embeddinggemma-300M-GGUF 0.3B • Updated Apr 29 • 335k • 52 ggml-org/embeddinggemma-300M-qat-q4_0-GGUF Feature Extraction • 0.3B • Updated Sep 15, 2025 • 3.62k • 6 ggml-org/embeddinggemma-300m-qat-q8_0-GGUF Feature Extraction • 0.3B • Updated Sep 15, 2025 • 34.1k • 21
ggml-org/embeddinggemma-300M-qat-q4_0-GGUF Feature Extraction • 0.3B • Updated Sep 15, 2025 • 3.62k • 6
ggml-org/embeddinggemma-300m-qat-q8_0-GGUF Feature Extraction • 0.3B • Updated Sep 15, 2025 • 34.1k • 21
Devstral 2 Collection for Devstral-Small-2-24B-Instruct-2512 models ggml-org/Devstral-Small-2-24B-Instruct-2512-GGUF 24B • Updated Dec 18, 2025 • 1.92k • 7 ggml-org/Devstral-2-123B-Instruct-2512-GGUF 125B • Updated Dec 19, 2025 • 13.1k • 3
Multimodal GGUFs Vision and audio models compatible with llama-server and llama-mtmd-cli GLM-V Collection 3 items • Updated Aug 25 • 14 Ministral 3 Collection 6 items • Updated Jul 30 • 4 Gemma 3 Collection 10 items • Updated Jul 30 • 24 Kimi-VL Collection 1 item • Updated Jul 30 • 2
GLM-V ggml-org/GLM-4.6V-Flash-GGUF 9B • Updated Jan 15 • 13.3k • 24 ggml-org/GLM-4.6V-GGUF 107B • Updated Jan 15 • 8.23k • 7 ggml-org/AutoGLM-Phone-9B-GGUF 9B • Updated Dec 17, 2025 • 908 • 3
Ministral 3 ggml-org/Ministral-3-3B-Reasoning-2512-GGUF Image-Text-to-Text • 3B • Updated Dec 2, 2025 • 905 • 2 ggml-org/Ministral-3-8B-Reasoning-2512-GGUF Image-Text-to-Text • 8B • Updated Dec 2, 2025 • 1.26k • 2 ggml-org/Ministral-3-14B-Reasoning-2512-GGUF Image-Text-to-Text • 14B • Updated Dec 2, 2025 • 1.09k • 3 ggml-org/Ministral-3-3B-Instruct-2512-GGUF Image-Text-to-Text • 3B • Updated Dec 2, 2025 • 1.08k • 4
ggml-org/Ministral-3-3B-Reasoning-2512-GGUF Image-Text-to-Text • 3B • Updated Dec 2, 2025 • 905 • 2
ggml-org/Ministral-3-8B-Reasoning-2512-GGUF Image-Text-to-Text • 8B • Updated Dec 2, 2025 • 1.26k • 2
ggml-org/Ministral-3-14B-Reasoning-2512-GGUF Image-Text-to-Text • 14B • Updated Dec 2, 2025 • 1.09k • 3
ggml-org/Ministral-3-3B-Instruct-2512-GGUF Image-Text-to-Text • 3B • Updated Dec 2, 2025 • 1.08k • 4
GPT OSS ggml-org/gpt-oss-120b-GGUF Text Generation • 117B • Updated Jul 28 • 146k • 79 ggml-org/gpt-oss-20b-GGUF Text Generation • 21B • Updated Jul 28 • 163k • 182
InternVL 3 and InternVL 2.5 ggml-org/InternVL3-1B-Instruct-GGUF 0.6B • Updated May 10, 2025 • 1.88k • 5 ggml-org/InternVL3-2B-Instruct-GGUF 2B • Updated May 10, 2025 • 2.15k • 5 ggml-org/InternVL3-8B-Instruct-GGUF 8B • Updated May 10, 2025 • 1.76k • 6 ggml-org/InternVL3-14B-Instruct-GGUF 15B • Updated May 10, 2025 • 881 • 4
Qwen 2 VL and Qwen 2.5 VL ggml-org/Qwen2.5-VL-3B-Instruct-GGUF 3B • Updated Apr 30, 2025 • 79.1k • 17 ggml-org/Qwen2.5-VL-7B-Instruct-GGUF 8B • Updated Apr 30, 2025 • 54.4k • 18 ggml-org/Qwen2.5-VL-32B-Instruct-GGUF 33B • Updated May 15, 2025 • 18.8k • 5 ggml-org/Qwen2-VL-2B-Instruct-GGUF 2B • Updated Apr 30, 2025 • 11.9k • 4
Qwen 3 ggml-org/Qwen3-0.6B-GGUF Text Generation • 0.8B • Updated Jul 16 • 40.4k • 18 ggml-org/Qwen3-1.7B-GGUF 2B • Updated Apr 28, 2025 • 86.9k • 14 ggml-org/Qwen3-4B-GGUF 4B • Updated Apr 28, 2025 • 12.8k • 8 ggml-org/Qwen3-8B-GGUF Text Generation • 8B • Updated Jul 28 • 2.7k • 8
SmolVLM GGUF ggml-org/SmolVLM2-2.2B-Instruct-GGUF 2B • Updated Apr 30, 2025 • 37.1k • 42 ggml-org/SmolVLM2-500M-Video-Instruct-GGUF 0.4B • Updated Apr 30, 2025 • 14k • 20 ggml-org/SmolVLM2-256M-Video-Instruct-GGUF Image-Text-to-Text • 0.2B • Updated Aug 23 • 5.6k • 13 ggml-org/SmolVLM-Instruct-GGUF 2B • Updated Apr 30, 2025 • 3.34k • 9
ggml-org/SmolVLM2-256M-Video-Instruct-GGUF Image-Text-to-Text • 0.2B • Updated Aug 23 • 5.6k • 13
GGUF LoRA adapters Adapters extracted from fine tuned models, using mergekit-extract-lora ggml-org/LoRA-Llama-3-Instruct-abliteration-8B-F16-GGUF 88.1M • Updated Nov 1, 2024 • 74 • 1 ggml-org/LoRA-Qwen2.5-1.5B-Instruct-abliterated-F16-GGUF 93.6M • Updated Jan 23, 2025 • 130 • 5 ggml-org/LoRA-Qwen2.5-3B-Instruct-abliterated-F16-GGUF 0.1B • Updated Jan 9, 2025 • 90 • 1 ggml-org/LoRA-Qwen2.5-7B-Instruct-abliterated-v3-F16-GGUF 90.9M • Updated Jan 8, 2025 • 55 • 4
ggml-org/LoRA-Qwen2.5-1.5B-Instruct-abliterated-F16-GGUF 93.6M • Updated Jan 23, 2025 • 130 • 5
llama.vim Recommended models for the llama.vim and llama.vscode plugins ggml-org/Qwen2.5-Coder-0.5B-Q8_0-GGUF Text Generation • 0.5B • Updated Jan 31, 2025 • 797 • 11 ggml-org/Qwen2.5-Coder-1.5B-Q8_0-GGUF Text Generation • 2B • Updated Oct 28, 2024 • 5.76k • 23 ggml-org/Qwen2.5-Coder-3B-Q8_0-GGUF Text Generation • 3B • Updated Nov 26, 2024 • 2.03k • 17 ggml-org/Qwen2.5-Coder-7B-Q8_0-GGUF Text Generation • 8B • Updated Oct 28, 2024 • 2.41k • 13
ggml-org/Qwen2.5-Coder-0.5B-Q8_0-GGUF Text Generation • 0.5B • Updated Jan 31, 2025 • 797 • 11
ggml-org/Qwen2.5-Coder-1.5B-Q8_0-GGUF Text Generation • 2B • Updated Oct 28, 2024 • 5.76k • 23
llama.cpp presets Models that are used for presets in llama.cpp. ggml-org/gte-small-Q8_0-GGUF Sentence Similarity • 33.2M • Updated Feb 6, 2025 • 129 • 3 ggml-org/bge-small-en-v1.5-Q8_0-GGUF Feature Extraction • 33.2M • Updated Feb 6, 2025 • 1.81k • 6 ggml-org/e5-small-v2-Q8_0-GGUF Sentence Similarity • 33.2M • Updated Feb 6, 2025 • 140 • 1
ggml-org/bge-small-en-v1.5-Q8_0-GGUF Feature Extraction • 33.2M • Updated Feb 6, 2025 • 1.81k • 6
VAD Voice Activity Detection (VAD) models for whisper.cpp. ggml-org/whisper-vad Updated Nov 17, 2025 • 26