Inference Providers
Active filters: Q8
AIconjured/embeddinggemma-300M-NVFP4-Q8-GGUF
0.3B • Updated • 1
cibernicola/FLOR-6.3B-xat-Q8_0
Text Generation
• 6B • Updated • 74
cibernicola/FLOR-1.3B-xat-Q8
Text Generation
• 1B • Updated • 35
cibernicola/FLOR-6.3B-xat-Q5_K
Text Generation
• 6B • Updated • 42
prithivMLmods/Qwen2.5-Coder-7B-Instruct-GGUF
Text Generation
• 8B • Updated • 248
• 2
prithivMLmods/Qwen2.5-Coder-7B-GGUF
Text Generation
• 8B • Updated • 828
• 3
prithivMLmods/Qwen2.5-Coder-3B-GGUF
Text Generation
• 3B • Updated • 127
• 4
prithivMLmods/Qwen2.5-Coder-1.5B-GGUF
Text Generation
• 2B • Updated • 618
• 5
prithivMLmods/Qwen2.5-Coder-1.5B-Instruct-GGUF
Text Generation
• 2B • Updated • 181
• 3
prithivMLmods/Qwen2.5-Coder-3B-Instruct-GGUF
Text Generation
• 3B • Updated • 134
• 5
prithivMLmods/Llama-3.2-3B-GGUF
Text Generation
• 3B • Updated • 135
• 2
harisnaeem/Phi-4-mini-instruct-GGUF-Q8
Text Generation
• 4B • Updated • 10
ykarout/llama3-deepseek_Q8
Text Generation
• 8B • Updated • 1
michelkao/Ollama-3.2-GGUF
Text Generation
• 3B • Updated • 307
SiddhJagani/gpt-oss-20b-no-think-mlx-Q8
Text Generation
• 21B • Updated • 47
• 1
0.1B • Updated • 9
Terminator278/Qwen2.5-Coder-3B-GGUF
Text Generation
• 3B • Updated • 78