Inference Providers
Active filters: arm64
AEON-7/Qwen3.6-35B-A3B-heretic-NVFP4
Image-Text-to-Text
• 21B • Updated • 436k
• 65
Hellohal2064/vllm-dgx-spark-gb10
Text Generation
• Updated • 9
cudabenchmarktest/personaplex-7b-turbo2bit
vlad-m-dev/distiluse-base-multilingual-v2-merged-onnx
Feature Extraction
• Updated • 1
onnx-community/distiluse-base-multilingual-v2-merged-onnx
Feature Extraction
• Updated • 1
halley-ai/gpt-oss-20b-MLX-4bit-gs32
Text Generation
• 21B • Updated • 149
• 3
halley-ai/gpt-oss-20b-MLX-6bit-gs32
Text Generation
• 21B • Updated • 37
• 1
halley-ai/gpt-oss-20b-MLX-5bit-gs32
Text Generation
• 21B • Updated • 61
• 1
halley-ai/gpt-oss-120b-MLX-8bit-gs32
Text Generation
• 117B • Updated • 55
• 1
halley-ai/gpt-oss-120b-MLX-bf16
Text Generation
• 117B • Updated • 167
• 3
halley-ai/gpt-oss-120b-MLX-6bit-gs64
Text Generation
• 117B • Updated • 37
• 1
halley-ai/Qwen3-Next-80B-A3B-Instruct-MLX-4bit-gs64
Text Generation
• 80B • Updated • 35
• 1
halley-ai/Qwen3-Next-80B-A3B-Instruct-MLX-5bit-gs32
Text Generation
• 80B • Updated • 24
• 1
halley-ai/Qwen3-Next-80B-A3B-Instruct-MLX-6bit-gs64
Text Generation
• 80B • Updated • 31
• 1
mjbommar/glaurung-binary-tokenizer-001
Feature Extraction
• Updated mjbommar/glaurung-binary-tokenizer-002
Feature Extraction
• Updated • 1
thehighnotes/vllm-jetson-orin
Text Generation
• Updated Text Generation
• Updated • 1
AEON-7/Gemma-4-26B-A4B-it-Uncensored-NVFP4
Text Generation
• 15B • Updated • 10.4k
• 22
AEON-7/Gemma-4-31B-it-DECKARD-HERETIC-Uncensored-NVFP4
Text Generation
• 18B • Updated • 7.3k
• 13
AEON-7/Gemma-4-31B-it-DECKARD-HERETIC-Uncensored-NVFP4-SVDQuant
Text Generation
• 19B • Updated • 198
• 2
AEON-7/DFlash-Qwen3.5-27B-Uncensored-NVFP4
Text Generation
• 17B • Updated • 148
• 2
AEON-7/supergemma4-26b-abliterated-multimodal-nvfp4
Text Generation
• 15B • Updated • 211
• 6
AEON-7/Gemma-4-E4B-DECKARD-HERETIC-NVFP4
Text Generation
• 6B • Updated • 871
• 1
AEON-7/Gemma-4-E4B-it-Uncensored-NVFP4
Text Generation
• 6B • Updated • 304
• 3
AEON-7/Gemma-4-E4B-DECKARD-HERETIC-Uncensored-NVFP4
Text Generation
• 6B • Updated • 147
• 1
AEON-7/gemma-4-31B-it-speculator.eagle3-NVFP4
Text Generation
• 2B • Updated • 384
• 5
AEON-7/Qwen3.6-27B-AEON-Ultimate-Uncensored-BF16
Text Generation
• 27B • Updated • 8.02k
• • 147
blckrvrfx/edge-multimodal-embeddings
Feature Extraction
• Updated AEON-7/Nemotron-3-Nano-Omni-AEON-Ultimate-Uncensored-BF16
Any-to-Any
• 33B • Updated • 172
• 6