Inference Providers
Active filters: gpu
AEON-7/Qwen3.6-35B-A3B-heretic-NVFP4
Image-Text-to-Text
• 21B • Updated • 354k
• 62
Hellohal2064/vllm-dgx-spark-gb10
Text Generation
• Updated • 9
Jay0515/onnxruntime-gpu-aarch64-cuda13-sm121
Other
• Updated • 7
AEON-7/Qwen3.6-27B-AEON-Ultimate-Uncensored-BF16
Text Generation
• 27B • Updated • 9.79k
• • 146
clxudfast/huihui4-8b-a4b-v2-Q4_K_M
8B • Updated • 5
• 1
HarmenWessels/gemma-4-12B-it-qat-int4-ov
Image-Text-to-Text
• Updated • 209
• 3
litert-community/yolox-nano-litert
Object Detection
• Updated • 1.17k
• 1
litert-community/real-esrgan-x4v3-litert
Image-to-Image
• Updated • 123
• 1
litert-community/Qwen3-Embedding-0.6B-LiteRT
Sentence Similarity
• Updated • 40
• 1
Text Generation
• Updated • 1
Vishal74/Seq2SeqModel_LSTM
Updated
Tech-Meld/gpus-everywhere
Text-to-Image
• Updated • 9
• • 1
vhab10/llama_3.1_8b_Q4_K_M-gguf
Text Generation
• 8B • Updated • 1.69k
Text Generation
• 4B • Updated • 16
mradermacher/Loxa-4B-GGUF
4B • Updated • 881
mradermacher/Loxa-4B-i1-GGUF
4B • Updated • 103
Text Generation
• 4B • Updated • 9
mradermacher/CodeLoxa-4B-GGUF
4B • Updated • 874
• 1
mradermacher/CodeLoxa-4B-i1-GGUF
4B • Updated • 193
Text Generation
• 2B • Updated • 15
• 1
mradermacher/Loxa-1.6B-GGUF
2B • Updated • 69
mradermacher/Loxa-1.6B-i1-GGUF
2B • Updated • 162
frameai/Loxa-1.6B-uncensored
Text Generation
• 2B • Updated • 14
• 1
mradermacher/Loxa-1.6B-uncensored-GGUF
2B • Updated • 74
• 2
mradermacher/Loxa-1.6B-uncensored-i1-GGUF
2B • Updated • 133
ConfidentialMind/gte-multilingual-reranker-base-onnx-op14-opt-gpu-int8
Sentence Similarity
• Updated • 4
• 1
ConfidentialMind/gte-multilingual-reranker-base-onnx-op14-opt-gpu
Sentence Similarity
• Updated • 9
ConfidentialMind/gte-multilingual-reranker-base-onnx-op19-opt-gpu
Sentence Similarity
• Updated • 5
Robotics
• Updated