Inference Providers
Active filters: fp4
Text Generation
• 19B • Updated • 23
• 3
qingcheng-ai/Qwen3-32B-fp4
Text Generation
• 19B • Updated • 108
• 4
qingcheng-ai/Qwen3-8B-fp4
Text Generation
• 5B • Updated • 31
• 2
RedHatAI/Qwen3-30B-A3B-NVFP4
Text Generation
• 17B • Updated • 3.07k
• 3
RedHatAI/Llama-3.1-70B-Instruct-NVFP4
Text Generation
• 41B • Updated • 535
• 1
RedHatAI/Llama-3.1-70B-Instruct-NVFP4A16
Text Generation
• 41B • Updated • 33
Text Generation
• 19B • Updated • 408k
• 9
RedHatAI/Qwen3-32B-NVFP4A16
Text Generation
• 19B • Updated • 191
• 2
nvidia/Qwen3-235B-A22B-NVFP4
Text Generation
• 133B • Updated • 26.2k
• 23
nvidia/Qwen3-30B-A3B-NVFP4
Text Generation
• 16B • Updated • 21.3k
• 38
RedHatAI/Llama-4-Scout-17B-16E-Instruct-NVFP4
Text Generation
• 64B • Updated • 718
• 3
apolloparty/Qwen3-4B-NVFP4A16
2B • Updated • 38
Tonic/petite-elle-L-aime-3-sft
Text Generation
• 3B • Updated • 104
• 1
mradermacher/petite-elle-L-aime-3-sft-GGUF
Text Generation
• 3B • Updated • 623
• 1
nm-testing/DeepSeek-R1-Distill-Qwen-32B-NVFP4
Text Generation
• 19B • Updated • 226
• 3
Text Generation
• 2B • Updated • 23
2imi9/Qwen3-1.7B-NVFP4A16
Text Generation
• 1B • Updated • 34
• 1
ELVISIO/Qwen3-8B-NVFP4A16
Text Generation
• 5B • Updated • 62
RedHatAI/Llama-3.3-70B-Instruct-NVFP4
Text Generation
• 41B • Updated • 14.4k
• 2
AlekseyCalvin/QWEN_IMAGE_fp4_w_AbliteratedTE_Diffusers
Text-to-Image
• 11B • Updated • 75
• 9
imgailab/flux1-trtx-dev-fp4-blackwell
Updated • 16
• 1
imgailab/flux1-trtx-schnell-fp4-blackwell
Updated • 12
• 1
llmat/Mistral-7B-Instruct-v0.3-NVFP4
Text Generation
• 4B • Updated • 9.06k
llmat/Mistral-Small-Instruct-2409-NVFP4
Text Generation
• 13B • Updated • 23
2imi9/gpt-oss-20B-NVFP4A16-BF16
Text Generation
• 21B • Updated • 112
• 5
nvidia/Phi-4-multimodal-instruct-NVFP4
4B • Updated • 1.7k
• 15
nvidia/Phi-4-reasoning-plus-NVFP4
8B • Updated • 452
• 11
nvidia/Llama-3.1-8B-Instruct-NVFP4
5B • Updated • 47.1k
• 15
Text Generation
• 5B • Updated • 138k
• 24
Text Generation
• 8B • Updated • 422k
• 17