Inference Providers
Active filters: 2-bit
TensorFold/GLM-5.3-Flash-MLX-2bit-MTP
Image-Text-to-Text
• 352B • Updated • 1.27k
• 3
pugant/Qwen3.8-Flash-Next-ROCMFP4_STRIX_LEAN-GGUF
Text Generation
• 177B • Updated • 2.07k
• 6
Olt1z/DSV4-Flash-Vision-ablit-EXL3-Kalibrated
Image-Text-to-Text
• 64B • Updated • 92
• 1
VinceTrune/Qwen3.8-Flash-Next-Abliterated-EXL3-2.05bpw
Text Generation
• 18B • Updated • 56
• 3
satgeze/DeepSeek-v4.1-Flash-EXL3-2.0bpw-Ablit-EngramQ4-SM120-Dual-RTX-Pro-6000
189B • Updated • 347
• 3
prism-ml/Ternary-Bonsai-2-27B-gguf-dev
Text Generation
• 27B • Updated • 8.02k
• 33
r0b0tlab/Qwen3.8-Flash-Next-EXL3-2.50bpw
Image-Text-to-Text
• 22B • Updated • 386
• 10
decent-jawfish/bonsai-2-27b-mtp
Text Generation
• 27B • Updated • 4.83k
• 15
genevera/GLM-5.3-Flash-Uncensored-EXL3-2.5bpw
53B • Updated • 99
• 1
benthecarman/MiMo-V2.6-Flash-RL-exl3
Text Generation
• 46B • Updated • 1.49k
• 8
klee100/Qwen3.8-Flash-Next-Uncensored-AutoRound-3bpw-MTP
54B • Updated • 197
• 3
Cropduster69/GLM-5.3-Flash-oQ2e-mtp
Image-Text-to-Text
• 321B • Updated • 12
• 1
Ayushnangia/bloom3B-2bit-gptq
Text Generation
• Updated • 22
kaitchup/Llama-2-7b-gptq-2bit
Text Generation
• 7B • Updated • 23
• 1
nesty/sg-bart-large-4096-gptq-2bit
Arylwen/instruct-palmyra-20b-gptq-2
Text Generation
• Updated • 13
Text Generation
• 0.1B • Updated • 19
Sujan42024/dlite-v2-1_5b-2bitQuantization
Text Generation
• 2B • Updated • 19
kaitchup/Mistral-7B-v0.1-gptq-2bit
Text Generation
• 7B • Updated • 21
kaitchup/Llama-2-13b-hf-gptq-2bit
Text Generation
• 13B • Updated • 17
kaitchup/Llama-2-7b-hf-gptq-2bit
Text Generation
• 7B • Updated • 91
Text Generation
• 3B • Updated • 435
• 6
MaziyarPanahi/Tess-XS-v1-3-yarn-128K-Mistral-7B-Instruct-v0.1-GGUF
Text Generation
• 7B • Updated • 757
• 6
MaziyarPanahi/zephyr-7b-beta-Mistral-7B-Instruct-v0.1-GGUF
Text Generation
• 7B • Updated • 901
• 4
MaziyarPanahi/zephyr-7b-beta-Mistral-7B-Instruct-v0.2-GGUF
Text Generation
• 7B • Updated • 1.83k
• 5
MaziyarPanahi/NeuralPipe-7B-slerp-GGUF
Text Generation
• 7B • Updated • 332
• 3
MaziyarPanahi/NeuralPipe-7B-slerp-v0.2-GGUF
Text Generation
• 7B • Updated • 727
• 2
MaziyarPanahi/Mistral-7B-Instruct-v0.1-16k-Mistral-7B-Instruct-v0.2-slerp-GGUF
Text Generation
• 7B • Updated • 1.2k
• 2
MaziyarPanahi/rfp_instruct_model-Mistral-7B-Instruct-v0.2-slerp-GGUF
Text Generation
• 7B • Updated • 705
• 2
MaziyarPanahi/Mistral-7B-Instruct-SQL-Mistral-7B-Instruct-v0.2-slerp-GGUF
Text Generation
• 7B • Updated • 435
• 1