Inference Providers
Active filters: quark
amd/Qwen3.8-27B-Quark-AWQ-MXFP4
Image-Text-to-Text
• 16B • Updated • 16k
• 24
just1moremodel/Qwen3.8-27B-Uncensored-MXFP4-awq
Text Generation
• 16B • Updated • 450
• 8
amd/Qwen3.8-27B-Quark-AWQ-INT4-W4A16
Image-Text-to-Text
• 6B • Updated • 102k
• 18
amd/Qwen3.8-Flash-Next-Quark-MXFP4
Image-Text-to-Text
• 119B • Updated • 119
• 2
gearwave00001/Huihui-Qwen3.8-27B-Quark-AWQ-MXFP4-MtpFp8
16B • Updated • 86
• 2
magiccodingman/Qwen3.8-Flash-Next-Quark-MXFP4-fp8
Image-Text-to-Text
• 119B • Updated • 37
• 2
ukisai/Swift-Qwen3.8-27b-int4-AMD
Image-Text-to-Text
• 6B • Updated • 75
• 2
nekofish/Qwen3.8-27B-MXFP6
Image-Text-to-Text
• 22B • Updated • 2
magiccodingman/Qwen3.8-27B-MXFP4-MagicQuant-GGUF
Image-Text-to-Text
• 27B • Updated • 7.7k
• 5
OneNexus/GLM-5.3-Flash-MXFP4
Image-Text-to-Text
• 169B • Updated • 1.25k
• 3
amd/GLM-5.3-Quark-MXFP4-AttnFP8
Text Generation
• 384B • Updated • 10.5k
• 2
Zhangdanyang/Kimi-K3-DSpark-FP8-PTPC
Text Generation
• 4B • Updated • 2
just1moremodel/Qwen3.8-27B-TURBO-Fable-MXFP4-awq
Image-Text-to-Text
• 16B • Updated • 1.1k
• 3
fxmarty/llama-tiny-testing-quark-indev
1.03M • Updated • 9
fxmarty/llama-tiny-int4-per-group-sym
1.03M • Updated • 9
fxmarty/llama-tiny-w-fp8-a-fp8
1.03M • Updated • 9
fxmarty/llama-tiny-w-fp8-a-fp8-o-fp8
1.03M • Updated • 11
fxmarty/llama-tiny-w-int8-per-tensor
1.03M • Updated • 11
fxmarty/llama-small-int4-per-group-sym-awq
16.7M • Updated • 15
fxmarty/quark-legacy-int8
1.03M • Updated • 9
fxmarty/llama-tiny-w-int8-b-int8-per-tensor
1.03M • Updated • 9
fxmarty/llama-small-int4-per-group-sym-awq-old
16.7M • Updated • 8
amd-quark/llama-tiny-w-int8-per-tensor
1.03M • Updated • 19
amd-quark/llama-tiny-w-int8-b-int8-per-tensor
1.03M • Updated • 19
amd-quark/llama-tiny-w-fp8-a-fp8
1.03M • Updated • 18
amd-quark/llama-tiny-w-fp8-a-fp8-o-fp8
1.03M • Updated • 17
amd-quark/llama-tiny-int4-per-group-sym
1.03M • Updated • 15
amd-quark/llama-small-int4-per-group-sym-awq
16.7M • Updated • 22
amd-quark/quark-legacy-int8
1.03M • Updated • 7
amd/Llama-3.1-8B-Instruct-FP8-KV-Quark-test
8B • Updated • 2.67k