Edit Models filters
Model Tree
Apps
Inference Providers
One-click Deployment
Models
419
Active filters: cuda
justinchuby/Muse-Glimmer-30B-ONNX-INT4-CUDA
Image-Text-to-Text • Updated
RinKana/Amadeus-Hinton-Karl-1.3M-Encoder-Decoder-Model
Updated
XENONNONE/Inflect-Micro-v2-ONNX
Text-to-Speech • Updated • 25
webbrain-one/Ling-3.0-tiny-ONNX
Text Generation • Updated • 721
mradermacher/KernelBench-RLVR-120b-GGUF
Reinforcement Learning • 117B • Updated • 569
Brooooooklyn/Qwen3.8-27B-NVFP4-mlx
Image-Text-to-Text • 27B • Updated • 702 • 1
mradermacher/KernelBench-RLVR-120b-i1-GGUF
Reinforcement Learning • 117B • Updated • 6.28k
DiogenesChen122/Qwen3.6-27B-Lora-20260817
Text Generation • Updated • 10
LW657/Bonsai-8B-gguf
Text Generation • 8B • Updated • 223
NoxNotreve/Bonsai-4B-gguf
Text Generation • 4B • Updated • 118
NoxNotreve/Bonsai-8B-gguf
Text Generation • 8B • Updated • 127
NoxNotreve/Bonsai-27B-gguf
Text Generation • 27B • Updated • 373
AethronPhantom/pyc-kernels
Updated
squ0sh/Bonsai-27B-mlx-1bit
Text Generation • 2B • Updated • 398
r3lax/escha-runtime-qwen3dense
Updated
r3lax/escha-runtime-qwen3moe
Updated • 1
hp0303/Qwen3.8-27B-Texelator-AWBC4
Text Generation • Updated • 13
THISITLLM/bounded-moe
Updated
ZoneTwelve/cifar10-cuda-models
Image Classification • Updated
waqasm86/kaggle-vllm-binaries
Updated
fraQtl/fraqtl-membrane-llamacpp-runtime
Updated
miyuki17/openevo-scientific-runtime-docker
Updated
ngoctham/Breeze-TTS-2
Text-to-Speech • 3B • Updated • 11
DiogenesChen122/Qwen3.8-27B-Lora-20260826
Text Generation • Updated • 1
DiogenesChen122/Qwen3.6-27B-Lora-20260826
Text Generation • Updated
engharat2/Qwen3.8-27B-QUASAR-NVFP4-NINFER
Text Generation • Updated • 351
SceneWorks/ltx-2.5-mlx
Text-to-Video • Updated
DiogenesChen122/Gemma-4-31B-Lora-20260828
Text Generation • Updated • 7