Edit Models filters
Model Tree
Apps
Inference Providers
One-click Deployment
Models
52
Active filters: sparse-attention
barelymining/ComfyUI-MiniMax-H3-FastVideo
Updated • 2.15k • 8
garnermccloud/Qwen3.8-Flash-Next-MLX-SSD-Stream
Image-Text-to-Text • 130B • Updated • 820 • 7
ddalcu/Qwen3.8-Flash-Next-MLX-Serve-mixed-4-8bit
Text Generation • 133B • Updated • 5.75k • 13
ddalcu/Qwen3.8-Flash-Next-MLX-Serve-4bit
Text Generation • 129B • Updated • 2.42k • 6
Lynxpda/micro-qwen4exp
0.2B • Updated • 239 • 1
cerebras/Llama-3-CBHybridL-8B
Text Generation • 8B • Updated • 73
cerebras/Llama-3-CBHybridM-8B
Text Generation • 8B • Updated • 66
seconds-0/nsa-117m-byte
Text Generation • 78.3M • Updated • 16
Enxin/VideoNSA
Video-Text-to-Text • 9B • Updated • 160 • 2
openbmb/InfLLM-V2-Long-Sparse-Base
8B • Updated • 150 • 7
sxiong/DHSA-Gemma2-2b-it-BF16
Updated • 8
AXONVERTEX-AI-RESEARCH/InfLLM-V2-Long-Sparse-Base-Q8_0-GGUF
8B • Updated • 16
dororodoroddo/BORA-1.1B-A0.4B-checkpoint
Text Generation • Updated • 31 • 2
amewebstudio/sparseflow-chat
Updated • 1 • 1
amewebstudio/sparseflow-chat-v8
Updated • 12
chaojixiaokeai/CortexNet
Updated
rp440/Qwen3-8b-DSA-index
Text Generation • Updated
smithblack-0/SHRAM
Text Generation • Updated • 18
Alwahsh/Meta-Llama-3.1-8B-Instruct-Butler
Text Generation • 8B • Updated • 22
datasysdev/ann-sparseattention
4B • Updated • 12
sst12345/liveditor
AMLab-UvA/mosaic
Updated • 7
yunyangge/OSP-Next
Text-to-Video • Updated • • 2
smithblack-0/SHRAM-dev
Text Generation • Updated • 20
mesklintech/mesko-tts
Text-to-Speech • Updated • 2 • 2
Vineetha00/synapnet-edge
Updated
Vineetha00/synapnet
Updated
sneedjak/Adelic-Gemma-4-31B-it
31B • Updated • 3
libertywing/FlashMemory-Deepseek-V4
Text Generation • Updated • 22