Edit Models filters
Model Tree
Apps
Inference Providers
One-click Deployment
Models
52
Active filters: sparse-attention
ddalcu/Qwen3.8-Flash-Next-MLX-Serve-mixed-4-8bit
Text Generation • 21B • Updated • 2.42k • 10
ddalcu/Qwen3.8-Flash-Next-MLX-Serve-4bit
Text Generation • 21B • Updated • 1.15k • 5
garnermccloud/Qwen3.8-Flash-Next-MLX-SSD-Stream
Image-Text-to-Text • 25B • Updated • 78 • 2
tencent/HiLS-Attention-7B
Text Generation • 7B • Updated • 369 • 24
cerebras/Llama-3-CBHybridL-8B
Text Generation • 8B • Updated • 55
cerebras/Llama-3-CBHybridM-8B
Text Generation • 8B • Updated • 49
seconds-0/nsa-117m-byte
Text Generation • 78.3M • Updated • 17
Enxin/VideoNSA
Video-Text-to-Text • 9B • Updated • 123 • 2
openbmb/InfLLM-V2-Long-Sparse-Base
8B • Updated • 365 • 7
sxiong/DHSA-Gemma2-2b-it-BF16
Updated • 7
AXONVERTEX-AI-RESEARCH/InfLLM-V2-Long-Sparse-Base-Q8_0-GGUF
8B • Updated • 9
dororodoroddo/BORA-1.1B-A0.4B-checkpoint
Text Generation • Updated • 27 • 2
amewebstudio/sparseflow-chat
Updated • 1 • 1
amewebstudio/sparseflow-chat-v8
Updated • 11
chaojixiaokeai/CortexNet
Updated
rp440/Qwen3-8b-DSA-index
Text Generation • Updated
smithblack-0/SHRAM
Text Generation • Updated • 11
Alwahsh/Meta-Llama-3.1-8B-Instruct-Butler
Text Generation • 8B • Updated • 20
datasysdev/ann-sparseattention
4B • Updated • 12
sst12345/liveditor
AMLab-UvA/mosaic
Updated • 7
yunyangge/OSP-Next
Text-to-Video • Updated • • 2
smithblack-0/SHRAM-dev
Text Generation • Updated • 13
mesklintech/mesko-tts
Text-to-Speech • Updated • 2 • 2
Vineetha00/synapnet-edge
Updated
Vineetha00/synapnet
Updated
sneedjak/Adelic-Gemma-4-31B-it
31B • Updated • 1
libertywing/FlashMemory-Deepseek-V4
Text Generation • Updated • 22
xin1u/UltraFlash
Text-to-Video • Updated • 1