Inference Providers
Active filters: sparse
tensorblock/Llama-2-7b-pruned50-retrained-GGUF
Text Generation
• 7B • Updated • 35
mradermacher/phi-2-pruned50-GGUF
3B • Updated • 103
mradermacher/llama2.c-stories110M-pruned50-GGUF
0.1B • Updated • 180
mradermacher/OpenHermes-2.5-Mistral-7B-pruned50-GGUF
7B • Updated • 76
• 1
mradermacher/MiniChat-2-3B-pruned2.4-GGUF
3B • Updated • 87
mradermacher/OpenHermes-2.5-Mistral-7B-pruned50-i1-GGUF
7B • Updated • 65
mradermacher/llama2.c-stories110M-pruned50-i1-GGUF
0.1B • Updated • 172
mradermacher/OpenHermes-2.5-Mistral-7B-pruned2.4-GGUF
7B • Updated • 128
mradermacher/OpenHermes-2.5-Mistral-7B-pruned2.4-i1-GGUF
7B • Updated • 197
tensorblock/OpenHermes-2.5-Mistral-7B-pruned2.4-GGUF
tensorblock/OpenHermes-2.5-Mistral-7B-pruned50-GGUF
mradermacher/Llama-2-7b-dolphin-open_platypus-pruned_70-GGUF
7B • Updated • 47
mradermacher/Llama-2-7b-dolphin-open_platypus-pruned_50-GGUF
7B • Updated • 74
mradermacher/Nous-Hermes-2-Yi-34B-pruned2.4-GGUF
34B • Updated • 41
mradermacher/Nous-Hermes-2-Yi-34B-pruned50-GGUF
34B • Updated • 35
opensearch-project/opensearch-neural-sparse-encoding-multilingual-v1
Feature Extraction
• 0.2B • Updated • 9.93k
• • 23
mradermacher/opensearch-neural-sparse-encoding-doc-v2-mini-GGUF
22.6M • Updated • 118
mradermacher/SparseLlama-3-8B-pruned_50.2of4-GGUF
8B • Updated • 42
• 1
opensearch-project/opensearch-neural-sparse-encoding-doc-v3-distill
Feature Extraction
• 67M • Updated • 2.63k
• • 10
tjingrant/sparsellm-1b-40p
1B • Updated • 4
tjingrant/sparsellm-1b-60p-small-dense
0.7B • Updated • 3
tjingrant/sparsellm-1b-80p
1B • Updated • 5
tjingrant/sparsellm-1b-60p
1B • Updated • 4
tjingrant/sparsellm-1b-20p
1B • Updated • 6
tjingrant/sparsellm-1b-80p-small-dense
0.5B • Updated • 1
tjingrant/sparsellm-1b-40p-small-dense
0.9B • Updated • 7
tjingrant/sparsellm-1b-20p-small-dense
1B • Updated • 5
tensorblock/RedHatAI_llama2.c-stories110M-pruned50-GGUF
0.1B • Updated • 15
sparse-encoder-testing/splade-bert-tiny-nq
Feature Extraction
• 4.42M • Updated • 83.5k
tomaarsen/inference-free-splade-bert-tiny-nq-3e-3-lambda-corpus
Feature Extraction
• Updated • 5