Inference Providers
Active filters: sparse
tensorblock/Llama-2-7b-pruned50-retrained-GGUF
Text Generation
• 7B • Updated • 114
mradermacher/phi-2-pruned50-GGUF
3B • Updated • 312
mradermacher/llama2.c-stories110M-pruned50-GGUF
0.1B • Updated • 172
mradermacher/OpenHermes-2.5-Mistral-7B-pruned50-GGUF
7B • Updated • 248
• 1
mradermacher/MiniChat-2-3B-pruned2.4-GGUF
3B • Updated • 255
mradermacher/OpenHermes-2.5-Mistral-7B-pruned50-i1-GGUF
7B • Updated • 600
mradermacher/llama2.c-stories110M-pruned50-i1-GGUF
0.1B • Updated • 298
mradermacher/OpenHermes-2.5-Mistral-7B-pruned2.4-GGUF
7B • Updated • 280
mradermacher/OpenHermes-2.5-Mistral-7B-pruned2.4-i1-GGUF
7B • Updated • 595
tensorblock/OpenHermes-2.5-Mistral-7B-pruned2.4-GGUF
7B • Updated • 137
tensorblock/OpenHermes-2.5-Mistral-7B-pruned50-GGUF
7B • Updated • 108
mradermacher/Llama-2-7b-dolphin-open_platypus-pruned_70-GGUF
7B • Updated • 291
mradermacher/Llama-2-7b-dolphin-open_platypus-pruned_50-GGUF
7B • Updated • 294
mradermacher/Nous-Hermes-2-Yi-34B-pruned2.4-GGUF
34B • Updated • 222
mradermacher/Nous-Hermes-2-Yi-34B-pruned50-GGUF
34B • Updated • 195
ibm-granite/granite-embedding-30m-sparse
Feature Extraction
• 30.3M • Updated • 12.8k
• • 26
opensearch-project/opensearch-neural-sparse-encoding-multilingual-v1
Feature Extraction
• 0.2B • Updated • 25.6k
• • 24
mradermacher/opensearch-neural-sparse-encoding-doc-v2-mini-GGUF
22.6M • Updated • 260
mradermacher/SparseLlama-3-8B-pruned_50.2of4-GGUF
8B • Updated • 274
• 1
opensearch-project/opensearch-neural-sparse-encoding-doc-v3-distill
Feature Extraction
• 67M • Updated • 34.4k
• • 10
tjingrant/sparsellm-1b-40p
1B • Updated • 21
tjingrant/sparsellm-1b-60p-small-dense
0.7B • Updated • 8
tjingrant/sparsellm-1b-80p
1B • Updated • 17
tjingrant/sparsellm-1b-60p
1B • Updated • 15
tjingrant/sparsellm-1b-20p
1B • Updated • 22
tjingrant/sparsellm-1b-80p-small-dense
0.5B • Updated • 11
tjingrant/sparsellm-1b-40p-small-dense
0.9B • Updated • 12
tjingrant/sparsellm-1b-20p-small-dense
1B • Updated • 13
tensorblock/RedHatAI_llama2.c-stories110M-pruned50-GGUF
0.1B • Updated • 120
sparse-encoder-testing/splade-bert-tiny-nq
Feature Extraction
• 4.42M • Updated • 91.8k