Inference Providers
Active filters: sparse
RedHatAI/Llama-2-7b-pruned50-retrained
Text Generation
• 7B • Updated • 30
RedHatAI/Llama-2-7b-pruned70-retrained
Text Generation
• 7B • Updated • 178
• 1
RedHatAI/Llama-2-7b-ultrachat200k-pruned_50
Text Generation
• 7B • Updated • 26
RedHatAI/Llama-2-7b-ultrachat200k-pruned_70
Text Generation
• 7B • Updated • 21
RedHatAI/Llama-2-7b-ultrachat200k-pruned_50-quantized-deepsparse
Text Generation
• Updated • 19
RedHatAI/Llama-2-7b-ultrachat200k-pruned_70-quantized-deepsparse
Text Generation
• Updated • 21
RedHatAI/Llama-2-7b-evol-code-alpaca-pruned_50
Text Generation
• 7B • Updated • 29
RedHatAI/Llama-2-7b-evol-code-alpaca-pruned_70
Text Generation
• 7B • Updated • 24
RedHatAI/Llama-2-7b-evol-code-alpaca-pruned_50-quantized-deepsparse
Text Generation
• Updated • 19
RedHatAI/Llama-2-7b-evol-code-alpaca-pruned_70-quantized-deepsparse
Text Generation
• Updated • 21
RedHatAI/Llama-2-7b-dolphin-open_platypus-pruned_50
Text Generation
• 7B • Updated • 26
RedHatAI/Llama-2-7b-dolphin-open_platypus-pruned_70
Text Generation
• 7B • Updated • 20
RedHatAI/Llama-2-7b-dolphin-open_platypus-pruned_50-quantized-deepsparse
Text Generation
• Updated • 21
RedHatAI/Llama-2-7b-dolphin-open_platypus-pruned_70-quantized-deepsparse
Text Generation
• Updated • 29
• 1
kettleguts/zephyr-7b-beta_sparse05
Text Generation
• 7B • Updated • 59
dtransposed/llama2.c-stories110M-pruned50-compressed-tensors
Text Generation
• 0.1B • Updated • 19
RedHatAI/Llama-2-7b-gsm8k-pruned_50
Text Generation
• 7B • Updated • 28
• 1
RedHatAI/Llama-2-7b-gsm8k-pruned_70
Text Generation
• 7B • Updated • 18
mradermacher/Llama-2-7b-pruned70-retrained-gsm8k-GGUF
7B • Updated • 1.7k
RedHatAI/SparseLlama-3-8B-pruned_50.2of4
Text Generation
• 8B • Updated • 17
vuiseng9/ov-mpt-7b-gsm8k-sparse70
Text Generation
• Updated opensearch-project/opensearch-neural-sparse-encoding-v2-distill
Feature Extraction
• 67M • Updated • 12.4k
• • 10
opensearch-project/opensearch-neural-sparse-encoding-doc-v2-distill
Feature Extraction
• 67M • Updated • 26.8k
• • 19
opensearch-project/opensearch-neural-sparse-encoding-doc-v2-mini
Feature Extraction
• 22.7M • Updated • 20.7k
• • 6
mradermacher/Nous-Hermes-2-SOLAR-10.7B-pruned2.4-GGUF
11B • Updated • 162
mradermacher/Nous-Hermes-2-SOLAR-10.7B-pruned2.4-i1-GGUF
11B • Updated • 282
tensorblock/llama2.c-stories110M-pruned50-GGUF
0.1B • Updated • 15
tensorblock/Llama-2-7b-pruned50-retrained-GGUF
Text Generation
• 7B • Updated • 33
mradermacher/phi-2-pruned50-GGUF
3B • Updated • 233
mradermacher/llama2.c-stories110M-pruned50-GGUF
0.1B • Updated • 211