RedHatAI/Meta-Llama-3-8B-Instruct-quantized.w4a16 Text Generation • 8B • Updated Jul 18, 2024 • 334 • 2
RedHatAI/Meta-Llama-3-70B-Instruct-quantized.w8a16 Text Generation • 71B • Updated Jul 18, 2024 • 34 • 6
RedHatAI/Meta-Llama-3-8B-Instruct-quantized.w8a16 Text Generation • 8B • Updated Jul 18, 2024 • 277 • 3
RedHatAI/SparseLLama-2-7b-ultrachat_200k-pruned_50.2of4 Text Generation • 7B • Updated Jul 7, 2024 • 41
RedHatAI/SparseLlama-2-7b-evolcodealpaca-pruned_50.2of4 Text Generation • 7B • Updated Jul 3, 2024 • 28
RedHatAI/SparseLlama-2-7b-cnn-daily-mail-pruned_50.2of4 Text Generation • 7B • Updated May 21, 2024 • 26
RedHatAI/Llama-2-7b-cnn-daily-mail-pruned_70-quantized-deepsparse Text Generation • Updated May 17, 2024 • 16
RedHatAI/Llama-2-7b-cnn-daily-mail-pruned_50-quantized-deepsparse Text Generation • Updated May 17, 2024 • 17
RedHatAI/Llama-2-7b-dolphin-open_platypus-pruned_70-quantized-deepsparse Text Generation • Updated May 16, 2024 • 30 • 1
RedHatAI/Llama-2-7b-dolphin-open_platypus-pruned_50-quantized-deepsparse Text Generation • Updated May 16, 2024 • 20
RedHatAI/Llama-2-7b-ultrachat200k-pruned_70-quantized-deepsparse Text Generation • Updated May 15, 2024 • 17