RedHatAI/whisper-large-v3-turbo-FP8-dynamic Automatic Speech Recognition • 0.9B • Updated Apr 22, 2025 • 1.97k • 6
RedHatAI/whisper-large-v3-turbo-quantized.w8a8 Automatic Speech Recognition • 0.9B • Updated Apr 22, 2025 • 780 • 4
RedHatAI/whisper-large-v3-turbo-quantized.w4a16 Automatic Speech Recognition • 0.9B • Updated Apr 28 • 2.12k • 8
RedHatAI/whisper-large-v3-quantized.w4a16 Automatic Speech Recognition • 2B • Updated Apr 22, 2025 • 1.81k • 3
RedHatAI/whisper-large-v3-FP8-dynamic Automatic Speech Recognition • 2B • Updated Apr 22, 2025 • 1.19k • 5
RedHatAI/whisper-large-v3-quantized.w8a8 Automatic Speech Recognition • 2B • Updated Apr 22, 2025 • 1.88k • 1
RedHatAI/whisper-medium-quantized.w8a8 Automatic Speech Recognition • 0.8B • Updated Apr 22, 2025 • 71
RedHatAI/whisper-tiny-quantized.w8a8 Automatic Speech Recognition • 57.8M • Updated Apr 22, 2025 • 27 • 1
RedHatAI/whisper-small-quantized.w8a8 Automatic Speech Recognition • 0.3B • Updated Apr 22, 2025 • 1.06k
RedHatAI/whisper-large-v2-quantized.w8a8 Automatic Speech Recognition • 2B • Updated Apr 22, 2025 • 27
RedHatAI/whisper-medium-quantized.w4a16 Automatic Speech Recognition • 0.8B • Updated Apr 22, 2025 • 59
RedHatAI/whisper-small-quantized.w4a16 Automatic Speech Recognition • 0.3B • Updated Apr 22, 2025 • 46 • 1
RedHatAI/whisper-large-v2-quantized.w4a16 Automatic Speech Recognition • 2B • Updated Apr 22, 2025 • 148 • 1
RedHatAI/whisper-large-v2-quantized.w4a16 Automatic Speech Recognition • 2B • Updated Apr 22, 2025 • 148 • 1
The Optimal BERT Surgeon: Scalable and Accurate Second-Order Pruning for Large Language Models Paper • 2203.07259 • Published Mar 14, 2022 • 4