-
RedHatAI/Mistral-Small-3.2-24B-Instruct-2506-NVFP4
Text Generation • 14B • Updated • 6.11k • 3 -
RedHatAI/Qwen3-VL-235B-A22B-Instruct-NVFP4
Text Generation • 133B • Updated • 3.78k • 4 -
RedHatAI/Qwen3-235B-A22B-Instruct-2507-NVFP4
Text Generation • 136B • Updated • 889 • 4 -
RedHatAI/Qwen3-235B-A22B-NVFP4
Text Generation • 136B • Updated • 243
AI & ML interests
OpenSource and AI
Recent Activity
View all activity
September 2025 Collection of third-party generative AI models validated by Red Hat AI for use across the Red Hat AI Product Portfolio.
-
RedHatAI/DeepSeek-R1-0528-quantized.w4a16
Text Generation • 104B • Updated • 639 • 12 -
RedHatAI/Qwen3-8B-FP8-dynamic
Text Generation • 8B • Updated • 11.8k • 9 -
RedHatAI/Kimi-K2-Instruct-quantized.w4a16
Text Generation • 146B • Updated • 200 • 12 -
RedHatAI/gemma-3n-E4B-it-FP8-dynamic
Text Generation • 8B • Updated • 138k • 3
May 2025 Collection of third-party generative AI models validated by Red Hat AI for use across the Red Hat AI Product Portfolio.
-
RedHatAI/Llama-4-Scout-17B-16E-Instruct-FP8-dynamic
Image-Text-to-Text • 109B • Updated • 26.4k • 27 -
RedHatAI/Llama-4-Scout-17B-16E-Instruct-quantized.w4a16
Image-Text-to-Text • 20B • Updated • 18.9k • 12 -
RedHatAI/Llama-4-Scout-17B-16E-Instruct
Image-Text-to-Text • 109B • Updated • 3.81k -
RedHatAI/Llama-4-Maverick-17B-128E-Instruct
Image-Text-to-Text • 402B • Updated • 43 • 2
Collection of quantized Gemma 3 models created by Google.
-
RedHatAI/gemma-3-27b-it-quantized.w4a16
Any-to-Any • 7B • Updated • 12.9k • 10 -
RedHatAI/gemma-3-12b-it-quantized.w4a16
Any-to-Any • 4B • Updated • 1.15k • 2 -
RedHatAI/gemma-3-4b-it-quantized.w4a16
Any-to-Any • 2B • Updated • 1.03k • 2 -
RedHatAI/gemma-3-1b-it-quantized.w8a8
Text Generation • 1B • Updated • 58.2k
Quantized variants of the Llama 4 release by Meta.
-
RedHatAI/Llama-4-Scout-17B-16E-Instruct-FP8-dynamic
Image-Text-to-Text • 109B • Updated • 26.4k • 27 -
RedHatAI/Llama-4-Scout-17B-16E-Instruct-quantized.w4a16
Image-Text-to-Text • 20B • Updated • 18.9k • 12 -
RedHatAI/Llama-4-Maverick-17B-128E-Instruct-FP8
Image-Text-to-Text • 402B • Updated • 275 • 2 -
RedHatAI/Llama-4-Maverick-17B-128E-Instruct-quantized.w4a16
Image-Text-to-Text • 59B • Updated • 218 • 1
Quantized variants of Mistral Small 3.1 (2503) Instruct.
-
RedHatAI/Mistral-Small-3.1-24B-Instruct-2503-FP8-dynamic
Image-Text-to-Text • 24B • Updated • 18.7k • 9 -
RedHatAI/Mistral-Small-3.1-24B-Instruct-2503-quantized.w8a8
Image-Text-to-Text • 24B • Updated • 4.83k • 5 -
RedHatAI/Mistral-Small-3.1-24B-Instruct-2503-quantized.w4a16
Image-Text-to-Text • 5B • Updated • 31.8k • 10
Quantized variants of Meta Llama 3.3 multilingual large language model (LLM) is an instruction tuned generative model in 70B (text in/text out).
Quantized Granite models from IBM Research.
-
RedHatAI/granite-3.1-8b-instruct-quantized.w8a8
Text Generation • 8B • Updated • 276 • 2 -
RedHatAI/granite-3.1-2b-base-quantized.w8a8
Text Generation • 3B • Updated • 38 -
RedHatAI/granite-3.1-8b-instruct-quantized.w4a16
Text Generation • 1B • Updated • 604 • 1 -
RedHatAI/granite-3.1-2b-instruct-quantized.w8a8
Text Generation • 3B • Updated • 28
October 2025 Collection of third-party generative AI models validated by Red Hat AI for use across the Red Hat AI Product Portfolio.
-
RedHatAI/gpt-oss-120b
Text Generation • 120B • Updated • 721 • 3 -
RedHatAI/gpt-oss-20b
Text Generation • 22B • Updated • 8.81k • 5 -
RedHatAI/Qwen3-Coder-480B-A35B-Instruct-FP8
Text Generation • 480B • Updated • 117 • 2 -
RedHatAI/whisper-large-v3-turbo-quantized.w4a16
Automatic Speech Recognition • 0.2B • Updated • 226 • 6
Collection of quantized whisper models created by OpenAI
-
RedHatAI/whisper-large-v3-turbo-quantized.w4a16
Automatic Speech Recognition • 0.2B • Updated • 226 • 6 -
RedHatAI/whisper-large-v3-turbo-quantized.w8a8
Automatic Speech Recognition • 0.9B • Updated • 74 • 4 -
RedHatAI/whisper-large-v3-turbo-FP8-dynamic
Automatic Speech Recognition • 0.9B • Updated • 482 • 6 -
RedHatAI/whisper-tiny-FP8-Dynamic
Automatic Speech Recognition • 57.8M • Updated • 133
Collection of quantized Qwen 3 models from Alibaba Cloud.
Quantized variants of Phi-4 family of small language and multi-modal models by Microsoft.
Quantized variants of Qwen 2.5 Instruct and Qwen VL models
-
RedHatAI/Qwen2.5-VL-7B-Instruct-quantized.w8a8
Image-to-Text • 8B • Updated • 1.6k • 8 -
RedHatAI/Qwen2.5-VL-7B-Instruct-quantized.w4a16
Image-to-Text • 3B • Updated • 765 • 7 -
RedHatAI/Qwen2.5-7B-quantized.w8a8
Text Generation • 8B • Updated • 84 • 1 -
RedHatAI/Qwen2.5-VL-72B-Instruct-FP8-dynamic
Image-to-Text • 73B • Updated • 8.63k • 14
Collection of kernels from vLLM built using https://github.com/huggingface/kernel-builder
-
RedHatAI/Mistral-Small-3.2-24B-Instruct-2506-NVFP4
Text Generation • 14B • Updated • 6.11k • 3 -
RedHatAI/Qwen3-VL-235B-A22B-Instruct-NVFP4
Text Generation • 133B • Updated • 3.78k • 4 -
RedHatAI/Qwen3-235B-A22B-Instruct-2507-NVFP4
Text Generation • 136B • Updated • 889 • 4 -
RedHatAI/Qwen3-235B-A22B-NVFP4
Text Generation • 136B • Updated • 243
October 2025 Collection of third-party generative AI models validated by Red Hat AI for use across the Red Hat AI Product Portfolio.
-
RedHatAI/gpt-oss-120b
Text Generation • 120B • Updated • 721 • 3 -
RedHatAI/gpt-oss-20b
Text Generation • 22B • Updated • 8.81k • 5 -
RedHatAI/Qwen3-Coder-480B-A35B-Instruct-FP8
Text Generation • 480B • Updated • 117 • 2 -
RedHatAI/whisper-large-v3-turbo-quantized.w4a16
Automatic Speech Recognition • 0.2B • Updated • 226 • 6
September 2025 Collection of third-party generative AI models validated by Red Hat AI for use across the Red Hat AI Product Portfolio.
-
RedHatAI/DeepSeek-R1-0528-quantized.w4a16
Text Generation • 104B • Updated • 639 • 12 -
RedHatAI/Qwen3-8B-FP8-dynamic
Text Generation • 8B • Updated • 11.8k • 9 -
RedHatAI/Kimi-K2-Instruct-quantized.w4a16
Text Generation • 146B • Updated • 200 • 12 -
RedHatAI/gemma-3n-E4B-it-FP8-dynamic
Text Generation • 8B • Updated • 138k • 3
May 2025 Collection of third-party generative AI models validated by Red Hat AI for use across the Red Hat AI Product Portfolio.
-
RedHatAI/Llama-4-Scout-17B-16E-Instruct-FP8-dynamic
Image-Text-to-Text • 109B • Updated • 26.4k • 27 -
RedHatAI/Llama-4-Scout-17B-16E-Instruct-quantized.w4a16
Image-Text-to-Text • 20B • Updated • 18.9k • 12 -
RedHatAI/Llama-4-Scout-17B-16E-Instruct
Image-Text-to-Text • 109B • Updated • 3.81k -
RedHatAI/Llama-4-Maverick-17B-128E-Instruct
Image-Text-to-Text • 402B • Updated • 43 • 2
Collection of quantized Gemma 3 models created by Google.
-
RedHatAI/gemma-3-27b-it-quantized.w4a16
Any-to-Any • 7B • Updated • 12.9k • 10 -
RedHatAI/gemma-3-12b-it-quantized.w4a16
Any-to-Any • 4B • Updated • 1.15k • 2 -
RedHatAI/gemma-3-4b-it-quantized.w4a16
Any-to-Any • 2B • Updated • 1.03k • 2 -
RedHatAI/gemma-3-1b-it-quantized.w8a8
Text Generation • 1B • Updated • 58.2k
Collection of quantized whisper models created by OpenAI
-
RedHatAI/whisper-large-v3-turbo-quantized.w4a16
Automatic Speech Recognition • 0.2B • Updated • 226 • 6 -
RedHatAI/whisper-large-v3-turbo-quantized.w8a8
Automatic Speech Recognition • 0.9B • Updated • 74 • 4 -
RedHatAI/whisper-large-v3-turbo-FP8-dynamic
Automatic Speech Recognition • 0.9B • Updated • 482 • 6 -
RedHatAI/whisper-tiny-FP8-Dynamic
Automatic Speech Recognition • 57.8M • Updated • 133
Quantized variants of the Llama 4 release by Meta.
-
RedHatAI/Llama-4-Scout-17B-16E-Instruct-FP8-dynamic
Image-Text-to-Text • 109B • Updated • 26.4k • 27 -
RedHatAI/Llama-4-Scout-17B-16E-Instruct-quantized.w4a16
Image-Text-to-Text • 20B • Updated • 18.9k • 12 -
RedHatAI/Llama-4-Maverick-17B-128E-Instruct-FP8
Image-Text-to-Text • 402B • Updated • 275 • 2 -
RedHatAI/Llama-4-Maverick-17B-128E-Instruct-quantized.w4a16
Image-Text-to-Text • 59B • Updated • 218 • 1
Collection of quantized Qwen 3 models from Alibaba Cloud.
Quantized variants of Mistral Small 3.1 (2503) Instruct.
-
RedHatAI/Mistral-Small-3.1-24B-Instruct-2503-FP8-dynamic
Image-Text-to-Text • 24B • Updated • 18.7k • 9 -
RedHatAI/Mistral-Small-3.1-24B-Instruct-2503-quantized.w8a8
Image-Text-to-Text • 24B • Updated • 4.83k • 5 -
RedHatAI/Mistral-Small-3.1-24B-Instruct-2503-quantized.w4a16
Image-Text-to-Text • 5B • Updated • 31.8k • 10
Quantized variants of Phi-4 family of small language and multi-modal models by Microsoft.
Quantized variants of Meta Llama 3.3 multilingual large language model (LLM) is an instruction tuned generative model in 70B (text in/text out).
Quantized variants of Qwen 2.5 Instruct and Qwen VL models
-
RedHatAI/Qwen2.5-VL-7B-Instruct-quantized.w8a8
Image-to-Text • 8B • Updated • 1.6k • 8 -
RedHatAI/Qwen2.5-VL-7B-Instruct-quantized.w4a16
Image-to-Text • 3B • Updated • 765 • 7 -
RedHatAI/Qwen2.5-7B-quantized.w8a8
Text Generation • 8B • Updated • 84 • 1 -
RedHatAI/Qwen2.5-VL-72B-Instruct-FP8-dynamic
Image-to-Text • 73B • Updated • 8.63k • 14
Quantized Granite models from IBM Research.
-
RedHatAI/granite-3.1-8b-instruct-quantized.w8a8
Text Generation • 8B • Updated • 276 • 2 -
RedHatAI/granite-3.1-2b-base-quantized.w8a8
Text Generation • 3B • Updated • 38 -
RedHatAI/granite-3.1-8b-instruct-quantized.w4a16
Text Generation • 1B • Updated • 604 • 1 -
RedHatAI/granite-3.1-2b-instruct-quantized.w8a8
Text Generation • 3B • Updated • 28
Collection of kernels from vLLM built using https://github.com/huggingface/kernel-builder