Back to Vllm

XPU - Intel® GPUs

docs/models/hardware_supported_models/xpu.md

0.20.15.5 KB
Original Source

XPU - Intel® GPUs

Validated Hardware

Hardware
Intel® Arc™ Pro B-Series Graphics

Text-only Language Models

ModelArchitectureFP16Dynamic FP8MXFP4
openai/gpt-oss-20bGPTForCausalLM
openai/gpt-oss-120bGPTForCausalLM
deepseek-ai/DeepSeek-R1-Distill-Llama-8BLlamaForCausalLM
deepseek-ai/DeepSeek-R1-Distill-Qwen-14BQwenForCausalLM
deepseek-ai/DeepSeek-R1-Distill-Qwen-32BQwenForCausalLM
deepseek-ai/DeepSeek-R1-Distill-Llama-70BLlamaForCausalLM
Qwen/Qwen2.5-72B-InstructQwen2ForCausalLM
Qwen/Qwen3-14BQwen3ForCausalLM
Qwen/Qwen3-32BQwen3ForCausalLM
Qwen/Qwen3-30B-A3BQwen3ForCausalLM
Qwen/Qwen3-30B-A3B-GPTQ-Int4Qwen3ForCausalLM
Qwen/Qwen3-coder-30B-A3B-InstructQwen3ForCausalLM
Qwen/QwQ-32BQwenForCausalLM
deepseek-ai/DeepSeek-V2-LiteDeepSeekForCausalLM
meta-llama/Llama-3.1-8B-InstructLlamaForCausalLM
baichuan-inc/Baichuan2-13B-ChatBaichuanForCausalLM
THUDM/GLM-4-9B-chatGLMForCausalLM
THUDM/CodeGeex4-All-9BCodeGeexForCausalLM
chuhac/TeleChat2-35BLlamaForCausalLM (TeleChat2 based on Llama arch)
01-ai/Yi1.5-34B-ChatYiForCausalLM
THUDM/CodeGeex4-All-9BCodeGeexForCausalLM
deepseek-ai/DeepSeek-Coder-33B-baseDeepSeekCoderForCausalLM
baichuan-inc/Baichuan2-13B-ChatBaichuanForCausalLM
meta-llama/Llama-2-13b-chat-hfLlamaForCausalLM
THUDM/CodeGeex4-All-9BCodeGeexForCausalLM
Qwen/Qwen1.5-14B-ChatQwenForCausalLM
Qwen/Qwen1.5-32B-ChatQwenForCausalLM

Multimodal Language Models

ModelArchitectureFP16Dynamic FP8MXFP4
OpenGVLab/InternVL3_5-8BInternVLForConditionalGeneration
OpenGVLab/InternVL3_5-14BInternVLForConditionalGeneration
OpenGVLab/InternVL3_5-38BInternVLForConditionalGeneration
Qwen/Qwen2-VL-7B-InstructQwen2VLForConditionalGeneration
Qwen/Qwen2.5-VL-72B-InstructQwen2VLForConditionalGeneration
Qwen/Qwen2.5-VL-32B-InstructQwen2VLForConditionalGeneration
THUDM/GLM-4v-9BGLM4vForConditionalGeneration
openbmb/MiniCPM-V-4MiniCPMVForConditionalGeneration

Embedding and Reranker Language Models

ModelArchitectureFP16Dynamic FP8MXFP4
Qwen/Qwen3-Embedding-8BQwen3ForTextEmbedding
Qwen/Qwen3-Reranker-8BQwen3ForSequenceClassification

✅ Runs and optimized.
🟨 Runs and correct but not optimized to green yet.
❌ Does not pass accuracy test or does not run.