---
title: 昇腾模型查询助手-昇腾社区
description: 提供昇腾推理和训练场景支持的模型清单及获取链接，方便开发者查询与使用。
keywords: 昇腾模型查询助手,昇腾社区,MindIE,SGLang,MoE,Mixture,Experts,Baichuan  Bloom  ChatGLM  DeepSeek  LLaMA  Qwen
url: https://www.hiascend.com/software/mindie/modellist/show
section: (其他)
---

# 昇腾模型查询助手-昇腾社区

URL: https://www.hiascend.com/software/mindie/modellist/show
描述: 提供昇腾推理和训练场景支持的模型清单及获取链接，方便开发者查询与使用。
关键词: 昇腾模型查询助手,昇腾社区,MindIE,SGLang,MoE,Mixture,Experts,Baichuan  Bloom  ChatGLM  DeepSeek  LLaMA  Qwen

昇腾模型查询助手

0/100

范围说明

本页面中呈现release版本支持的模型，如需查询和体验更多昇腾支持模型，请前往昇腾模型生态全景平台 [了解详情](https://ai.gitcode.com/ascend-model-ecosystem)。

按训练查询

按推理查询

按产品查询

推理引擎

MindIE

vLLM

SGLang

版本

3.0.0

2.3.0

2.2.RC1

2.1.RC2

2.1.RC1

2.0.RC2

2.0.RC1

1.0.RC3

1.0.0

模型类型

大语言模型列表

多模态理解模型列表

多模态生成模型列表

模型系列

MoE（Mixture-of-Experts，混合专家模型）

Baichuan

Bloom

ChatGLM

DeepSeek

LLaMA

Qwen

导出表格

| 模型名称 | 支持引擎 | 版本 | 模型类型 | 模型系列 | 产品 | 模型数据 | 量化 | 服务化 | 长序列支持 | 提示词最大长度 | 权重链接 |
| --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | ---|

| Qwen3-235B-A22B | MindIE | 3.0.0 | 大语言模型列表 | MoE（Mixture-of-Experts，混合专家模型） | Atlas 800I A2 推理服务器（64G）：支持的卡数为16Atlas 800I A3 超节点服务器：支持的卡数为8Atlas 300I Duo 推理卡：不支持 | FP16：仅Atlas 800I A2 推理服务器（64G）和Atlas 800I A3 超节点服务器支持BF16：仅Atlas 800I A2 推理服务器（64G）和Atlas 800I A3 超节点服务器支持 | W8A8量化：支持 | MindIE Motor：支持 | 不支持 |  | 获取权重 |
| --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | --- | ---|
| Qwen3-30B-A3B | MindIE | 3.0.0 | 大语言模型列表 | MoE（Mixture-of-Experts，混合专家模型） | Atlas 800I A2 推理服务器（64G）：支持的卡数为2或4，推荐使用4卡Atlas 800I A3 超节点服务器：支持的卡数为2或4，推荐使用4卡Atlas 300I Duo 推理卡：支持的卡数为2 | FP16：支持BF16：仅Atlas 800I A2 推理服务器（64G）和Atlas 800I A3 超节点服务器支持 | W8A8量化：仅Atlas 800I A2 推理服务器（64G）和Atlas 800I A3 超节点服务器支持 | MindIE Motor：支持 | 不支持 |  | 获取权重 |
| DeepSeek-R1-0528 | MindIE | 3.0.0 | 大语言模型列表 | MoE（Mixture-of-Experts，混合专家模型） | （W8A8）Atlas 800I A2 推理服务器（64G）：支持的卡数为16Atlas 800I A3 超节点服务器：支持的卡数为8Atlas 300I Duo 推理卡：不支持 | FP16：支持BF16：支持 | W4A8量化：仅Atlas 800I A2 推理服务器（64G）和Atlas 800I A3 超节点服务器支持W8A8量化：仅Atlas 800I A2 推理服务器（64G）和Atlas 800I A3 超节点服务器支持W8A8C8量化：仅Atlas 800I A2 推理服务器（64G）和Atlas 800I A3 超节点服务器支持 | MindIE Motor：支持 | Atlas 800I A2 推理服务器（64G）支持的长度最长为128K |  | DeepSeek-R1DeepSeek-R1-0528DeepSeek-R1-0528-w8a8 |
| DeepSeek-V2-236B | MindIE | 3.0.0 | 大语言模型列表 | MoE（Mixture-of-Experts，混合专家模型） | Atlas 800I A2 推理服务器（64G）：支持的卡数为16Atlas 800I A3 超节点服务器：支持的卡数为8Atlas 300I Duo 推理卡：不支持 | FP16：仅Atlas 800I A2 推理服务器（64G）和Atlas 800I A3 超节点服务器支持BF16：仅Atlas 800I A2 推理服务器（64G）和Atlas 800I A3 超节点服务器支持 | W8A8量化：仅Atlas 800I A2 推理服务器（64G）和Atlas 800I A3 超节点服务器支持 | MindIE Motor：支持 | Atlas 800I A2 推理服务器（64G）支持的长度最长为128K |  | 获取权重 |
| DeepSeek-V3-0324 | MindIE | 3.0.0 | 大语言模型列表 | MoE（Mixture-of-Experts，混合专家模型） | （W8A8）Atlas 800I A2 推理服务器（64G）支持的卡数为16Atlas 800I A3 超节点服务器：支持的卡数为8Atlas 300I Duo 推理卡：不支持 | FP16：支持BF16：支持 | W4A8量化：仅Atlas 800I A2 推理服务器（64G）和Atlas 800I A3 超节点服务器支持W8A8量化：仅Atlas 800I A2 推理服务器（64G）和Atlas 800I A3 超节点服务器支持W8A8C8量化：仅Atlas 800I A2 推理服务器（64G）和Atlas 800I A3 超节点服务器支持 | MindIE Motor：支持 | Atlas 800I A2 推理服务器（64G）支持的长度最长为128K |  | DeepSeek-V3DeepSeek-V3-0324 |
| DeepSeek-V3.1 | MindIE | 3.0.0 | 大语言模型列表 | MoE（Mixture-of-Experts，混合专家模型） | （W8A8）Atlas 800I A2 推理服务器（64G）支持的卡数为16Atlas 800I A3 超节点服务器：支持的卡数为8Atlas 300I Duo 推理卡：不支持 | FP16：支持BF16：支持 | W4A8量化：仅Atlas 800I A2 推理服务器（64G）和Atlas 800I A3 超节点服务器支持W8A8量化：仅Atlas 800I A2 推理服务器（64G）和Atlas 800I A3 超节点服务器支持W8A8C8量化：仅Atlas 800I A2 推理服务器（64G）和Atlas 800I A3 超节点服务器支持 | MindIE Motor：支持 | Atlas 800I A2 推理服务器（64G）支持的长度最长为128K |  | 获取权重 |
| Mixtral-8x7B-Instruct-V0.1 | MindIE | 3.0.0 | 大语言模型列表 | MoE（Mixture-of-Experts，混合专家模型） | Atlas 800I A2 推理服务器：支持的卡数为8Atlas 800I A3 超节点服务器：支持的卡数为4Atlas 300I Duo 推理卡：支持的卡数为4 | FP16：支持BF16：仅Atlas 800I A2 推理服务器和Atlas 800I A3 超节点服务器支持 | W8A8量化：仅Atlas 800I A2 推理服务器和Atlas 800I A3 超节点服务器支持 | MindIE Motor：支持 | 不支持 |  | 获取权重 |
| Mixtral-8x22B-Instruct-V0.1 | MindIE | 3.0.0 | 大语言模型列表 | MoE（Mixture-of-Experts，混合专家模型） | Atlas 800I A2 推理服务器（64G）：支持的卡数为8Atlas 800I A3 超节点服务器：支持的卡数为4Atlas 300I Duo 推理卡：不支持 | FP16：仅Atlas 800I A2 推理服务器（64G）和Atlas 800I A3 超节点服务器支持BF16：仅Atlas 800I A2 推理服务器（64G）和Atlas 800I A3 超节点服务器支持 | 不支持 | MindIE Motor：支持 | 不支持 |  | 获取权重 |
| Kimi K2 | MindIE | 3.0.0 | 大语言模型列表 | MoE（Mixture-of-Experts，混合专家模型） | （W8A8）Atlas 800I A2 推理服务器（64G）支持的卡数为32Atlas 800I A3 超节点服务器：支持的卡数为16Atlas 300I Duo 推理卡：不支持 | FP16：支持BF16：支持 | W8A8量化：仅Atlas 800I A2 推理服务器（64G）和Atlas 800I A3 超节点服务器支持 | MindIE Motor：支持 | 不支持 |  | 获取权重 |
| GLM4.5 | MindIE | 3.0.0 | 大语言模型列表 | MoE（Mixture-of-Experts，混合专家模型） | （W8A8）Atlas 800I A2 推理服务器（64G）支持的卡数为16Atlas 800I A3 超节点服务器：待测试验证Atlas 300I Duo 推理卡：不支持 | FP16：支持BF16：支持 | W8A8量化：仅Atlas 800I A2 推理服务器（64G）支持 | MindIE Motor：支持 | 不支持 |  | 获取权重 |

共 12 条

10条/页

20条/页

30条/页

40条/页

1

2

前往

更多模型

本页面中呈现release版本支持的模型，如需查询和体验更多昇腾支持模型，请前往昇腾模型生态全景平台。

立即前往
