[Feature] Support AWQ MoE W4A16 Quantization (#142)

Signed-off-by: tangshiwen <tangshiwen@baidu.com> Co-authored-by: Li Wei <liwei.109@outlook.com>
2026-01-26 18:56:05 +08:00
parent 2a998286c0
commit 0711c1abfa
7 changed files with 639 additions and 126 deletions
--- a/docs/source/user_guide/feature_guide/quantization.md
+++ b/docs/source/user_guide/feature_guide/quantization.md
@@ -31,7 +31,7 @@ Like vLLM, we now support quantization methods such as compressed-tensors, AWQ,
      <td style="padding: 10px; border: 1px solid #000;">✅</td>
      <td style="padding: 10px; border: 1px solid #000;">✅</td>
      <td style="padding: 10px; border: 1px solid #000;">✅</td>
-      <td style="padding: 10px; border: 1px solid #000;">WIP</td>
+      <td style="padding: 10px; border: 1px solid #000;">✅</td>
      <td style="padding: 10px; border: 1px solid #000;">✅</td>
      <td style="padding: 10px; border: 1px solid #000;">WIP</td>
    </tr>