This PR add temporary qtype woq_int4 to avoid affecting other qtype and models. Co-authored-by: leonardozcm <leonardo1997zcm@gmail.com> |
||
|---|---|---|
| .. | ||
| llm | ||
This PR add temporary qtype woq_int4 to avoid affecting other qtype and models. Co-authored-by: leonardozcm <leonardo1997zcm@gmail.com> |
||
|---|---|---|
| .. | ||
| llm | ||