This PR add temporary qtype woq_int4 to avoid affecting other qtype and models. Co-authored-by: leonardozcm <leonardo1997zcm@gmail.com> |
||
|---|---|---|
| .. | ||
| model | ||
| __init__.py | ||
| convert.py | ||
| convert_model.py | ||
| quantize.py | ||
This PR add temporary qtype woq_int4 to avoid affecting other qtype and models. Co-authored-by: leonardozcm <leonardo1997zcm@gmail.com> |
||
|---|---|---|
| .. | ||
| model | ||
| __init__.py | ||
| convert.py | ||
| convert_model.py | ||
| quantize.py | ||