ipex-llm/python/llm/example/CPU/HF-Transformers-AutoModels/Model
2024-04-02 19:54:30 -07:00
..
aquila Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
aquila2 Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
baichuan Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
baichuan2 Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
bluelm Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
chatglm Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
chatglm2 Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
chatglm3 Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
codellama Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
codeshell Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
deciLM-7b Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
deepseek Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
deepseek-moe Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
distil-whisper Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
dolly_v1 Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
dolly_v2 Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
falcon Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
flan-t5 Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
fuyu Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
gemma Update pip install to use --extra-index-url for ipex package (#10557) 2024-03-28 09:56:23 +08:00
internlm Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
internlm-xcomposer Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
internlm2 Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
llama2 Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
mistral Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
mixtral Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
moss update readme (#10632) 2024-04-02 19:54:30 -07:00
mpt Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
phi-1_5 Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
phi-2 Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
phixtral Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
phoenix Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
qwen Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
qwen-vl Fix Qwen-VL example problem (#10582) 2024-04-02 12:17:30 -07:00
qwen1.5 Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
redpajama Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
replit Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
skywork Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
solar Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
starcoder Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
vicuna Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
whisper Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
wizardcoder-python Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
yi Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
yuan2 Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
ziya Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00
README.md Update_document by heyang (#30) 2024-03-25 10:06:02 +08:00

IPEX-LLM Transformers INT4 Optimization for Large Language Model

You can use IPEX-LLM to run any Huggingface Transformer models with INT4 optimizations on either servers or laptops. This directory contains example scripts to help you quickly get started using IPEX-LLM to run some popular open-source models in the community. Each model has its own dedicated folder, where you can find detailed instructions on how to install and run it.

To run the examples, we recommend using Intel® Xeon® processors (server), or >= 12th Gen Intel® Core™ processor (client).

For OS, IPEX-LLM supports Ubuntu 20.04 or later (glibc>=2.17), CentOS 7 or later (glibc>=2.17), and Windows 10/11.

Best Known Configuration on Linux

For better performance, it is recommended to set environment variables on Linux with the help of IPEX-LLM:

pip install ipex-llm
source ipex-llm-init