text-generation-inference/server/text_generation_server/layers
Wang, Yi 3ac7df2b6d
hotfix : enable intel ipex cpu and xpu in python3.11 ()
enable intel ipex cpu and xpu in python3.11

Signed-off-by: Wang, Yi A <yi.a.wang@intel.com>
2024-09-12 17:23:49 +02:00
..
attention hotfix : enable intel ipex cpu and xpu in python3.11 () 2024-09-12 17:23:49 +02:00
awq feat: add ruff and resolve issue () 2024-07-26 10:29:09 -04:00
gptq Upgrading exl2. () 2024-08-14 11:58:08 +02:00
marlin Handle GPTQ-Marlin loading in GPTQMarlinWeightLoader () 2024-07-31 13:08:41 +02:00
__init__.py feat: add ruff and resolve issue () 2024-07-26 10:29:09 -04:00
bnb.py feat: add ruff and resolve issue () 2024-07-26 10:29:09 -04:00
conv.py Refactor layers. () 2024-05-13 12:44:30 +02:00
eetq.py feat(fp8): use fbgemm kernels and load fp8 weights directly () 2024-07-20 19:02:04 +02:00
exl2.py Add support for Deepseek V2 () 2024-07-19 17:23:20 +02:00
fp8.py feat: add ruff and resolve issue () 2024-07-26 10:29:09 -04:00
layernorm.py Removing IPEX_AVAIL. () 2024-06-25 13:20:57 +02:00
linear.py feat: add ruff and resolve issue () 2024-07-26 10:29:09 -04:00
lora.py feat: add ruff and resolve issue () 2024-07-26 10:29:09 -04:00
medusa.py Prefix caching () 2024-08-20 11:15:30 +02:00
mlp.py Tied embeddings in MLP speculator. () 2024-08-29 17:44:54 +02:00
rotary.py feat: add ruff and resolve issue () 2024-07-26 10:29:09 -04:00
speculative.py feat: add ruff and resolve issue () 2024-07-26 10:29:09 -04:00
tensor_parallel.py feat: add ruff and resolve issue () 2024-07-26 10:29:09 -04:00