text-generation-inference/server/text_generation_server/layers/moe
2025-01-20 13:55:54 +00:00
..
__init__.py Add fp8 support moe models 2025-01-20 13:55:54 +00:00
fp8.py Add fp8 support moe models 2025-01-20 13:55:54 +00:00
gptq_marlin.py Add support for fused MoE Marlin for AWQ (#2616) 2024-10-08 11:56:41 +02:00
unquantized.py Add fp8 support moe models 2025-01-20 13:55:54 +00:00