text-generation-inference/server/text_generation_server/layers/moe
2025-01-28 17:05:53 +00:00
..
__init__.py add deepseekv3 2025-01-28 17:05:53 +00:00
fp8.py add deepseekv3 2025-01-28 17:05:53 +00:00
gptq_marlin.py Add support for fused MoE Marlin for AWQ (#2616) 2024-10-08 11:56:41 +02:00
unquantized.py add deepseekv3 2025-01-28 17:05:53 +00:00