text-generation-inference/backends
2025-06-06 15:31:04 +00:00
..
client Revert "feat: improve qwen2-vl startup " (#2924) 2025-01-17 12:09:05 -05:00
gaudi Upgrade to new vllm extension ops for Gaudi backend (fix issue in exponential bucketing) (#3239) 2025-05-22 15:29:16 +02:00
grpc-metadata Upgrading our rustc version. (#2908) 2025-01-15 17:04:03 +01:00
llamacpp Add option to configure prometheus port (#3187) 2025-04-23 20:43:25 +05:30
neuron refactor(neuron): remove obsolete code paths 2025-06-06 15:31:04 +00:00
trtllm Add option to configure prometheus port (#3187) 2025-04-23 20:43:25 +05:30
v2 Add option to configure prometheus port (#3187) 2025-04-23 20:43:25 +05:30
v3 Deepseek R1 for Gaudi backend (#3211) 2025-05-19 16:36:39 +02:00