OlivierDehaene
|
6796d38c6d
|
feat(router): add cors allow origin options (#73)
|
2023-02-17 18:22:00 +01:00 |
|
OlivierDehaene
|
5437d49beb
|
feat(router): add max_total_tokens and empty_input validation (#68)
closes #65
|
2023-02-15 21:56:59 +01:00 |
|
OlivierDehaene
|
9af454142a
|
feat: add distributed tracing (#62)
|
2023-02-13 13:02:45 +01:00 |
|
OlivierDehaene
|
b3b7ea0d74
|
feat: Use json formatter by default in docker image
|
2022-11-02 17:29:56 +01:00 |
|
OlivierDehaene
|
3cf6368c77
|
feat(server): Support all AutoModelForCausalLM on a best effort basis
|
2022-10-28 19:24:00 +02:00 |
|
OlivierDehaene
|
beb552127a
|
feat(client): Simplify sharded logic
|
2022-10-22 23:40:05 +02:00 |
|
OlivierDehaene
|
c837893370
|
feat(router): Add max_waiting_tokens
|
2022-10-21 16:40:05 +02:00 |
|
Olivier Dehaene
|
f16f2f5ae1
|
v0.1.0
|
2022-10-20 19:14:44 +02:00 |
|
Olivier Dehaene
|
92c1ecd008
|
feat: Add arguments to CLI
|
2022-10-17 18:27:33 +02:00 |
|
Olivier Dehaene
|
5e5d8766a2
|
feat: Improve error handling
|
2022-10-17 14:59:00 +02:00 |
|
Olivier Dehaene
|
bf99afe916
|
feat: Docker image
|
2022-10-14 15:56:21 +02:00 |
|
Olivier Dehaene
|
39df4d9975
|
Use axum
|
2022-10-11 18:14:39 +02:00 |
|
Olivier Dehaene
|
4c693e6524
|
Refactored gRPC interface
Added validation logic
|
2022-10-11 16:50:54 +02:00 |
|
Olivier Dehaene
|
fa9a088467
|
Add load testing
|
2022-10-11 10:36:51 +02:00 |
|
Olivier Dehaene
|
295831a481
|
Init
|
2022-10-08 12:30:12 +02:00 |
|