v0.0.2-alpha13
Add safety server to inference (#2449)
Unify title update and visibility update inference endpoints (#2627)
Introduce model configs to abstract pairings of models and hardware (#2194)
Computing message queue positions (#2235)
Fixed worker requirements wrt transformers (#2621)
adjusted token buffer to pop EOS
fixed inference deploy
inference: allow user change chat title (#2496)