v0.0.3-alpha18
Add safety server to inference (#2449)
Add sampling presets based on community feedback (#2723)
Introduce model configs to abstract pairings of models and hardware (#2194)
Added max messages and max message length settings for inference (#2774)
Implement auto scroll when stream message (#2839)
adjusted token buffer to pop EOS
add translate button to important readme's (#2816)
inference: allow user change chat title (#2496)