mirror of
https://github.com/bentoml/OpenLLM.git
synced 2026-04-27 18:41:34 -04:00
Fixes quantization_config and low_cpu_mem_usage to be available on PyTorch implementation only See changelog for more details on #28
Fixes quantization_config and low_cpu_mem_usage to be available on PyTorch implementation only See changelog for more details on #28