vectorch-ai / ScaleLLM

A high-performance inference system for large language models, designed for production environments.
https://docs.vectorch.com/
Apache License 2.0
316 stars 23 forks source link

feat: added with statement support to release memory and exposed help function for tokenizer #231

Closed guocuimi closed 4 weeks ago

guocuimi commented 4 weeks ago