bd-iaas-us / vllm

A high-throughput and memory-efficient inference and serving engine for LLMs
https://docs.vllm.ai
Apache License 2.0
1 stars 0 forks source link

Longrope design #7

Closed chizhang118 closed 1 month ago

chizhang118 commented 2 months ago

longrope design - https://bytedance.larkoffice.com/wiki/K4EZwogNfiYgxskvDe6cNvEMnLe