alibaba / rtp-llm

RTP-LLM: Alibaba's high-performance LLM inference engine for diverse applications.
Apache License 2.0
544 stars 50 forks source link

feat: add cpu attention api #80

Closed wenhuanh closed 4 months ago