alibaba

alibaba / rtp-llm

#3
1,149199+4 todayCuda

RTP-LLM: Alibaba's high-performance LLM inference engine for diverse applications.

📊 Project Info

Language
Cuda
Stars
1,149
Forks
199
Today
+4
Ranking
#3
Collection
Language
Trending Date
May 31, 2026
Last Push
5/31/2026

🏷️ Topics

gptinferencellamallmllm-servingllmopsmodel-serving

📸 Screenshots

rtp-llm screenshot 1