Trending repositories: llm-serving
1 tracked repository tagged with llm-serving, ordered by stars. Use the topic filters below to narrow further.
Filter by topic
1 of 1 repository
thu-pacman/chitu
High-performance inference framework for large language models, focusing on efficiency, flexibility, and availability.
AI summary: A high-performance LLM inference engine supporting diverse GPUs and complex deployments.
3,155infrastructurePythonApache-2.0