vllm:基于 PyTorch 生态的 LLM 推理与服务库项目
Star
0
Fork
0
Code
Introduction
Code
Issues
Pull Requests
Pipeline
Actions
Discussion
Wiki
Members
1
Analysis
Settings
Star
0
Fork
0
A high-throughput and memory-efficient inference and serving engine for LLMs
main
Branch
176
Tags
99
IDE
ZIP
Clone
README
Introduction
A high-throughput and memory-efficient inference and serving engine for LLMs
Apache_License_v2.0
Python
12.29 K
Commits
Customize your domain
README
Rulesets
Report repository
Report repository