vLLM
High-throughput, memory-efficient inference and serving engine for LLMs
A high-throughput and memory-efficient inference and serving engine for LLMs
Last updated: 2026-07-25T05:35:42.501859+00:00
Evidence Chain
| Field | Value | Source | Tier | Fetched | Confidence |
|---|---|---|---|---|---|
| archived | No | https://api.github.com/repos/vllm-project/vllm | 1 | 2026-07-25T05:35:42.501859+00:00 | 0.99 |
| description | A high-throughput and memory-efficient inference and serving engine for LLMs | https://api.github.com/repos/vllm-project/vllm | 1 | 2026-07-25T05:35:42.501859+00:00 | 0.99 |
| license | Apache-2.0 | https://api.github.com/repos/vllm-project/vllm | 1 | 2026-07-25T05:35:42.501859+00:00 | 0.99 |
| license_hint_from_doc | Apache-2.0 | D:\我的研究\未來計畫區\網路爬蟲_AI爬蟲與Agent自動化搜尋技術整理_2026-07-14.md | 3 | 2026-07-25T05:35:42.501859+00:00 | 0.8 |
| open_issues | 6042 | https://api.github.com/repos/vllm-project/vllm | 1 | 2026-07-25T05:35:42.501859+00:00 | 0.99 |
| page_title | GitHub - vllm-project/vllm: A high-throughput and memory-efficient inference and serving engine for LLMs | https://github.com/vllm-project/vllm | 2 | 2026-07-25T05:35:42.501859+00:00 | 0.9 |
| pushed_at | 2026-07-25T04:21:36Z | https://api.github.com/repos/vllm-project/vllm | 1 | 2026-07-25T05:35:42.501859+00:00 | 0.99 |
| stars | 87101 | https://api.github.com/repos/vllm-project/vllm | 1 | 2026-07-25T05:35:42.501859+00:00 | 0.99 |
| topics | amd, blackwell, cuda, deepseek, deepseek-v3, gpt, gpt-oss, inference, kimi, llama, llm, llm-serving, model-serving, moe, openai, pytorch, qwen, qwen3, tpu, transformer | https://api.github.com/repos/vllm-project/vllm | 1 | 2026-07-25T05:35:42.501859+00:00 | 0.99 |