Universal Dynamic Curated DirectoryUDCD MVP
← Back to home Archive

vLLM

High-throughput, memory-efficient inference and serving engine for LLMs

A high-throughput and memory-efficient inference and serving engine for LLMs

Local Model Tools

GitHub

Last updated: 2026-07-25T05:35:42.501859+00:00

Evidence Chain

FieldValueSourceTierFetchedConfidence
archived No https://api.github.com/repos/vllm-project/vllm 1 2026-07-25T05:35:42.501859+00:00 0.99
description A high-throughput and memory-efficient inference and serving engine for LLMs https://api.github.com/repos/vllm-project/vllm 1 2026-07-25T05:35:42.501859+00:00 0.99
license Apache-2.0 https://api.github.com/repos/vllm-project/vllm 1 2026-07-25T05:35:42.501859+00:00 0.99
license_hint_from_doc Apache-2.0 D:\我的研究\未來計畫區\網路爬蟲_AI爬蟲與Agent自動化搜尋技術整理_2026-07-14.md 3 2026-07-25T05:35:42.501859+00:00 0.8
open_issues 6042 https://api.github.com/repos/vllm-project/vllm 1 2026-07-25T05:35:42.501859+00:00 0.99
page_title GitHub - vllm-project/vllm: A high-throughput and memory-efficient inference and serving engine for LLMs https://github.com/vllm-project/vllm 2 2026-07-25T05:35:42.501859+00:00 0.9
pushed_at 2026-07-25T04:21:36Z https://api.github.com/repos/vllm-project/vllm 1 2026-07-25T05:35:42.501859+00:00 0.99
stars 87101 https://api.github.com/repos/vllm-project/vllm 1 2026-07-25T05:35:42.501859+00:00 0.99
topics amd, blackwell, cuda, deepseek, deepseek-v3, gpt, gpt-oss, inference, kimi, llama, llm, llm-serving, model-serving, moe, openai, pytorch, qwen, qwen3, tpu, transformer https://api.github.com/repos/vllm-project/vllm 1 2026-07-25T05:35:42.501859+00:00 0.99