AICreatorHub
NewsToolsPromptsModelsGuidesRepos
Trending1,050,000 Tokens: GPT-6 Astra vs Claude Fable 5.1 in 3 Numbers
Search…
Search…NewsToolsPromptsModelsGuidesRepos
AICreatorHub

India's bilingual AI knowledge hub.

ExploreNewsToolsModelsGuidesPrompts
DiscoverDealsHire an AI ExpertStoreAI Tool QuizBest AI For...
LegalAboutContactPrivacy PolicyTermsDisclaimer
FollowX / TwitterYouTubeRSS
© 2026 AICreatorHub. All rights reserved.
HomeReposvllm
vllm-project

vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

View on GitHub Website
Stars
90,690
Forks
21,550
Language
Python
Licence
Apache-2.0
amdblackwellcudadeepseekdeepseek-v3gptgpt-ossinference
Created 9 Feb 2023Last push 1 Sept 2026
Share:

Stars and forks are a snapshot taken when this directory was last refreshed, not a live count — the figure here matches the one in our articles and videos. Last refreshed 5 Sept 2026.