Skip to content
Navigation Menu
Sign in
Appearance settings
Platform
AI CODE CREATION
GitHub Copilot
Write better code with AI
GitHub Copilot app
Direct agents from issue to merge
MCP Registry
Integrate external tools
DEVELOPER WORKFLOWS
Actions
Automate any workflow
Codespaces
Instant dev environments
Issues
Plan and track work
Code Review
Manage code changes
Code Quality
Enforce quality at merge
APPLICATION SECURITY
GitHub Advanced Security
Find and fix vulnerabilities
Code security
Secure your code as you build
Secret protection
Stop leaks before they start
EXPLORE
Why GitHub
Documentation
Blog
Changelog
Marketplace
View all features
Solutions
BY COMPANY SIZE
Enterprises
Small and medium teams
Startups
Nonprofits
BY USE CASE
App Modernization
DevSecOps
DevOps
CI/CD
View all use cases
BY INDUSTRY
Healthcare
Financial services
Manufacturing
Government
View all industries
View all solutions
Resources
EXPLORE BY TOPIC
AI
Software Development
DevOps
Security
View all topics
EXPLORE BY TYPE
Customer stories
Events & webinars
Ebooks & reports
Business insights
GitHub Skills
SUPPORT & SERVICES
Documentation
Customer support
Community forum
Trust center
Partners
View all resources
Open Source
COMMUNITY
GitHub Sponsors
Fund open source developers
PROGRAMS
Security Lab
Maintainer Community
Accelerator
GitHub Stars
Archive Program
REPOSITORIES
Topics
Trending
Collections
Enterprise
ENTERPRISE SOLUTIONS
Enterprise platform
AI-powered developer platform
AVAILABLE ADD-ONS
GitHub Advanced Security
Enterprise-grade security features
Copilot for Business
Enterprise-grade AI features
Premium Support
Enterprise-grade 24/7 support
Pricing
Search
/
Sign in
Sign up
Appearance settings
You signed in with another tab or window.
Reload
to refresh your session.
You signed out in another tab or window.
Reload
to refresh your session.
You switched accounts on another tab or window.
Reload
to refresh your session.
Dismiss alert
{{ message }}
Uh oh!
There was an error while loading.
Please reload this page
.
vllm-project
/
vllm
Public
Uh oh!
There was an error while loading.
Please reload this page
.
Notifications
You must be signed in to change notification settings
Fork
21k
Star
89.7k
Code
Issues
2.2k
Pull requests
4.7k
Discussions
Actions
Projects
Security and quality
62
Insights
Additional navigation options
Code
Issues
Pull requests
Discussions
Actions
Projects
Security and quality
Insights
Actions: vllm-project/vllm
Actions
All workflows
Workflows
Add label on auto-merge enabled
Add label on auto-merge enabled
Buf
Buf
CI catalog
CI catalog
Close inactive issues and PRs
Close inactive issues and PRs
CodeQL
CodeQL
CodeQL Advanced
CodeQL Advanced
Copilot
Copilot
Copilot cloud agent
Copilot cloud agent
Copilot code review
Copilot code review
Dependabot Updates
Dependabot Updates
Show more workflows...
Management
Caches
All workflows
All workflows
Actions
Loading...
Loading
Sorry, something went wrong.
Uh oh!
There was an error while loading.
Please reload this page
.
will be ignored since log searching is not yet available
Showing runs from all workflows
will be ignored since log searching is not yet available
229,167 workflow run results
229,167 workflow run results
Workflow
Filter by Workflow
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching workflows.
Event
Filter by Event
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching events.
Status
Filter by Status
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching statuses.
Branch
Filter by Branch
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching branches.
Actor
Filter by Actor
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching users.
[Bug]: AriaForConditionalGeneration produces incoherent garbage output at tensor-parallel size > 1 (reproduces on CPU, no GPU/accelerator needed)
Label issues based on keywords
#14453:
Issue
#49122
opened by
wagnerpatriota
5s
5s
View workflow file
Exp/hisparse routed experts
pre-commit
#170242:
Pull request
#49121
labeled by
mergify
Bot
12s
S1ro1:exp/hisparse-routed-experts
S1ro1:exp/hisparse-routed-experts
12s
View #49121
View workflow file
Exp/hisparse routed experts
pre-commit
#170240:
Pull request
#49121
labeled by
mergify
Bot
2s
S1ro1:exp/hisparse-routed-experts
S1ro1:exp/hisparse-routed-experts
2s
View #49121
View workflow file
Exp/hisparse routed experts
pre-commit
#170241:
Pull request
#49121
labeled by
mergify
Bot
3s
S1ro1:exp/hisparse-routed-experts
S1ro1:exp/hisparse-routed-experts
3s
View #49121
View workflow file
Exp/hisparse routed experts
pre-commit
#170239:
Pull request
#49121
opened by
S1ro1
7s
S1ro1:exp/hisparse-routed-experts
S1ro1:exp/hisparse-routed-experts
7s
View #49121
View workflow file
Exp/hisparse routed experts
New PR Bot
#8993:
Pull request
#49121
opened by
S1ro1
10s
10s
View #49121
View workflow file
[CI/Build][BugFix][The Rock][AMD] Add spawn method in vision examples…
pre-commit
#170238:
Commit
ace9fda
pushed by
AndreasKaratzas
7m 24s
main
main
7m 24s
View workflow file
Push on main
CodeQL
#21953:
by
AndreasKaratzas
6m 39s
main
main
6m 39s
[Model] Honor fp32 head_dtype in Inkling muP logits (target + MTP draft)
New PR Bot
#8992:
Pull request
#49120
opened by
KKothuri
6s
6s
View #49120
View workflow file
[Model] Honor fp32 head_dtype in Inkling muP logits (target + MTP draft)
pre-commit
#170235:
Pull request
#49120
opened by
KKothuri
6s
KKothuri:fix/inkling-fp32-head-dtype
KKothuri:fix/inkling-fp32-head-dtype
6s
View #49120
View workflow file
fix(spec_decode): bypass embedding dim check for MTP speculative decoding methods
pre-commit
#169925:
Pull request
#49036
synchronize by
ArjunPakhan
6s
ArjunPakhan:fix/gemma4-mtp-v1-reduction-dim
ArjunPakhan:fix/gemma4-mtp-v1-reduction-dim
6s
View #49036
View workflow file
fix(spec_decode): bypass embedding dim check for MTP speculative decoding methods
pre-commit
#169924:
Pull request
#49036
labeled by
mergify
Bot
9s
ArjunPakhan:fix/gemma4-mtp-v1-reduction-dim
ArjunPakhan:fix/gemma4-mtp-v1-reduction-dim
9s
View #49036
View workflow file
fix(spec_decode): bypass embedding dim check for MTP speculative decoding methods
pre-commit
#169923:
Pull request
#49036
labeled by
mergify
Bot
2s
ArjunPakhan:fix/gemma4-mtp-v1-reduction-dim
ArjunPakhan:fix/gemma4-mtp-v1-reduction-dim
2s
View #49036
View workflow file
fix(spec_decode): bypass embedding dim check for MTP speculative decoding methods
pre-commit
#169922:
Pull request
#49036
opened by
ArjunPakhan
6s
ArjunPakhan:fix/gemma4-mtp-v1-reduction-dim
ArjunPakhan:fix/gemma4-mtp-v1-reduction-dim
6s
View #49036
View workflow file
fix(spec_decode): bypass embedding dim check for MTP speculative decoding methods
New PR Bot
#8929:
Pull request
#49036
opened by
ArjunPakhan
9s
9s
View #49036
View workflow file
Push on main
CodeQL
#21940:
by
vllm-bot
5m 39s
main
main
5m 39s
[Bugfix] Bump tml-fa4 for cutlass-dsl 4.6 API compatibility (#48988)
pre-commit
#169921:
Commit
c7ce03b
pushed by
vllm-bot
6m 20s
main
main
6m 20s
View workflow file
fix: handle missing parent modules in _has_module
New PR Bot
#8928:
Pull request
#49035
opened by
ShuhaoZhangTony
6s
6s
View #49035
View workflow file
fix: handle missing parent modules in _has_module
pre-commit
#169920:
Pull request
#49035
opened by
ShuhaoZhangTony
6s
ShuhaoZhangTony:feature/has-module-missing-parent
ShuhaoZhangTony:feature/has-module-missing-parent
6s
View #49035
View workflow file
fix(v1): avoid false shutdown failures on clean exit
pre-commit
#169919:
Pull request
#49034
labeled by
mergify
Bot
5s
vLLM-HUST:feature/v1-clean-shutdown
vLLM-HUST:feature/v1-clean-shutdown
5s
View #49034
View workflow file
fix(v1): avoid false shutdown failures on clean exit
New PR Bot
#8927:
Pull request
#49034
opened by
ShuhaoZhangTony
9s
9s
View #49034
View workflow file
fix(v1): avoid false shutdown failures on clean exit
pre-commit
#169918:
Pull request
#49034
opened by
ShuhaoZhangTony
9s
vLLM-HUST:feature/v1-clean-shutdown
vLLM-HUST:feature/v1-clean-shutdown
9s
View #49034
View workflow file
[Do not merge!] [Build] Migrate vendored DeepGEMM from pybind to TORCH_LIBRARY (abi3)
pre-commit
#169917:
Pull request
#48962
labeled by
Harry-Chen
5m 59s
cleonard530:deep_gemm_migration_to_torch_library
cleonard530:deep_gemm_migration_to_torch_library
5m 59s
View #48962
View workflow file
[Bugfix][Spec Decode] Reach TP consensus on the drafter-pass gate before launching drafter collectives (fixes rank-divergent wedge at max_model_len boundary)
pre-commit
#169916:
Pull request
#49027
synchronize by
marksunner
10s
marksunner:pr/spec-drafter-gate-tp-consensus
marksunner:pr/spec-drafter-gate-tp-consensus
10s
View #49027
View workflow file
[XPU] [MoE] add quant input when prepare for fusedmoe
pre-commit
#169915:
Pull request
#47122
synchronize by
mergify
Bot
11m 3s
zufangzhu:zufang/support_quant_act_moe
zufangzhu:zufang/support_quant_act_moe
11m 3s
View #47122
View workflow file
Previous
1
2
3
4
5
…
9166
9167
Next
You can’t perform that action at this time.