Skip to content
Navigation Menu
Sign in
Appearance settings
Platform
AI CODE CREATION
GitHub Copilot
Write better code with AI
GitHub Copilot app
Direct agents from issue to merge
MCP Registry
Integrate external tools
DEVELOPER WORKFLOWS
Actions
Automate any workflow
Codespaces
Instant dev environments
Issues
Plan and track work
Code Review
Manage code changes
Code Quality
Enforce quality at merge
APPLICATION SECURITY
GitHub Advanced Security
Find and fix vulnerabilities
Code security
Secure your code as you build
Secret protection
Stop leaks before they start
EXPLORE
Why GitHub
Documentation
Blog
Changelog
Marketplace
View all features
Solutions
BY COMPANY SIZE
Enterprises
Small and medium teams
Startups
Nonprofits
BY USE CASE
App Modernization
DevSecOps
DevOps
CI/CD
View all use cases
BY INDUSTRY
Healthcare
Financial services
Manufacturing
Government
View all industries
View all solutions
Resources
EXPLORE BY TOPIC
AI
Software Development
DevOps
Security
View all topics
EXPLORE BY TYPE
Customer stories
Events & webinars
Ebooks & reports
Business insights
GitHub Skills
SUPPORT & SERVICES
Documentation
Customer support
Community forum
Trust center
Partners
View all resources
Open Source
COMMUNITY
GitHub Sponsors
Fund open source developers
PROGRAMS
Security Lab
Maintainer Community
GitHub Stars
Archive Program
REPOSITORIES
Topics
Trending
Collections
Enterprise
ENTERPRISE SOLUTIONS
Enterprise platform
AI-powered developer platform
AVAILABLE ADD-ONS
GitHub Advanced Security
Enterprise-grade security features
Copilot for Business
Enterprise-grade AI features
Premium Support
Enterprise-grade 24/7 support
Pricing
Search
/
Sign in
Sign up
Appearance settings
You signed in with another tab or window.
Reload
to refresh your session.
You signed out in another tab or window.
Reload
to refresh your session.
You switched accounts on another tab or window.
Reload
to refresh your session.
Dismiss alert
{{ message }}
0xjunhao
/
vllm
Public
forked from
vllm-project/vllm
Notifications
You must be signed in to change notification settings
Fork
0
Star
1
Code
Pull requests
0
Actions
Projects
Security and quality
0
Insights
Additional navigation options
Code
Pull requests
Actions
Projects
Security and quality
Insights
Actions: 0xjunhao/vllm
Actions
All workflows
Workflows
Add label on auto-merge enabled
Add label on auto-merge enabled
Buf
Buf
Close inactive issues and PRs
Close inactive issues and PRs
Label issues based on keywords
Label issues based on keywords
New PR Bot
New PR Bot
Notify CI authorization
Notify CI authorization
PR title
PR title
pre-commit
pre-commit
Record CI approval
Record CI approval
Run CI from PR comment
Run CI from PR comment
Show more workflows...
Management
Caches
pre-commit
pre-commit
Actions
Loading...
Loading
Sorry, something went wrong.
Uh oh!
There was an error while loading.
Please reload this page
.
will be ignored since log searching is not yet available
Show workflow options
Create status badge
Create status badge
Loading
Uh oh!
There was an error while loading.
Please reload this page
.
pre-commit.yml
will be ignored since log searching is not yet available
18 workflow runs
18 workflow runs
Event
Filter by Event
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching events.
Status
Filter by Status
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching statuses.
Branch
Filter by Branch
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching branches.
Actor
Filter by Actor
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching users.
[Bugfix][Spec Decode] Honour the draft's attention_backend on Model R…
pre-commit
#30:
Commit
2524051
pushed by
0xjunhao
Skipped
main
main
Skipped
View workflow file
[Bugfix] Thread kv_transfer_params into engine for /inference/v1/gene…
pre-commit
#29:
Commit
c71f6f8
pushed by
0xjunhao
1d 0h 0m 2s
main
main
1d 0h 0m 2s
View workflow file
[Frontend] Preserve bare Inkling text in Python and Rust parsers (#50…
pre-commit
#28:
Commit
629a938
pushed by
0xjunhao
1d 0h 0m 2s
main
main
1d 0h 0m 2s
View workflow file
[XPU][CI]Adjust timeout_in_minutes in Intel GPU CI (#48418)
pre-commit
#27:
Commit
93e3bc8
pushed by
0xjunhao
1d 0h 0m 1s
main
main
1d 0h 0m 1s
View workflow file
Upgrade tpu-inference to v0.19.0 (#41844)
pre-commit
#26:
Commit
ca3e62d
pushed by
0xjunhao
5m 16s
main
main
5m 16s
View workflow file
[NVFP4] NVFP4 MOE emulation fallback for H100/MI300/MI350, standardiz…
pre-commit
#25:
Commit
d622e27
pushed by
0xjunhao
4m 8s
main
main
4m 8s
View workflow file
[CI/Build] Improve stability of CPU tests (#39966)
pre-commit
#24:
Commit
324a3d2
pushed by
0xjunhao
5m 31s
main
main
5m 31s
View workflow file
[Model] Sync upstream BT=chunk_size fix for GDN chunk_fwd_kernel_o, s…
pre-commit
#23:
Commit
b779eb3
pushed by
0xjunhao
4m 30s
main
main
4m 30s
View workflow file
[Async][Spec Decoding] Zero-bubble async scheduling + spec decoding (…
pre-commit
#22:
Commit
fafe76b
pushed by
0xjunhao
3m 13s
main
main
3m 13s
View workflow file
[Bugfix] Fix ROCm crash in qwen3_next multi-stream events (#36795) (#…
pre-commit
#21:
Commit
a913b61
pushed by
0xjunhao
4m 41s
main
main
4m 41s
View workflow file
[UX] Add vLLM model inspection view (#29450)
pre-commit
#20:
Commit
d5ec6c0
pushed by
0xjunhao
3m 30s
main
main
3m 30s
View workflow file
[ROCM][CI] Fix AMD Examples Test Group (#30276)
pre-commit
#19:
Commit
2cc5aff
pushed by
0xjunhao
4m 42s
main
main
4m 42s
View workflow file
Re-enable FlashInfer for Llama4 on Blackwell in e2e fusion tests (#28…
pre-commit
#18:
Commit
61728cd
pushed by
0xjunhao
4m 39s
main
main
4m 39s
View workflow file
[Bug] Fix DeepEP low latency `assert self.batched_router_logits.size(…
pre-commit
#17:
Commit
fcb1d57
pushed by
0xjunhao
4m 29s
main
main
4m 29s
View workflow file
Use UV_LINK_MODE=copy in Dockerfile to avoid hardlink fail (#22128)
pre-commit
#16:
Commit
c494f96
pushed by
0xjunhao
7m 3s
main
main
7m 3s
View workflow file
[Bugfix][V1][P/D]Fix the uneven polling issue in the toy proxy for P2…
pre-commit
#15:
Commit
c09efff
pushed by
0xjunhao
6m 45s
main
main
6m 45s
View workflow file
[Bug] Update auto_tune.sh to separate benchmarking and profiling. (#2…
pre-commit
#14:
Commit
309c1bb
pushed by
0xjunhao
6m 39s
main
main
6m 39s
View workflow file
feat: Add Support GPTQ Quantization MOE on ROCM vllm serve (#21733)
pre-commit
#13:
Commit
3654847
pushed by
0xjunhao
7m 33s
main
main
7m 33s
View workflow file
You can’t perform that action at this time.