Skip to content
Navigation Menu
Sign in
Appearance settings
Platform
AI CODE CREATION
GitHub Copilot
Write better code with AI
GitHub Copilot app
Direct agents from issue to merge
MCP Registry
Integrate external tools
DEVELOPER WORKFLOWS
Actions
Automate any workflow
Codespaces
Instant dev environments
Issues
Plan and track work
Code Review
Manage code changes
Code Quality
Enforce quality at merge
APPLICATION SECURITY
GitHub Advanced Security
Find and fix vulnerabilities
Code security
Secure your code as you build
Secret protection
Stop leaks before they start
EXPLORE
Why GitHub
Documentation
Blog
Changelog
Marketplace
View all features
Solutions
BY COMPANY SIZE
Enterprises
Small and medium teams
Startups
Nonprofits
BY USE CASE
App Modernization
DevSecOps
DevOps
CI/CD
View all use cases
BY INDUSTRY
Healthcare
Financial services
Manufacturing
Government
View all industries
View all solutions
Resources
EXPLORE BY TOPIC
AI
Software Development
DevOps
Security
View all topics
EXPLORE BY TYPE
Customer stories
Events & webinars
Ebooks & reports
Business insights
GitHub Skills
SUPPORT & SERVICES
Documentation
Customer support
Community forum
Trust center
Partners
View all resources
Open Source
COMMUNITY
GitHub Sponsors
Fund open source developers
PROGRAMS
Security Lab
Maintainer Community
GitHub Stars
Archive Program
REPOSITORIES
Topics
Trending
Collections
Enterprise
ENTERPRISE SOLUTIONS
Enterprise platform
AI-powered developer platform
AVAILABLE ADD-ONS
GitHub Advanced Security
Enterprise-grade security features
Copilot for Business
Enterprise-grade AI features
Premium Support
Enterprise-grade 24/7 support
Pricing
Search
/
Sign in
Sign up
Appearance settings
You signed in with another tab or window.
Reload
to refresh your session.
You signed out in another tab or window.
Reload
to refresh your session.
You switched accounts on another tab or window.
Reload
to refresh your session.
Dismiss alert
{{ message }}
llsj14
/
vllm
Public
forked from
vllm-project/vllm
Notifications
You must be signed in to change notification settings
Fork
0
Star
1
Code
Issues
1
Pull requests
0
Actions
Projects
Security and quality
0
Insights
Additional navigation options
Code
Issues
Pull requests
Actions
Projects
Security and quality
Insights
Actions: llsj14/vllm
Actions
All workflows
Workflows
Add label on auto-merge enabled
Add label on auto-merge enabled
Cleanup PR Body
Cleanup PR Body
Label issues based on keywords
Label issues based on keywords
macOS Apple Silicon Smoke Test
macOS Apple Silicon Smoke Test
PR Reminder Comment Bot
PR Reminder Comment Bot
pre-commit
pre-commit
Close inactive issues and PRs
Close inactive issues and PRs
Disabled
Show more workflows...
Management
Caches
pre-commit
pre-commit
Actions
Loading...
Loading
Sorry, something went wrong.
Uh oh!
There was an error while loading.
Please reload this page
.
will be ignored since log searching is not yet available
Show workflow options
Create status badge
Create status badge
Loading
Uh oh!
There was an error while loading.
Please reload this page
.
pre-commit.yml
will be ignored since log searching is not yet available
15 workflow runs
15 workflow runs
Event
Filter by Event
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching events.
Status
Filter by Status
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching statuses.
Branch
Filter by Branch
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching branches.
Actor
Filter by Actor
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching users.
[Model] Add LoRA support for Whisper models (#29856)
pre-commit
#26:
Commit
3b23d57
pushed by
llsj14
4m 23s
main
main
4m 23s
View workflow file
fix: set output token ids with async scheduling according to sampling params
pre-commit
#25:
Pull request
#2
opened by
llsj14
3m 37s
fix/require-output-tok-ids
fix/require-output-tok-ids
3m 37s
View #2
View workflow file
[Frontend] Add MCP tool streaming support to Responses API (#31761)
pre-commit
#24:
Commit
a4ec0c5
pushed by
llsj14
3m 31s
main
main
3m 31s
View workflow file
[CI][Bugfix] Fix token counting in chunked prefill compl test (#31630)
pre-commit
#23:
Commit
4f9ce35
pushed by
llsj14
3m 36s
main
main
3m 36s
View workflow file
add aarnphm and chaunceyjiang to the new tool_parser directory (#31088)
pre-commit
#22:
Commit
bb80f69
pushed by
llsj14
3m 33s
main
main
3m 33s
View workflow file
[Doc] Add documents for multi-node distributed serving with MP backen…
pre-commit
#21:
Commit
7c16f3f
pushed by
llsj14
4m 40s
main
main
4m 40s
View workflow file
[V0 deprecation] Remove VLLM_USE_V1 usage in platform and v1 module (…
pre-commit
#20:
Commit
30a14b0
pushed by
llsj14
4m 46s
main
main
4m 46s
View workflow file
Fix MiniMax-M2 rmsnorm precision and remove useless code (#27627)
pre-commit
#19:
Commit
d6704dd
pushed by
llsj14
4m 20s
main
main
4m 20s
View workflow file
[MISC] Rename the torch profiler filename as instance_id+rank_id for …
pre-commit
#18:
Commit
7685201
pushed by
llsj14
3m 3s
main
main
3m 3s
View workflow file
[PERF] [Qwen3-next] Speed up gated RMSNorm (#26207)
pre-commit
#17:
Commit
82e64c7
pushed by
llsj14
4m 39s
main
main
4m 39s
View workflow file
[BugFix] enable DOTALL to match multi-line tool_call parameters in ex…
pre-commit
#16:
Commit
2b85697
pushed by
llsj14
8m 7s
main
main
8m 7s
View workflow file
Revert "[PERF] Use faster way of decode in tokenizer: avoid useless l…
pre-commit
#15:
Commit
b4e9fd8
pushed by
llsj14
12m 26s
main
main
12m 26s
View workflow file
[Misc] Minor code cleanup for _get_prompt_logprobs_dict (#23064)
pre-commit
#14:
Commit
8ea0c27
pushed by
llsj14
7m 11s
main
main
7m 11s
View workflow file
[Bugfix] fix qwen3 moe fp8 accuracy issue (#23031)
pre-commit
#13:
Commit
a258ad8
pushed by
llsj14
6m 44s
main
main
6m 44s
View workflow file
[Frontend] Expose do_log_stats interval to env (#22905)
pre-commit
#12:
Commit
a0632a3
pushed by
llsj14
7m 46s
main
main
7m 46s
View workflow file
You can’t perform that action at this time.