Skip to content
Navigation Menu
Sign in
Appearance settings
Platform
AI CODE CREATION
GitHub Copilot
Write better code with AI
GitHub Copilot app
Direct agents from issue to merge
MCP Registry
Integrate external tools
DEVELOPER WORKFLOWS
Actions
Automate any workflow
Codespaces
Instant dev environments
Issues
Plan and track work
Code Review
Manage code changes
Code Quality
Enforce quality at merge
APPLICATION SECURITY
GitHub Advanced Security
Find and fix vulnerabilities
Code security
Secure your code as you build
Secret protection
Stop leaks before they start
EXPLORE
Why GitHub
Documentation
Blog
Changelog
Marketplace
View all features
Solutions
BY COMPANY SIZE
Enterprises
Small and medium teams
Startups
Nonprofits
BY USE CASE
App Modernization
DevSecOps
DevOps
CI/CD
View all use cases
BY INDUSTRY
Healthcare
Financial services
Manufacturing
Government
View all industries
View all solutions
Resources
EXPLORE BY TOPIC
AI
Software Development
DevOps
Security
View all topics
EXPLORE BY TYPE
Customer stories
Events & webinars
Ebooks & reports
Business insights
GitHub Skills
SUPPORT & SERVICES
Documentation
Customer support
Community forum
Trust center
Partners
View all resources
Open Source
COMMUNITY
GitHub Sponsors
Fund open source developers
PROGRAMS
Security Lab
Maintainer Community
GitHub Stars
Archive Program
REPOSITORIES
Topics
Trending
Collections
Enterprise
ENTERPRISE SOLUTIONS
Enterprise platform
AI-powered developer platform
AVAILABLE ADD-ONS
GitHub Advanced Security
Enterprise-grade security features
Copilot for Business
Enterprise-grade AI features
Premium Support
Enterprise-grade 24/7 support
Pricing
Search
/
Sign in
Sign up
Appearance settings
You signed in with another tab or window.
Reload
to refresh your session.
You signed out in another tab or window.
Reload
to refresh your session.
You switched accounts on another tab or window.
Reload
to refresh your session.
Dismiss alert
{{ message }}
aeon-x
vllm
Repository navigation
Code
Pull requests
1
(1)
Actions
Projects
Security and quality
Insights
More
items
Actions: aeon-x/vllm
Actions
All workflows
Workflows
Add label on auto-merge enabled
Add label on auto-merge enabled
Buf
Buf
Close inactive issues and PRs
Close inactive issues and PRs
Label issues based on keywords
Label issues based on keywords
New PR Bot
New PR Bot
Notify CI authorization
Notify CI authorization
PR title
PR title
pre-commit
pre-commit
Record CI approval
Record CI approval
Run CI from PR comment
Run CI from PR comment
Show more workflows...
Management
Caches
pre-commit
pre-commit
Actions
Loading...
Loading
Sorry, something went wrong.
Uh oh!
There was an error while loading.
Please reload this page
.
will be ignored since log searching is not yet available
Show workflow options
Create status badge
Create status badge
Loading
Uh oh!
There was an error while loading.
Please reload this page
.
pre-commit.yml
will be ignored since log searching is not yet available
50 workflow runs
50 workflow runs
Event
Filter by Event
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching events.
Status
Filter by Status
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching statuses.
Branch
Filter by Branch
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching branches.
Actor
Filter by Actor
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching users.
[Metrics] Expose cached prompt tokens by cache tier (#56318)
pre-commit
#50:
Commit
e37e51d
pushed by
aeon-x
2s
main
main
2s
View workflow file
[Bugfix][Spec Decode] Keep heterogeneous-vocab draft models on Model …
pre-commit
#49:
Commit
ad0f67a
pushed by
aeon-x
1s
main
main
1s
View workflow file
[Perf][Spec Decode] Avoid triton recompiles in the acceptance estimat…
pre-commit
#48:
Commit
3d5f4d4
pushed by
aeon-x
2s
main
main
2s
View workflow file
[Kimi-K3][AMD] Return KDA and MLA projection outputs directly (#50592)
pre-commit
#47:
Commit
382970e
pushed by
aeon-x
2s
main
main
2s
View workflow file
[Bugfix] NemotronHMTP: add hf_to_vllm_mapper so quant exclusions reac…
pre-commit
#46:
Commit
85c1365
pushed by
aeon-x
1d 23h 16m 3s
main
main
1d 23h 16m 3s
View workflow file
[Bugfix][ROCm] Preserve AITER unified-attention metadata during graph…
pre-commit
#45:
Commit
76ff0cd
pushed by
aeon-x
1d 4h 37m 7s
main
main
1d 4h 37m 7s
View workflow file
[Bugfix][CPU] Fix several bugs (#54042)
pre-commit
#44:
Commit
fdbf2dd
pushed by
aeon-x
6h 38m 57s
main
main
6h 38m 57s
View workflow file
[Bugfix][MLA] Fix BLHNC addressing for FlashInfer sparse MLA (#54465)
pre-commit
#43:
Commit
5707355
pushed by
aeon-x
5h 59m 21s
main
main
5h 59m 21s
View workflow file
[Bugfix][Spec Decode] Keep default CUDA graph sizes memory-safe (#54418)
pre-commit
#42:
Commit
b2dc864
pushed by
aeon-x
6h 22m 26s
main
main
6h 22m 26s
View workflow file
[Bugfix][Kernel] Keep packed GDN decode beta in FP32 (#53877)
pre-commit
#41:
Commit
56058fd
pushed by
aeon-x
1d 2h 52m 56s
main
main
1d 2h 52m 56s
View workflow file
[CI] Add explicit step keys to 18 hardware test steps (#54330)
pre-commit
#40:
Commit
8fa4c6c
pushed by
aeon-x
8h 23m 9s
main
main
8h 23m 9s
View workflow file
[CI][Ray] Fix flaky multi-node assignment test after placement-group …
pre-commit
#39:
Commit
680e217
pushed by
aeon-x
6h 0m 1s
main
main
6h 0m 1s
View workflow file
[CI][Test] Deflake test_mem.py sleep-mode asserts via allocator booke…
pre-commit
#38:
Commit
6cddad4
pushed by
aeon-x
6h 0m 3s
main
main
6h 0m 3s
View workflow file
[Mypy] Fix typing for J models (#54130)
pre-commit
#37:
Commit
fd5d3ae
pushed by
aeon-x
1d 4h 59m 13s
main
main
1d 4h 59m 13s
View workflow file
[ROCm][CI] Stage E gating (#50920)
pre-commit
#36:
Commit
cacc429
pushed by
aeon-x
6h 10m 11s
main
main
6h 10m 11s
View workflow file
[Rust Frontend] Take the raw buffer in mm tensor lowering when possib…
pre-commit
#35:
Commit
4c6c9d5
pushed by
aeon-x
6h 6m 48s
main
main
6h 6m 48s
View workflow file
[BugFix] Bind RayExecutorV2 TCPStore before publishing its port (#50969)
pre-commit
#34:
Commit
67e86d1
pushed by
aeon-x
6h 0m 1s
main
main
6h 0m 1s
View workflow file
[Bugfix] Restore multimodal support on the plain "vllm" throughput ba…
pre-commit
#33:
Commit
d1922cb
pushed by
aeon-x
1d 5h 15m 49s
main
main
1d 5h 15m 49s
View workflow file
[ROCm][CI] Keep startup profiling from aborting when free memory grow…
pre-commit
#32:
Commit
9236159
pushed by
aeon-x
6h 0m 14s
main
main
6h 0m 14s
View workflow file
[Kimi-K3] Merge MLA gate into QKV-A projection (#54015)
pre-commit
#31:
Commit
6ec92bc
pushed by
aeon-x
5h 59m 51s
main
main
5h 59m 51s
View workflow file
[Rust Frontend] Align OpenAI request and response edge cases (#53218)
pre-commit
#30:
Commit
9db222c
pushed by
aeon-x
5h 59m 56s
main
main
5h 59m 56s
View workflow file
[Bugfix] Update FlashMLA for sparse decode workspace fix (#53755)
pre-commit
#29:
Commit
75dea9b
pushed by
aeon-x
1d 5h 13m 0s
main
main
1d 5h 13m 0s
View workflow file
[Tools][Recipes] Improve sweep recommendations and short-alias parsin…
pre-commit
#28:
Commit
79651d6
pushed by
aeon-x
6h 2m 56s
main
main
6h 2m 56s
View workflow file
[Bugfix] Release CUDA graph profiling memory before KV cache allocati…
pre-commit
#27:
Commit
94d96e2
pushed by
aeon-x
5h 56m 28s
main
main
5h 56m 28s
View workflow file
[Model] Remove unused DeepSeek V4 top-k buffer helper (#53697)
pre-commit
#26:
Commit
0a5ad6f
pushed by
aeon-x
6h 3m 4s
main
main
6h 3m 4s
View workflow file
Previous
1
2
Next
You can’t perform that action at this time.