Skip to content
Navigation Menu
Sign in
Appearance settings
Platform
AI CODE CREATION
GitHub Copilot
Write better code with AI
GitHub Copilot app
Direct agents from issue to merge
MCP Registry
Integrate external tools
DEVELOPER WORKFLOWS
Actions
Automate any workflow
Codespaces
Instant dev environments
Issues
Plan and track work
Code Review
Manage code changes
Code Quality
Enforce quality at merge
APPLICATION SECURITY
GitHub Advanced Security
Find and fix vulnerabilities
Code security
Secure your code as you build
Secret protection
Stop leaks before they start
EXPLORE
Why GitHub
Documentation
Blog
Changelog
Marketplace
View all features
Solutions
BY COMPANY SIZE
Enterprises
Small and medium teams
Startups
Nonprofits
BY USE CASE
App Modernization
DevSecOps
DevOps
CI/CD
View all use cases
BY INDUSTRY
Healthcare
Financial services
Manufacturing
Government
View all industries
View all solutions
Resources
EXPLORE BY TOPIC
AI
Software Development
DevOps
Security
View all topics
EXPLORE BY TYPE
Customer stories
Events & webinars
Ebooks & reports
Business insights
GitHub Skills
SUPPORT & SERVICES
Documentation
Customer support
Community forum
Trust center
Partners
View all resources
Open Source
COMMUNITY
GitHub Sponsors
Fund open source developers
PROGRAMS
Security Lab
Maintainer Community
GitHub Stars
Archive Program
REPOSITORIES
Topics
Trending
Collections
Enterprise
ENTERPRISE SOLUTIONS
Enterprise platform
AI-powered developer platform
AVAILABLE ADD-ONS
GitHub Advanced Security
Enterprise-grade security features
Copilot for Business
Enterprise-grade AI features
Premium Support
Enterprise-grade 24/7 support
Pricing
Search
/
Sign in
Sign up
Appearance settings
You signed in with another tab or window.
Reload
to refresh your session.
You signed out in another tab or window.
Reload
to refresh your session.
You switched accounts on another tab or window.
Reload
to refresh your session.
Dismiss alert
{{ message }}
baidu
/
vLLM-Kunlun
Public
Notifications
You must be signed in to change notification settings
Fork
103
Star
472
Code
Issues
71
Pull requests
31
Discussions
Actions
Projects
Wiki
Security and quality
0
Insights
Additional navigation options
Code
Issues
Pull requests
Discussions
Actions
Projects
Wiki
Security and quality
Insights
All pull requests
New pull request
Search pull requests
is
:
pr
state
:
open
is:pr state:open
Clear filter
Search
Pull requests
Open
31
(31)
Closed
328
(328)
Author
Label
Projects
Milestones
Reviews
Assignee
Sort by
Newest
descending
More items
Comfortable display density
Compact display density
[Feature] Enable MRV2 hybrid prefix caching and optimize metadata paths on Kunlun
enhancement
kunlun-xpu
needs review
tests
#492
·
czyu-nju
opened
Sep 24, 2026
·
·
[Feature] Route Model Runner V2 through xspeedgate_ops native ops
bug
kunlun-xpu
needs review
tests
#490
·
Joeegin
opened
Sep 20, 2026
Contributor
·
·
4
[Refactor] Register derived RoPE implementations through OOT
kunlun-xpu
needs review
#488
·
GrootLiu
opened
Sep 16, 2026
Collaborator
·
·
[WIP][Feature] Support MoE expert LoRA inference
kunlun-xpu
needs review
#487
·
loveleaves
opened
Sep 16, 2026
·
·
4
[Refactor] Unify fused MoE routing and execution pipeline
documentation
kunlun-xpu
needs review
tests
#486
·
GrootLiu
opened
Sep 16, 2026
Collaborator
·
·
1
Kimi k3
kunlun-xpu
needs review
tests
#482
·
baoqian426
opened
Sep 10, 2026
Contributor
·
·
2
MiniMax-M3 Support
enhancement
kunlun-xpu
needs review
#478
·
xyDong0223
opened
Sep 8, 2026
Collaborator
·
·
12
Fix/minimax 2 5 environment
#458
·
xyDong0223
opened
Sep 2, 2026
Collaborator
·
·
5
[Doc] Add automated OpenWiki updates
#456
·
xyDong0223
opened
Sep 2, 2026
Collaborator
·
·
4
feat(dsv4): support DeepSeek-V4-Flash inference on Kunlun XPU
#439
·
BoundlessWindMoon
opened
Aug 20, 2026
·
·
8
[Feature] Support DeepSeek-V4-Flash on Kunlun — FP8 + INT8 MoE Native Support
#402
·
BoundlessWindMoon
opened
Jul 29, 2026
·
·
5
feat: add MiniMax M3 + M2.7 support (Kunlun XPU)
#396
·
NaphJohn
opened
Jul 24, 2026
Contributor
·
·
4
[Feature] Support Qwen3.5 MoE MTP speculative decoding
#384
·
sdh1014
opened
Jul 5, 2026
·
·
4
[BUGFIX]int8_moe_backend hook
#381
·
Marshall-Ge
opened
Jun 26, 2026
Contributor
·
·
2
feat: upgrade vllm_kunlun to support vllm 0.23.0 qwen3-8b
#380
·
1916hcc
opened
Jun 26, 2026
Contributor
·
·
[Benchmark] Adapt attention and fused MoE benchmarks to vLLM-Kunlun
#364
·
KinChow
opened
May 26, 2026
·
·
7
[Bugfix] Force xgrammar to use torch_native backend on XPU
#353
·
GrootLiu
opened
May 8, 2026
Collaborator
·
·
4
[Doc] Align developer guide navigation titles
#350
·
Lidang-Jiang
opened
May 6, 2026
Contributor
·
·
[Doc] Clean up contributing guide navigation
#349
·
Lidang-Jiang
opened
May 6, 2026
Contributor
·
·
[Doc] Add model source guidance
#348
·
Lidang-Jiang
opened
May 6, 2026
Contributor
·
·
[Doc] Replace placeholder Docker image tags
#347
·
Lidang-Jiang
opened
May 6, 2026
Contributor
·
·
[Doc] Update documentation copyright year
#346
·
Lidang-Jiang
opened
May 6, 2026
Contributor
·
·
[Feature] W4A16 support for MoE models
#345
·
Qeeweew
opened
May 6, 2026
·
·
[Doc] Align environment variable docs with runtime envs
#342
·
Lidang-Jiang
opened
Apr 28, 2026
Contributor
·
·
[Feature] Add DeepSeek V3.2 W8A8 INT8 model support
#339
·
Lidang-Jiang
opened
Apr 24, 2026
Contributor
·
·
Previous
1
2
Next
You can’t perform that action at this time.