Skip to content

Pull requests: InternLM/lmdeploy

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

refactor: split api server endpoints improvement
#4797 opened Jul 28, 2026 by lvhan028 Collaborator Loading…
feat: add Rust TurboMind API server
#4794 opened Jul 28, 2026 by lvhan028 Collaborator Draft
[Feat]: Support output input logprobs enhancement New feature or request
#4793 opened Jul 28, 2026 by RunningLeon Collaborator Loading…
[Bugfix] Fix PyTorch H2D input lifetime across CUDA streams Bug:P1
#4792 opened Jul 28, 2026 by grimoire Collaborator Loading…
bump version to v0.15.0
#4791 opened Jul 28, 2026 by lvhan028 Collaborator Loading…
[WIP]: Support dflash for qwen3.5
#4789 opened Jul 27, 2026 by RunningLeon Collaborator Draft
optimize and modularize SSM prefix caching improvement
#4788 opened Jul 27, 2026 by grimoire Collaborator Loading…
Integrate DeepEPv2 enhancement New feature or request
#4783 opened Jul 25, 2026 by irexyc Collaborator Loading…
feat: support Intern-S2-Preview TS forecaster enhancement New feature or request
#4780 opened Jul 24, 2026 by CUHKSZzxy Collaborator Loading…
TEST: update turbomind qwen3.5 config
#4778 opened Jul 24, 2026 by littlegy Contributor Loading…
Upgrade to cu130
#4753 opened Jul 15, 2026 by RunningLeon Collaborator Loading…
feat: support GLM-5.2
#4737 opened Jul 7, 2026 by CUHKSZzxy Collaborator Loading…
refactor: isolate XML tool parser state
#4736 opened Jul 6, 2026 by lvhan028 Collaborator Loading…
Add TurboMind ViT support for InternVL and Qwen VL models
#4719 opened Jun 29, 2026 by irexyc Collaborator Loading…
refactor: rename quant policy to kv cache dtype
#4718 opened Jun 29, 2026 by CUHKSZzxy Collaborator Draft
refactor: rename vl package to multimodal BC-breaking
#4710 opened Jun 26, 2026 by CUHKSZzxy Collaborator Draft
feat: share multimodal hash helpers enhancement New feature or request
#4704 opened Jun 24, 2026 by CUHKSZzxy Collaborator Loading…
Fix stale prefix cache hit rate metric
#4699 opened Jun 22, 2026 by DavinciEvans Loading…
fix: gate multimodal preprocessing concurrency
#4687 opened Jun 17, 2026 by CUHKSZzxy Collaborator Loading…
Batch invariant support PART1
#4666 opened Jun 10, 2026 by grimoire Collaborator Draft
ProTip! Follow long discussions with comments:>50.