-
Notifications
You must be signed in to change notification settings - Fork 717
Pull requests: InternLM/lmdeploy
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
refactor: split api server endpoints
improvement
#4797
opened Jul 28, 2026 by
lvhan028
Collaborator
Loading…
SM90 native BF16/FP8 GEMM kernels, fused-SiLU quantization, and linear test harness
improvement
#4795
opened Jul 28, 2026 by
lzhangzz
Collaborator
Loading…
[Feat]: Support output input logprobs
enhancement
New feature or request
#4793
opened Jul 28, 2026 by
RunningLeon
Collaborator
Loading…
[Bugfix] Fix PyTorch H2D input lifetime across CUDA streams
Bug:P1
#4792
opened Jul 28, 2026 by
grimoire
Collaborator
Loading…
optimize and modularize SSM prefix caching
improvement
#4788
opened Jul 27, 2026 by
grimoire
Collaborator
Loading…
Integrate DeepEPv2
enhancement
New feature or request
#4783
opened Jul 25, 2026 by
irexyc
Collaborator
Loading…
[Fix] generate: propagate the disconnect 400 instead of returning null
#4782
opened Jul 24, 2026 by
AmirF194
Loading…
feat: support Intern-S2-Preview TS forecaster
enhancement
New feature or request
#4780
opened Jul 24, 2026 by
CUHKSZzxy
Collaborator
Loading…
[Fix] completions_v1: abort session and return 400 on client disconnect
#4777
opened Jul 23, 2026 by
AmirF194
Loading…
fix(security): reject pickled update_weights payloads by default
#4765
opened Jul 21, 2026 by
Solaris-star
Loading…
[Draft] Add PyTorch engine support for Hy3 BF16 and static FP8 (NO MTP)
#4763
opened Jul 20, 2026 by
yidingcheng0206
Collaborator
•
Draft
5 of 7 tasks
Add TurboMind ViT support for InternVL and Qwen VL models
#4719
opened Jun 29, 2026 by
irexyc
Collaborator
Loading…
feat: share multimodal hash helpers
enhancement
New feature or request
#4704
opened Jun 24, 2026 by
CUHKSZzxy
Collaborator
Loading…
fix: gate multimodal preprocessing concurrency
#4687
opened Jun 17, 2026 by
CUHKSZzxy
Collaborator
Loading…
Previous Next
ProTip!
Follow long discussions with comments:>50.