Skip to content

Pull requests: InternLM/lmdeploy

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

refactor(pytorch): derive CUDA step metadata from selected operators
#4805 opened Jul 30, 2026 by grimoire Collaborator Loading…
refactor: stream tool parameters incrementally improvement
#4802 opened Jul 29, 2026 by lvhan028 Collaborator Loading…
Ssm prefix cache non aligned
#4799 opened Jul 29, 2026 by grimoire Collaborator Draft
2 tasks
refactor: report cache usage directly improvement
#4798 opened Jul 29, 2026 by lvhan028 Collaborator Loading…
refactor: split api server endpoints improvement
#4797 opened Jul 28, 2026 by lvhan028 Collaborator Loading…
feat: add Rust TurboMind API server
#4794 opened Jul 28, 2026 by lvhan028 Collaborator Draft
[Feat]: Support output input logprobs enhancement New feature or request
#4793 opened Jul 28, 2026 by RunningLeon Collaborator Loading…
[Bugfix] Fix PyTorch H2D input lifetime across CUDA streams Bug:P1
#4792 opened Jul 28, 2026 by grimoire Collaborator Loading…
bump version to v0.15.0
#4791 opened Jul 28, 2026 by lvhan028 Collaborator Loading…
[WIP]: Support dflash for qwen3.5
#4789 opened Jul 27, 2026 by RunningLeon Collaborator Draft
optimize and modularize SSM prefix caching improvement
#4788 opened Jul 27, 2026 by grimoire Collaborator Loading…
Integrate DeepEPv2 enhancement New feature or request
#4783 opened Jul 25, 2026 by irexyc Collaborator Loading…
feat: support Intern-S2-Preview TS forecaster enhancement New feature or request
#4780 opened Jul 24, 2026 by CUHKSZzxy Collaborator Loading…
TEST: update turbomind qwen3.5 config
#4778 opened Jul 24, 2026 by littlegy Contributor Loading…
Upgrade to cu130
#4753 opened Jul 15, 2026 by RunningLeon Collaborator Loading…
feat: support GLM-5.2
#4737 opened Jul 7, 2026 by CUHKSZzxy Collaborator Loading…
Add TurboMind ViT support for InternVL and Qwen VL models enhancement New feature or request
#4719 opened Jun 29, 2026 by irexyc Collaborator Loading…
refactor: rename quant policy to kv cache dtype
#4718 opened Jun 29, 2026 by CUHKSZzxy Collaborator Draft
refactor: rename vl package to multimodal BC-breaking
#4710 opened Jun 26, 2026 by CUHKSZzxy Collaborator Draft
feat: share multimodal hash helpers enhancement New feature or request
#4704 opened Jun 24, 2026 by CUHKSZzxy Collaborator Loading…
ProTip! Exclude everything labeled bug with -label:bug.