forked from local-inference-lab/vllm
-
Notifications
You must be signed in to change notification settings - Fork 0
Pull requests: FujitsuPolycom/vllm
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
integration: validate GLM-5.2 sparse CKV stack on latest GG
verified
Validated by focused tests
#7
opened Jul 21, 2026 by
FujitsuPolycom
Owner
•
Draft
docs(deploy): add pinned GLM-5.2 sparse CKV integration release
verified
Validated by focused tests
#6
opened Jul 21, 2026 by
FujitsuPolycom
Owner
•
Draft
dcp: bulk-prefetch three shared CKV layers over copy-engine transport
verified
Validated by focused tests
#5
opened Jul 21, 2026 by
FujitsuPolycom
Owner
•
Draft
feat(dcp): add selected-record sparse CKV decode
verified
Validated by focused tests
#4
opened Jul 20, 2026 by
FujitsuPolycom
Owner
•
Draft
feat(dcp): add depth-N native CKV prefetch
verified
Validated by focused tests
#3
opened Jul 20, 2026 by
FujitsuPolycom
Owner
•
Draft
feat(dcp): replicate target sparse-indexer cache
verified
Validated by focused tests
#2
opened Jul 20, 2026 by
FujitsuPolycom
Owner
•
Draft
feat(kv-cache): allocate mixed MLA groups in lockstep
verified
Validated by focused tests
#1
opened Jul 20, 2026 by
FujitsuPolycom
Owner
•
Draft
ProTip!
Mix and match filters to narrow down what you’re looking for.