-
Notifications
You must be signed in to change notification settings - Fork 194
Pull requests: lightseekorg/tokenspeed
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
[AMD] Add initial support for pipelined GFX1250 Gluon MHA decode
#746
opened Jul 21, 2026 by
adityakankariya
Loading…
fix: clear padded breakable graph handoffs
#741
opened Jul 20, 2026 by
FlamingoPg
Contributor
Loading…
feat(profile): Link VizTracer to Proton CPU scopes
#740
opened Jul 20, 2026 by
antiagainst
Member
•
Draft
feat(scheduler): per-adapter KV prefix-cache namespace + max_loras batch cap
#735
opened Jul 19, 2026 by
qywu
Collaborator
Loading…
[wip] refactor(kernel): migrate GDN Triton kernels to tensor descriptors
#721
opened Jul 18, 2026 by
raikonenfnu
Contributor
Loading…
4 tasks done
[WIP][AMD] Implement MTP support for qwen3.5 MXFP4
#720
opened Jul 18, 2026 by
raikonenfnu
Contributor
Loading…
3 tasks
perf(deepseek-v4): capture DSA projections and cache writes into the prefill graph
#714
opened Jul 17, 2026 by
dongjiyingdjy
Contributor
Loading…
[WIP] feat(scheduler): live-tail allocation for sliding groups
#694
opened Jul 16, 2026 by
nperrin-fr
Collaborator
•
Draft
fix(scheduler): check host write-back capacity before mutating retrac…
#688
opened Jul 15, 2026 by
alexps9
Loading…
3 tasks done
feat(rl): stamp weight versions on generation responses
#672
opened Jul 14, 2026 by
HJSang
Collaborator
Loading…
5 tasks done
perf(deepseek-v4): select two-stage fused mHC launch by token count
#661
opened Jul 13, 2026 by
Xiangyi1996
Collaborator
•
Draft
fix(pd): correct DeepSeek V4 layerwise cache handoff
#660
opened Jul 13, 2026 by
lucifer1004
Contributor
•
Draft
feat(runtime): scheduler-driven full_refresh bit for page-table mirror
#651
opened Jul 12, 2026 by
raikonenfnu
Contributor
Loading…
feat(kernel): add SM120 FlashInfer MXFP4 MoE
#645
opened Jul 11, 2026 by
lucifer1004
Contributor
Loading…
fix(deepseek-v4): route large-token fused mHC around the allinone cliff
#622
opened Jul 9, 2026 by
Xiangyi1996
Collaborator
•
Draft
perf(deepseek-v4): widen sparse compress cache launch for large ratios
#621
opened Jul 9, 2026 by
Xiangyi1996
Collaborator
Loading…
feat: add DeepSeek V4 L2 KV cache offload and perf optimize.
#620
opened Jul 9, 2026 by
SimonCqk
Contributor
Loading…
fix(spec): restore sentinel padding + context clamps for every drafter
#618
opened Jul 9, 2026 by
jasl
Contributor
Loading…
Previous Next
ProTip!
Filter pull requests by the default branch with base:main.