Skip to content

Pull requests: lightseekorg/tokenspeed

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

fix: clear padded breakable graph handoffs
#741 opened Jul 20, 2026 by FlamingoPg Contributor Loading…
feat(lora): LoRA adapter serving
#738 opened Jul 20, 2026 by qywu Collaborator Loading…
feat(scheduler): per-adapter KV prefix-cache namespace + max_loras batch cap
#735 opened Jul 19, 2026 by qywu Collaborator Loading…
feat: add basic MiniMax M3 support
#733 opened Jul 19, 2026 by FlamingoPg Contributor Loading…
[wip] refactor(kernel): migrate GDN Triton kernels to tensor descriptors
#721 opened Jul 18, 2026 by raikonenfnu Contributor Loading…
4 tasks done
[WIP][AMD] Implement MTP support for qwen3.5 MXFP4
#720 opened Jul 18, 2026 by raikonenfnu Contributor Loading…
3 tasks
feat(rl): stamp weight versions on generation responses
#672 opened Jul 14, 2026 by HJSang Collaborator Loading…
5 tasks done
feat(deepseek-v4): enable SM120 serving
#648 opened Jul 11, 2026 by lucifer1004 Contributor Draft
feat(kernel): add SM120 FlashInfer MXFP4 MoE
#645 opened Jul 11, 2026 by lucifer1004 Contributor Loading…
[WIP] Try new Triton package
#631 opened Jul 10, 2026 by antiagainst Member Loading…
feat: add DeepSeek V4 L2 KV cache offload and perf optimize.
#620 opened Jul 9, 2026 by SimonCqk Contributor Loading…
fix(spec): restore sentinel padding + context clamps for every drafter
#618 opened Jul 9, 2026 by jasl Contributor Loading…
ProTip! Filter pull requests by the default branch with base:main.