Skip to content

Pull requests: lightseekorg/tokenspeed

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

fix(multi-node): harden Qwen3.5 multi-node execution
#780 opened Jul 23, 2026 by tuanzhangCS Contributor Loading…
fix(qwen3.5): allow MTP to be quantized
#777 opened Jul 23, 2026 by minedec Contributor Loading…
fix: Support NVIDIA Nemotron-3-Super-120B-A12B
#770 opened Jul 22, 2026 by lian125537 Loading…
feat(kernel): Add validation per family/mode for kernel registration
#769 opened Jul 22, 2026 by Max191 Contributor Loading…
Support MiniMax M3 CPU KVStore
#758 opened Jul 22, 2026 by FlamingoPg Contributor Draft
feat(lora): LoRA adapter serving
#738 opened Jul 20, 2026 by qywu Collaborator Loading…
feat(scheduler): per-adapter KV prefix-cache namespace + max_loras batch cap
#735 opened Jul 19, 2026 by qywu Collaborator Loading…
[wip] refactor(kernel): migrate GDN Triton kernels to tensor descriptors
#721 opened Jul 18, 2026 by raikonenfnu Contributor Loading…
4 tasks done
[WIP][AMD] Implement MTP support for qwen3.5 MXFP4
#720 opened Jul 18, 2026 by raikonenfnu Contributor Loading…
3 tasks
feat(rl): stamp weight versions on generation responses
#672 opened Jul 14, 2026 by HJSang Collaborator Loading…
5 tasks done
feat(deepseek-v4): enable SM120 serving
#648 opened Jul 11, 2026 by lucifer1004 Contributor Draft
feat(kernel): add SM120 FlashInfer MXFP4 MoE
#645 opened Jul 11, 2026 by lucifer1004 Contributor Loading…
ProTip! What’s not been updated in a month: updated:<2026-06-23.