Skip to content

Pull requests: lightseekorg/tokenspeed

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

ci: update default GPU runners
#832 opened Jul 28, 2026 by lightseek-bot Contributor Draft
perf(kimi3): retile warp decode for M2 and M4
#831 opened Jul 28, 2026 by panditsa Contributor Draft
feat(spec): support dspark
#829 opened Jul 28, 2026 by minedec Contributor Draft
fix(runtime): keep multimodal tensors inline across nodes
#826 opened Jul 28, 2026 by tuanzhangCS Contributor Loading…
feat: wire flashinfer autotuner
#820 opened Jul 27, 2026 by syuoni Member Draft
deps: test TensorRT-LLM rc22 kernel wheel
#818 opened Jul 27, 2026 by Xiangyi1996 Collaborator Draft
Xpu qwen
#811 opened Jul 27, 2026 by zhenwei-intel Loading…
Jenga: two level allocation
#804 opened Jul 25, 2026 by wangbo981016 Contributor Draft
feat: support torchspec training
#798 opened Jul 25, 2026 by Dogacel Contributor Loading…
feat(kernel): support small-batch Gluon MLA decode
#793 opened Jul 24, 2026 by Max191 Contributor Loading…
feat(kernel): Add validation per family/mode for kernel registration
#769 opened Jul 22, 2026 by Max191 Contributor Loading…
Support MiniMax M3 CPU KVStore
#758 opened Jul 22, 2026 by FlamingoPg Contributor Draft
feat(lora): LoRA adapter serving
#738 opened Jul 20, 2026 by qywu Collaborator Loading…
feat(scheduler): per-adapter KV prefix-cache namespace + max_loras batch cap
#735 opened Jul 19, 2026 by qywu Collaborator Loading…
[wip] refactor(kernel): migrate GDN Triton kernels to tensor descriptors
#721 opened Jul 18, 2026 by raikonenfnu Contributor Loading…
4 tasks done
[WIP][AMD] Implement MTP support for qwen3.5 MXFP4
#720 opened Jul 18, 2026 by raikonenfnu Contributor Loading…
3 tasks
ProTip! no:milestone will show everything without a milestone.