Skip to content

Pull requests: lightseekorg/tokenspeed

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

perf: fuse MiniMax sparse cache insertion
#839 opened Jul 29, 2026 by FlamingoPg Contributor Draft
perf(k3): optimize decode
#834 opened Jul 29, 2026 by nperrin-fr Collaborator Draft
[WIP] perf(kimi3): retile warp decode
#831 opened Jul 28, 2026 by panditsa Contributor Loading…
feat(dspark): Add DSpark support
#829 opened Jul 28, 2026 by minedec Contributor Draft
feat: wire flashinfer autotuner
#820 opened Jul 27, 2026 by syuoni Member Draft
deps: test TensorRT-LLM rc22 kernel wheel
#818 opened Jul 27, 2026 by Xiangyi1996 Collaborator Draft
feat: jenga two level allocation
#804 opened Jul 25, 2026 by wangbo981016 Contributor Loading…
feat: support torchspec training
#798 opened Jul 25, 2026 by Dogacel Contributor Loading…
feat(kernel): support small-batch Gluon MLA decode
#793 opened Jul 24, 2026 by Max191 Contributor Loading…
feat(kernel): Add validation per family/mode for kernel registration
#769 opened Jul 22, 2026 by Max191 Contributor Loading…
Support MiniMax M3 CPU KVStore
#758 opened Jul 22, 2026 by FlamingoPg Contributor Draft
feat(lora): LoRA adapter serving
#738 opened Jul 20, 2026 by qywu Collaborator Loading…
feat(scheduler): per-adapter KV prefix-cache namespace + max_loras batch cap
#735 opened Jul 19, 2026 by qywu Collaborator Loading…
[wip] refactor(kernel): migrate GDN Triton kernels to tensor descriptors
#721 opened Jul 18, 2026 by raikonenfnu Contributor Loading…
4 tasks done
[WIP][AMD] Implement MTP support for qwen3.5 MXFP4
#720 opened Jul 18, 2026 by raikonenfnu Contributor Loading…
3 tasks
ProTip! Updated in the last three days: updated:>2026-07-26.