Skip to content

Pull requests: lightseekorg/tokenspeed

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

feat(amd): add initial GFX1250 Gluon MLA kernels
#877 opened Jul 31, 2026 by Yu-Zhewen Contributor Loading…
feat: improve multi-node usability
#873 opened Jul 31, 2026 by syuoni Member Loading…
feat(DeepEP): support fp8 deepep path
#872 opened Jul 31, 2026 by tuanzhangCS Contributor Loading…
deps: bump TRTLLM kernel to 20260731
#870 opened Jul 31, 2026 by Xiangyi1996 Collaborator Draft
feat: add Qwen2.5-MoE model support
#866 opened Jul 31, 2026 by potoior Loading…
Remove Radix cache path and consolidate runtime cache handling
#864 opened Jul 31, 2026 by wangbo981016 Contributor Loading…
Support detailed token logprobs for RL and distillation
#861 opened Jul 31, 2026 by HJSang Collaborator Loading…
fix(runtime): reject ambiguous prompt sources
#854 opened Jul 30, 2026 by ShiroKSH Loading…
perf(kimi-k3): fuse AMD native MoE up projection epilogue
#841 opened Jul 29, 2026 by qedawkins Contributor Loading…
perf: fuse MiniMax sparse cache insertion
#839 opened Jul 29, 2026 by FlamingoPg Contributor Draft
[WIP] perf(kimi3): retile warp decode
#831 opened Jul 28, 2026 by panditsa Contributor Loading…
feat(dspark): Add DSpark support
#829 opened Jul 28, 2026 by minedec Contributor Draft
feat: wire flashinfer autotuner
#820 opened Jul 27, 2026 by syuoni Member Loading…
deps: test TensorRT-LLM rc22 kernel wheel
#818 opened Jul 27, 2026 by Xiangyi1996 Collaborator Draft
feat: support torchspec training
#798 opened Jul 25, 2026 by Dogacel Contributor Loading…
feat(kernel): support small-batch Gluon MLA decode
#793 opened Jul 24, 2026 by Max191 Contributor Loading…
ProTip! Type g i on any issue or pull request to go back to the issue listing page.