Skip to content

Pull requests: RL-Align/RL-Kernel

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

feat(ascend): Qwen-Image qk_rmsnorm & multi_axis_rope kernels (issue #386)
#411 opened Sep 13, 2026 by erfgss Contributor Loading…
4 of 7 tasks
Qwen image latent pack unpack multimodal Features, bugs, or optimizations specific to multimodal support.
#410 opened Sep 12, 2026 by nodeeeeee Loading…
style: format Python sources and benchmark results
#408 opened Sep 12, 2026 by maxiaosong1124 Collaborator Loading…
Add fixed-order MoE shared/residual merge reference deepseek-P6 DSv4 platform: cuda Specific optimizations or bugs in NVIDIA graphics cards (such as FlashInfer, TMA optimizations)
#404 opened Sep 11, 2026 by AsterWang Loading…
[DSv4][P5-4] MXFP8×MXFP4 Routed Expert grouped GEMM: strict CUDA kernel deepseek-P5 DSv4 platform: cuda Specific optimizations or bugs in NVIDIA graphics cards (such as FlashInfer, TMA optimizations)
#399 opened Sep 10, 2026 by Bignonia7 Loading…
feat(musa): add native deterministic gemm kernel MUSA
#395 opened Sep 9, 2026 by Arlo-mt Collaborator Loading…
Musa support native fused logp kernel MUSA
#392 opened Sep 8, 2026 by Arlo-mt Collaborator Loading…
[DSv4][P1-S0] Start kit for the P1 work package (mHC + RMSNorm) deepseek-P1 DSv4 platform: cuda Specific optimizations or bugs in NVIDIA graphics cards (such as FlashInfer, TMA optimizations)
#383 opened Sep 3, 2026 by zhangj1an Collaborator Loading…
10 tasks
feat: enable Triton kernels on MUSA MUSA
#375 opened Sep 1, 2026 by Arlo-mt Collaborator Loading…
[DSv4][P5-0] Start kit for the P5 work package (MXFP4 Routed Expert + LoRA + Shared Expert) deepseek-P5 DSv4 platform: cuda Specific optimizations or bugs in NVIDIA graphics cards (such as FlashInfer, TMA optimizations)
#368 opened Sep 1, 2026 by KJLdefeated Collaborator Loading…
[WS1] Add batch-invariant h_aggregate kernel deepseek-P1 DSv4 platform: cuda Specific optimizations or bugs in NVIDIA graphics cards (such as FlashInfer, TMA optimizations)
#366 opened Aug 30, 2026 by nodeeeeee Loading…
docs: add DCO 1.1 text and contributor sign-off guide type: ci-cd Modify GitHub Actions, automated tests, and packaging/deployment tasks.
#359 opened Aug 29, 2026 by Zhifu-Liu Contributor Loading…
[WS2][Logp] Deterministic config option for operator platform: cuda Specific optimizations or bugs in NVIDIA graphics cards (such as FlashInfer, TMA optimizations)
#314 opened Aug 16, 2026 by KJLdefeated Collaborator Loading…
[CI][refator]: migrate GPU workflow needs-gpu-ci
#309 opened Aug 14, 2026 by Flink-ddd Collaborator Loading…
ProTip! What’s not been updated in a month: updated:<2026-08-13.