Skip to content

Pull requests: RL-Align/RL-Kernel

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

Add VIME Qwen3-8B TP4/CP2 consistency experiment and results
#377 opened Sep 1, 2026 by inaniloquentee Collaborator Loading…
7 tasks done
Musa support native kernels
#376 opened Sep 1, 2026 by Arlo-mt Draft
feat: enable Triton kernels on MUSA MUSA
#375 opened Sep 1, 2026 by Arlo-mt Loading…
docs(readme): update feature and platform support documentation Improvements or additions to documentation
#374 opened Sep 1, 2026 by Flink-ddd Collaborator Loading…
[skill] add ws1 ascend kernel
#373 opened Sep 1, 2026 by zhangj1an Collaborator Loading…
[WS1][Ascend] [Qwen3-8b] Fused linear logp ops Ascend
#372 opened Sep 1, 2026 by zhangj1an Collaborator Loading…
[WS1][Ascend] [Qwen3-8b] LM head ops Ascend
#371 opened Sep 1, 2026 by zhangj1an Collaborator Loading…
[WS1][Ascend] [Qwen3-8b] Fused logp ops Ascend
#370 opened Sep 1, 2026 by zhangj1an Collaborator Loading…
[WS1][Ascend] [Qwen3-8b] Embedding ops Ascend
#369 opened Sep 1, 2026 by zhangj1an Collaborator Loading…
[DSv4][P5-0] Start kit for the P5 work package (MXFP4 Routed Expert + LoRA + Shared Expert) deepseek-P5 DSv4 platform: cuda Specific optimizations or bugs in NVIDIA graphics cards (such as FlashInfer, TMA optimizations)
#368 opened Sep 1, 2026 by KJLdefeated Collaborator Loading…
[WS1] Add batch-invariant h_aggregate kernel deepseek-P1 DSv4 platform: cuda Specific optimizations or bugs in NVIDIA graphics cards (such as FlashInfer, TMA optimizations)
#366 opened Aug 30, 2026 by nodeeeeee Loading…
feat(ascend): add batch-invariant RMSNorm Ascend C operator Ascend
#364 opened Aug 30, 2026 by erfgss Contributor Loading…
docs: add DCO 1.1 text and contributor sign-off guide type: ci-cd Modify GitHub Actions, automated tests, and packaging/deployment tasks.
#359 opened Aug 29, 2026 by Zhifu-Liu Contributor Loading…
[PERF][distributed]: optimize deterministic ROCm collectives with HIP IPC platform: rocm Specific tasks specific to AMD graphics cards (such as CK, bpreshuffle/FA)
#357 opened Aug 29, 2026 by maxiaosong1124 Collaborator Loading…
[FEAT][distributed]: add deterministic ROCm/RCCL transport collectives platform: rocm Specific tasks specific to AMD graphics cards (such as CK, bpreshuffle/FA)
#356 opened Aug 28, 2026 by Flink-ddd Collaborator Loading…
feat(ascend): add deterministic collective Ascend C kernel Ascend
#355 opened Aug 28, 2026 by zhangj1an Collaborator Loading…
[ROCm] Deterministic fused linear logp in Triton platform: rocm Specific tasks specific to AMD graphics cards (such as CK, bpreshuffle/FA)
#347 opened Aug 27, 2026 by KJLdefeated Collaborator Loading…
feat(ascend): add prefix-shared attention Ascend C kernel Ascend
#340 opened Aug 25, 2026 by zhangj1an Collaborator Loading…
feat(logprob): add deterministic ROCm vocab-parallel path platform: rocm Specific tasks specific to AMD graphics cards (such as CK, bpreshuffle/FA)
#328 opened Aug 21, 2026 by hihaluemen Contributor Loading…
feat(ffn): add deterministic distributed Triton FFN for ROCm platform: rocm Specific tasks specific to AMD graphics cards (such as CK, bpreshuffle/FA)
#325 opened Aug 20, 2026 by frank-2077 Collaborator Loading…
[WS1][kernels] Deterministic attention Ascend C kernel Ascend
#320 opened Aug 19, 2026 by zhangj1an Collaborator Loading…
feat(attention): add strict bitwise ROCm path platform: rocm Specific tasks specific to AMD graphics cards (such as CK, bpreshuffle/FA)
#319 opened Aug 19, 2026 by inaniloquentee Collaborator Loading…
[WS2][Logp] Deterministic config option for operator platform: cuda Specific optimizations or bugs in NVIDIA graphics cards (such as FlashInfer, TMA optimizations)
#314 opened Aug 16, 2026 by KJLdefeated Collaborator Loading…
[CI][refator]: migrate GPU workflow needs-gpu-ci
#309 opened Aug 14, 2026 by Flink-ddd Collaborator Loading…
[WS2][Mismatch] Add Qwen3 FFN implementation factor
#308 opened Aug 13, 2026 by bitborne Collaborator Loading…
ProTip! Exclude everything labeled bug with -label:bug.