Skip to content

Pull requests: RL-Align/RL-Kernel

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobodyLoading
Sort

Pull requests list

[PERF][distributed]: optimize deterministic ROCm collectives with HIP IPC platform: rocm Specific tasks specific to AMD graphics cards (such as CK, bpreshuffle/FA)
#357 opened Aug 29, 2026 by maxiaosong1124CollaboratorLoading…
[FEAT][distributed]: add deterministic ROCm/RCCL transport collectives platform: rocm Specific tasks specific to AMD graphics cards (such as CK, bpreshuffle/FA)
#356 opened Aug 28, 2026 by Flink-dddCollaboratorLoading…
bench(ffn): add H100 cuBLAS and vLLM comparisons platform: cuda Specific optimizations or bugs in NVIDIA graphics cards (such as FlashInfer, TMA optimizations)
#348 opened Aug 27, 2026 by frank-2077CollaboratorLoading…
[ROCm] Deterministic fused linear logp in Triton platform: rocm Specific tasks specific to AMD graphics cards (such as CK, bpreshuffle/FA)
#347 opened Aug 27, 2026 by KJLdefeatedCollaboratorLoading…
[Runtime][TP2 CP2] Publish strict bitwise runtime and user modes platform: cuda Specific optimizations or bugs in NVIDIA graphics cards (such as FlashInfer, TMA optimizations)
#344 opened Aug 26, 2026 by inaniloquenteeCollaboratorLoading…
feat(ascend): add prefix-shared attention Ascend C kernel Ascend
#340 opened Aug 25, 2026 by zhangj1anCollaboratorLoading…
feat(ascend): add SwiGLU operator Ascend
#334 opened Aug 23, 2026 by erfgssLoading…
feat(logprob): add deterministic ROCm vocab-parallel path platform: rocm Specific tasks specific to AMD graphics cards (such as CK, bpreshuffle/FA)
#328 opened Aug 21, 2026 by hihaluemenContributorLoading…
feat(ffn): add deterministic distributed Triton FFN for ROCm platform: rocm Specific tasks specific to AMD graphics cards (such as CK, bpreshuffle/FA)
#325 opened Aug 20, 2026 by frank-2077CollaboratorLoading…
[WS2][gemm][2th merge] ffn collectives on test
#321 opened Aug 19, 2026 by frank-2077CollaboratorLoading…
[WS1][kernels] Deterministic attention Ascend C kernel Ascend
#320 opened Aug 19, 2026 by zhangj1anCollaboratorLoading…
feat(attention): add strict bitwise ROCm path platform: rocm Specific tasks specific to AMD graphics cards (such as CK, bpreshuffle/FA)
#319 opened Aug 19, 2026 by inaniloquenteeCollaboratorLoading…
[WS2][Logp] Deterministic config option for operator platform: cuda Specific optimizations or bugs in NVIDIA graphics cards (such as FlashInfer, TMA optimizations)
#314 opened Aug 16, 2026 by KJLdefeatedCollaboratorLoading…
[CI][refator]: migrate GPU workflow needs-gpu-ci
#309 opened Aug 14, 2026 by Flink-dddCollaboratorLoading…
[WS2][Mismatch] Add Qwen3 FFN implementation factor
#308 opened Aug 13, 2026 by bitborneCollaboratorLoading…
[WS2][GEMM] Add Qwen3 FFN orchestration and consistency validation
#304 opened Aug 13, 2026 by bitborneCollaboratorLoading…
5 of 10 tasks
[WS2] Ablation Matrix API v2.0 on top of PR230
#288 opened Aug 9, 2026 by zhangj1anCollaboratorLoading…
[WS2][PR2][Logp] feat: add TP=1 logprob comparison harness
#262 opened Aug 4, 2026 by hihaluemenContributorLoading…
ProTip! Updated in the last three days: updated:>2026-08-26.