-
Notifications
You must be signed in to change notification settings - Fork 98
Pull requests: ROCm/FlyDSL
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
[ROCDL] Add arch-aware universal s_waitcnt
#915
opened Jul 28, 2026 by
sjfeng1999
Collaborator
Loading…
1 task
[Dialect][Perf] Don't merge mixed static/runtime offsets on LDS pointers
#914
opened Jul 27, 2026 by
Arist12
Contributor
Loading…
6 tasks done
Fix hierarchical reduced predicates in copy layout lowering
#912
opened Jul 27, 2026 by
HydraQYH
Loading…
1 task
[Kernel][Perf] Optimize paged-attention metadata decode
#910
opened Jul 27, 2026 by
fsx950223
Contributor
Loading…
5 tasks done
[Kernel] Migrate SWA decode to tile programming
#909
opened Jul 27, 2026 by
fsx950223
Contributor
Loading…
7 tasks done
Add Optimized MoE Routing Path
#901
opened Jul 24, 2026 by
amd-wsung102
Contributor
Loading…
1 task done
[fix] Autotuner: propagate the tuned function's return value
#899
opened Jul 24, 2026 by
aryaman-gupta
Loading…
[Kernel][MI350] Add 8 wave GQA sliding-window attention kernel
#894
opened Jul 23, 2026 by
amd-nprotaso
Loading…
[Test] Add gfx1250 WMMA lowering tests for additional dtypes
#891
opened Jul 23, 2026 by
AiyyappanMR
Loading…
2 of 3 tasks
gemm: add fp8 per-tensor grouped GEMM forward (M-grouped/MoE)
#887
opened Jul 23, 2026 by
kyle-256
Loading…
1 task done
[compiler] carry Python containers through dynamic if/for/while/ifexp
#874
opened Jul 20, 2026 by
xudoyuan
Collaborator
Loading…
1 task
[Kernel] Add optimized 4-wave MXFP8 GEMM kernel for gfx950
#872
opened Jul 18, 2026 by
aris134
Loading…
1 task done
[Kernel] fp8 conv3d: 8-wave GEMM pipeline + BIG_IN fix
#860
opened Jul 15, 2026 by
jiacao-amd
Contributor
•
Draft
3 tasks
[Perf] Optimize rmsnorm/layernorm
#848
opened Jul 14, 2026 by
cschenjunlin
Contributor
Loading…
1 task
Previous Next
ProTip!
Follow long discussions with comments:>50.