Skip to content

Pull requests: ROCm/FlyDSL

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

[Fix] conv3d fp8: guard padded M rows in the epilogue store
#969 opened Aug 5, 2026 by jiacao-amd Contributor Loading…
[Kernel][MI350] Add bias, alibi bias and sink to flash attention
#960 opened Aug 3, 2026 by amd-nprotaso Contributor Loading…
Add SVD quant for Quark
#957 opened Aug 2, 2026 by amd-xiaoyu12 Draft
1 task
BugFix for MegaMoE multi-gpu
#954 opened Aug 2, 2026 by GwilliamHu Member Loading…
[Tool] Add per-kernel ISA resource delta table
#928 opened Jul 30, 2026 by Phil-amd Member Loading…
3 of 4 tasks
Unify benchmark timing contracts and add calibrated CI gates
#924 opened Jul 30, 2026 by jhinpan Collaborator Loading…
6 tasks done
[llvm] bump up llvm and adapt FlyDSL to new api
#922 opened Jul 29, 2026 by jli-melchior Collaborator Loading…
1 task
[DSL] Preserve logical signedness of unsigned integer dtypes
#920 opened Jul 28, 2026 by Arist12 Contributor Loading…
6 tasks done
[Bugfix][Dialect] Reject vector operands in atomic copy atoms
#918 opened Jul 28, 2026 by AiyyappanMR Loading…
1 task done
[Dialect][Perf] Don't merge mixed static/runtime offsets on LDS pointers
#914 opened Jul 27, 2026 by Arist12 Contributor Loading…
6 tasks done
[Kernel][Perf] Optimize paged-attention metadata decode
#910 opened Jul 27, 2026 by fsx950223 Contributor Loading…
5 tasks done
Add fast_divmod magic-number division helper
#906 opened Jul 26, 2026 by kashif Contributor Loading…
Add Optimized MoE Routing Path
#901 opened Jul 24, 2026 by amd-wsung102 Contributor Loading…
1 task done
Enabling coexec llvm for Flydsl
#900 opened Jul 24, 2026 by omuhamma Draft
[Test] Add gfx1250 WMMA lowering tests for additional dtypes
#891 opened Jul 23, 2026 by AiyyappanMR Loading…
2 of 3 tasks
[Kernel][MI350] Add cache-aware 8-wave BF16/FP16 GEMM
#888 opened Jul 23, 2026 by zhanglx13 Draft
6 tasks done
gemm: add fp8 per-tensor grouped GEMM forward (M-grouped/MoE)
#887 opened Jul 23, 2026 by kyle-256 Loading…
1 task done
Add optional forward LSE output
#886 opened Jul 23, 2026 by AakarshAMD Loading…
a16w16 for gfx1250 on flydsl
#875 opened Jul 20, 2026 by omuhamma Loading…
ProTip! Type g p on any issue or pull request to go back to the pull request listing page.