-
Notifications
You must be signed in to change notification settings - Fork 94
Pull requests: NVIDIA/TensorRT-Edge-LLM
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
fix(attention): Scope FMHA cubin loading by mask type
#148
opened Jul 25, 2026 by
suharvest
Loading…
feat(builder): Set workspace RAM limits and enhance log level controls
#126
opened Jul 6, 2026 by
cjac
Loading…
7 tasks done
Strip /usr/include from apt pybind11's interface include dirs
#121
opened Jul 2, 2026 by
whitesscott
Loading…
4 tasks
Use <cmath> instead of <math.h> in context attention FMHA headers
#120
opened Jul 1, 2026 by
whitesscott
Loading…
3 tasks
fix: Respect native Orin CUDA architecture and CuTe link requirements
#118
opened Jun 28, 2026 by
suharvest
Loading…
fix: propagate CuTe DSL runtime link requirements for static libraries
#103
opened Jun 7, 2026 by
francismelon
Loading…
3 of 7 tasks
feat https://github.com/NVIDIA/TensorRT-Edge-LLM/issues/87: add CustomVoice language conditioning support for Qwen3-TTS
#98
opened May 27, 2026 by
suharvest
Loading…
6 of 8 tasks
fix https://github.com/NVIDIA/TensorRT-Edge-LLM/issues/87: hard-error instead of silent return when CuTe DSL GEMM is not compiled
#97
opened May 27, 2026 by
suharvest
Loading…
6 of 7 tasks
imageUtilKernels.cu: optimize initAttentionMaskKernel
#28
opened Jan 26, 2026 by
vitamin-chaos
Loading…
ProTip!
Follow long discussions with comments:>50.