Issues 共 1377
warpgroup_mma_wait does not properly synchronize across warpgroups
#11047 · saagarjha · 24 天前
assertion failure at mlir::triton::ColumnAction::apply in triton/lib/Tools/LinearLayout.cpp
#10910 · YuanchengJiang · 2026-07-16
[BUG] sm_103 (B300): `tl.dot` chunked-recurrence kernel is non-deterministic with `num_warps∈{4,8}, num_stages∈{2,3}` — reproduces on Triton 3.7.0, i.e. NOT fixed by #9615 / #9871's fix
#10590 · michaelroyzen · 2026-06-12
[v.3.7.1] Release Tracker
#10553 · atalman · 2026-06-09
Probabilistic illegal memory access from a kernel with multiple tl.dot on H100 (sm_90)
#10486 · tarinduj · 2026-06-04
`make_llir` ~8-minute O(n²) stall after LLVM bump #9746 (BuiltinFuncToLLVM / simplifyRegions) on AMD backend
#10465 · mgehre-amd · 2026-06-03
Inquiry: How does Triton handle precision differences caused by summation order in tl.sum?
#10438 · maxh2018 · 2026-06-01
AMD: `tritonamdgpu-canonicalize-pointers` asserts on `tl.where` between two base pointers (gfx1150)
#10381 · kashif · 2026-05-26
TritonGPURemoveLayoutConversions dominance failure compiling vLLM fused_moe_kernel on sm87
#10327 · massif-01 · 2026-05-15
Triton v3.7.0 - Folder examples/plugins not included into the release
#10262 · waltercool · 2026-05-08