Skip to content

Pull requests: NVIDIA/cutlass

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

Add missing template disambiguators
#3433 opened Aug 4, 2026 by Yinwei-Zhang Loading…
docs: fix broken CuTe DSL educational notebook links
#3431 opened Aug 3, 2026 by felipeofdev-ai Loading…
1 of 2 tasks
Add CuTeDSL math tests for sqrt operation
#3430 opened Aug 3, 2026 by Pritiks23 Loading…
Fix GC-cycle leak: drop frame self-references in CuTe DSL decorators
#3423 opened Jul 30, 2026 by thakkarV Contributor Loading…
[CuTeDSL] Add prepared-launch API for low-overhead kernel replay
#3419 opened Jul 29, 2026 by zkyue Contributor Loading…
Fix pre-Blackwell validation in the Ada FP8 GEMM example
#3411 opened Jul 27, 2026 by Lec16sf Loading…
Require CUDA 12.9 for SM121 default architectures
#3410 opened Jul 26, 2026 by davidkny22 Contributor Loading…
[CuTeDSL] Make autotuning more user friendly & new docs
#3408 opened Jul 24, 2026 by kainzhong Contributor Loading…
[CuTeDSL] Let cute.compile opt into the compile cache
#3402 opened Jul 22, 2026 by aryanputta Loading…
[examples] Fix stale thread layout comment in wgmma_sm90.cu
#3401 opened Jul 21, 2026 by SriRangaTarun Contributor Loading…
[Hopper CuTeDSL] Fix max reduction in fmha kernel
#3399 opened Jul 20, 2026 by Aladoro Loading…
Fix CuTe tuple protocol compatibility across CCCL versions
#3386 opened Jul 14, 2026 by xiufanl Contributor Loading…
Add tuning notes for Hopper 16-bit dense GEMM
#3373 opened Jul 7, 2026 by hamuzhan Loading…
ProTip! Type g i on any issue or pull request to go back to the issue listing page.