-
Notifications
You must be signed in to change notification settings - Fork 4.2k
Pull requests: microsoft/onnxruntime
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
Reject nodes that feed inputs to a zero-input operator schema
#32633
opened Sep 16, 2026 by
shiyi (shiyi9801)
Contributor
Loading…
[MLAS] Fix fp16 overflow/NaN bugs in ARM64 NEON Gelu/Erf and add extreme-value test coverage
#32631
opened Sep 16, 2026 by
kjg0724 (kjg0724)
Contributor
Loading…
[WebGPU] Refactor MatMul algorithm selection
#32630
opened Sep 16, 2026 by
Jiajia Qin (qjia7)
Contributor
•
Draft
[WebGPU] Fix mixed-type LayerNormalization shaders
#32629
opened Sep 16, 2026 by
Jiajia Qin (qjia7)
Contributor
•
Draft
[WebNN EP] Support NHWC layout in Resize by inferring resample2d axes
#32628
opened Sep 16, 2026 by
Wanming Lin (Honry)
Contributor
Loading…
Enable opt-in host-pageable CUDA GatherBlockQuantized
#32626
opened Sep 16, 2026 by
Copilot
AI
Loading…
[CUDA] Address cuDNN paged SDPA review feedback
#32624
opened Sep 16, 2026 by
Tianlei Wu (tianleiwu)
Contributor
Loading…
[WebGPU] checks to handle oob writes based on seqlens_k
ep:WebGPU
ort-web webgpu provider
#32623
opened Sep 15, 2026 by
Edward Chen (edgchen1)
Contributor
Loading…
Optimize CPU uint8 GatherBlockQuantized dequantization
#32621
opened Sep 15, 2026 by
Copilot
AI
Loading…
Add review guidance for deprecated components
#32620
opened Sep 15, 2026 by
Edward Chen (edgchen1)
Contributor
Loading…
[CUDA] Enable non-causal PagedAttention with small cache pages
#32619
opened Sep 15, 2026 by
Copilot
AI
Loading…
[CUDA][WebGPU] Add packed sparse attention indexer for continuous batching
#32618
opened Sep 15, 2026 by
Copilot
AI
•
3/3
Loading…
Add GQA workspace estimation
ep:CUDA
issues related to the CUDA execution provider
memory
#32617
opened Sep 15, 2026 by
Ti-Tai Wang (titaiwangms)
Contributor
Loading…
Improve CPU provider build time with Eigen PCH
#32616
opened Sep 15, 2026 by
Ti-Tai Wang (titaiwangms)
Contributor
Loading…
Optimize single-source WebGPU SparsePagedAttention
#32613
opened Sep 15, 2026 by
Copilot
AI
•
3/3
Loading…
Fix heap buffer overflow from reusing a packed sub-byte buffer for a full-byte tensor
#32611
opened Sep 15, 2026 by
Wei Wang (wangw-1991)
Contributor
Loading…
Fix out-of-bounds read in EmbedLayerNormalization shape inference for scalar beta
#32610
opened Sep 15, 2026 by
Wei Wang (wangw-1991)
Contributor
Loading…
Fix out-of-bounds read when a Loop/Scan body input is also an initializer
#32609
opened Sep 15, 2026 by
Wei Wang (wangw-1991)
Contributor
Loading…
[MLAS] Add RVV fused activation fast path on riscv64
#32608
opened Sep 15, 2026 by
ww8191201-coder
Loading…
Harden shape inference for contrib ops
#32607
opened Sep 15, 2026 by
shiyi (shiyi9801)
Contributor
Loading…
[WebNN EP] Support GroupNormalization, GroupNorm and SkipGroupNorm
#32606
opened Sep 15, 2026 by
Wanming Lin (Honry)
Contributor
Loading…
[WebGPU] Widen SubgroupMatrixMatMulNBits tile size to 128x256
#32605
opened Sep 15, 2026 by
Jianhui Dai (daijh)
Contributor
•
Draft
Previous Next
ProTip!
Type g p on any issue or pull request to go back to the pull request listing page.