onednn
https://github.com/oneapi-src/onednn
C++
oneAPI Deep Neural Network Library (oneDNN)
Triage Issues!
When you volunteer to triage issues, you'll receive an email each day with a link to an open issue that needs help in this project. You'll also receive instructions on how to triage issues.
Triage Docs!
Receive a documented method or class from your favorite GitHub repos in your inbox every day. If you're really pro, receive undocumented methods or classes and supercharge your commit history.
C++ not yet supported2 Subscribers
Add a CodeTriage badge to onednn
Help out
- Issues
- doc: remove outdated note
- xe: optimize dispatch_compile_params_t size
- scripts: verbose converter: add missing sparse encoding support
- cpu: aarch64: conv: add async runtime support
- gpu: sdpa: xe_hpg head_size=128 config is 1.6-1.8x slower than a 16-subgroup alternative on Arc A770
- cpu: aarch64: brgemm_matmul: make async compatible
- rfcs: proposal for extending async profiling to cpu threadpool
- [do-not-merge] [WIP] aarch64: brgemm: U8-by-S8 USMMLA direct BRGEMM and BRCONV
- [do-not-merge] [WIP] aarch64: brgemm: enable SMMLA for S8S8 BRGEMM and BRGCONV
- [do-not-merge] [WIP] aarch64: brgemm: BFMMLA BRGConv BRGMatMul Enablement
- Docs
- C++ not yet supported