tvm
https://github.com/apache/tvm
Python
Open deep learning compiler stack for cpu, gpu and specialized accelerators
Triage Issues!
When you volunteer to triage issues, you'll receive an email each day with a link to an open issue that needs help in this project. You'll also receive instructions on how to triage issues.
Triage Docs!
Receive a documented method or class from your favorite GitHub repos in your inbox every day. If you're really pro, receive undocumented methods or classes and supercharge your commit history.
Python not yet supported3 Subscribers
Add a CodeTriage badge to tvm
Help out
- Issues
- [Relax][cuDNN] Do not offload causal / non-fp16 attention, and fix the default softmax scale
- [Bug][Relax][ONNX] BinaryBase.base_impl calls .item() on a TIR PrimExpr
- [Bug][Relax][ONNX] Range._impl_v12 rejects runtime Call for `limit` argument
- [Bug] CUDA Check failed: block_realize == old_block_realize_.get() (0x88a32c0 vs. 0x88cd340)
- [Fix][S-TIR][DLight] Fall back from non-affine reduction write-back
- [Bug] CUDA Check failed: scope != "global" when using `argmin`
- [Bug][Relax] FuseOpsByPattern leaks native memory on every call (arena-allocated Group::attrs never destructed)
- [Fix][Codegen] Preserve NaNs in floating-point min and max
- [Bug] [ONNX Frontend] SparseConvolution and ScatterDense operators unsupported during import
- [Bug] `relax.build` crashes with a `ScheduleError` during the dlight GPU scheduling pass
- Docs
- Python not yet supported