CUTLASS 2.3 adds GEMMs targeting Sparse Tensor Cores on the NVIDIA Ampere Architecture, fast SGEMM, and small matrix classes, bug fixes, and performance enhancements. |
||
|---|---|---|
| .. | ||
| gemm_operation.py | ||
| generator.py | ||
| library.py | ||
| manifest.py | ||
CUTLASS 2.3 adds GEMMs targeting Sparse Tensor Cores on the NVIDIA Ampere Architecture, fast SGEMM, and small matrix classes, bug fixes, and performance enhancements. |
||
|---|---|---|
| .. | ||
| gemm_operation.py | ||
| generator.py | ||
| library.py | ||
| manifest.py | ||