onnxruntime
70b73afd - [ADD] fuse Matmul + fastgalu -> gemmfastgelu (#11699)

Commit
3 years ago
[ADD] fuse Matmul + fastgalu -> gemmfastgelu (#11699) **Description**: Describe your changes. fuse MatMul + FastGelu -> GemmFastGelu prepare for AMD optimized fused operator GemmFastGelu usage: python benchmark.py -g -m bert-base-cased --sequence_length 384 --batch_sizes 128 --provider=rocm -p fp16 --disable_embed_layer_norm --enable_gemm_fast_gelu **Motivation and Context** - Why is this change required? What problem does it solve? - If it fixes an open issue, please link to the issue here.
Author
Parents
Loading