Anndrey24 opened a new pull request, #17003:
URL: https://github.com/apache/tvm/pull/17003

   This commit adds a scalable `arm_cpu` conv2d NHWC schedule for fp32 which 
generates SME instructions by using the tensor intrinsics introduced in #16921.
   
   Alongside the SME schedule, the logic of the TE schedule 
`schedule_conv2d_gemm_native()` for both non-scalable and scalable vector 
implementations has also been translated into the new TIR schedule. This means 
that the TE compute definition `compute_conv2d_NHWC_hybrid()` is now compatible 
with both the original TE schedules (e.g. `schedule_conv2d_NHWC_hybrid()`) and 
the newly introduced TIR schedule `schedule_conv2d_NHWC_hybrid_TIR()`. The 
corresponding TOPI test has been extended to reflect that.
   
   cc @ekalda @lhutton1 


-- 
This is an automated message from the Apache Git Service.
To respond to the message, please log on to GitHub and use the
URL above to go to the specific comment.

To unsubscribe, e-mail: [email protected]

For queries about this service, please contact Infrastructure at:
[email protected]

Reply via email to