Lunderberg opened a new pull request, #16302:
URL: https://github.com/apache/tvm/pull/16302

   The updates to the cutlass kernels made in TVM PR #16244 require symbols 
provided in cuda 7.5+.  While the cuda architecture is specified by setting 
`NVCC_FLAGS` in the `CMakeLists.txt` for each kernel, cmake 3.18+ also sets it 
based on the
   `CMAKE_CUDA_ARCHITECTURES` value.  If not set, cmake will explicitly pass 
the compute capability as nvidia's default of 5.2, *EVEN IF* it has already 
been specified in `NVCC_FLAGS`.  Because the kernels cannot compile with 
compute capability of 5.2, this causes compilation errors.
   
   By setting `CMAKE_CUDA_ARCHITECTURES` to `OFF`, cmake does not add 5.2 as a 
target architecture.
   
   See https://cmake.org/cmake/help/latest/policy/CMP0104.html for details on 
CMake's policy for CUDA architecture flags.
   
   See https://cmake.org/cmake/help/latest/policy/CMP0104.html for the default 
CUDA architecture for each version of CUDA.


-- 
This is an automated message from the Apache Git Service.
To respond to the message, please log on to GitHub and use the
URL above to go to the specific comment.

To unsubscribe, e-mail: [email protected]

For queries about this service, please contact Infrastructure at:
[email protected]

Reply via email to