The GitHub Actions job "tvm-bot" on tvm.git/main has failed. Run started by GitHub user tlopex (triggered by tlopex).
Head commit for run: 1a764d799316c7d73123750a0a3f7ce40c7d5cba / Shushi Hong <[email protected]> [Tests] Reduce runtime of slow Python tests (#20006) This PR reduces the runtime of several slow Python test groups: - Parameterize LLVM division and CUDA vectorized-cast cases so pytest-xdist can schedule them independently. - Replace exhaustive ONNX execution with structural importer checks plus representative numerical cases, and avoid registering unsupported backend cases. - Reuse compiled paged-attention kernels across compatible test cases. Targeted measurements showed: - LLVM division: 25.34s → 19.87s - CUDA vectorized casts: 142.85s → 108.40s - Paged-attention CPU: 316.32s → 192.50s - ONNX Conv: 25.84s → 1.81s - ONNX Reduce: 20.66s → 4.04s This PR also fixes a latent CUDA Graph cleanup bug that could leave `cudaErrorStreamCaptureInvalidated` in the worker thread and cause unrelated subsequent GPU tests to fail. Report URL: https://github.com/apache/tvm/actions/runs/29456763618 With regards, GitHub Actions via GitBox --------------------------------------------------------------------- To unsubscribe, e-mail: [email protected] For additional commands, e-mail: [email protected]
