[Metal] Reduce number of threads for reduction layers #8206

echuraev · 2021-06-07T14:17:01Z

Reduced default number of threads in reduction kernels for Metal.
Default code generation generated thread block with the following size:
32x32x1. With this size number of threads per threadgroup was equal to
1024 (32 * 32 * 1). Sometimes device doesn't have enough resources and
in this case we will get an exception that the block size is greater
than value of maxTotalThreadsPerThreadgroup.
To prevent such situation we decrease default number of threads. With
this fix every model should work with default codegen and auto-tuning or
auto-scheduling will select the optimal number of threads.

Thanks for contributing to TVM! Please refer to guideline https://tvm.apache.org/docs/contribute/ for useful information and tips. After the pull request is submitted, please request code reviews from Reviewers by @ them in the pull request thread.

Reduced default number of threads in reduction kernels for Metal. Default code generation generated thread block with the following size: 32x32x1. With this size number of threads per threadgroup was equal to 1024 (32 * 32 * 1). Sometimes device doesn't have enough resources and in this case we will get an exception that the block size is greater than value of maxTotalThreadsPerThreadgroup. To prevent such situation we decrease default number of threads. With this fix every model should work with default codegen and auto-tuning or auto-scheduling will select the optimal number of threads.

echuraev marked this pull request as ready for review June 7, 2021 16:34

ZihengJiang added the status: need review label Jun 8, 2021

masahi approved these changes Jun 10, 2021

View reviewed changes

masahi merged commit 8ea6a30 into apache:main Jun 10, 2021

echuraev deleted the echuraev/metal_fix_max_total_threads branch September 24, 2021 10:37

junrushao mentioned this pull request Nov 1, 2021

Apache TVM v0.8 Release Note Candidate #9416

Closed

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

[Metal] Reduce number of threads for reduction layers #8206

[Metal] Reduce number of threads for reduction layers #8206

echuraev commented Jun 7, 2021

[Metal] Reduce number of threads for reduction layers #8206

[Metal] Reduce number of threads for reduction layers #8206

Conversation

echuraev commented Jun 7, 2021