Evgeny Mankov 1025341300 Device property maxThreadsPerMultiProcessor set equal to totalGlobalMem (HIP path).
Reason: maxThreadsPerMultiProcessor should be as the same as group memory size. Group memory will not be paged out, so, the physical memory size = total shared memory size = group region size.

NVCC path remains untouched: CUDA's device property maxThreadsPerMultiProcessor is reported.
2016-02-12 00:04:14 +03:00
S
Cur síos
Níor tugadh tuairisc
282 MiB
Teangacha
C++ 67.5%
C 20.6%
Python 6.6%
CMake 3.4%
Shell 0.6%
Eile 1.1%