Evgeny Mankov
c76791140d
Guard #ifdef USE_ROCR_20 is added for ROCR_20 device properties (memoryClockRate, memoryBusWidth)
...
By default isn't defined.
To add ROCR_20 support HIP have to be compiled as follows: make CXX_DEFINES+=-DUSE_ROCR_20
[ROCm/clr commit: d4b15399f5 ]
2016-02-19 13:27:03 +03:00
Evgeny Mankov
a17733dd80
Device property memoryBusWidth implementation.
...
+ Device property memoryBusWidth is added to hipDeviceProp_t struct.
+ Device attribute hipDeviceAttributeMemoryBusWidth is added to hipDeviceAttribute_t struct.
+ Tests update.
[ROCm/clr commit: da8169dd89 ]
2016-02-18 18:15:01 +03:00
Evgeny Mankov
a47073f25d
Device property memoryClockRate implementation.
...
+ Device property memoryClockRate is added to hipDeviceProp_t struct.
+ Device attribute hipDeviceAttributeMemoryClockRate is added to hipDeviceAttribute_t struct.
+ Tests update.
+ Rename hipDevAttrConcurrentKernels to hipDeviceAttributeConcurrentKernels.
[ROCm/clr commit: 8aace64dce ]
2016-02-18 17:25:28 +03:00
Evgeny Mankov
3faa6fd86c
Attribute hipDevAttrConcurrentKernels for obtaining Device property concurrentKernels is added.
...
[ROCm/clr commit: d4bd94e9a0 ]
2016-02-18 14:34:18 +03:00
Evgeny Mankov
6add51ef8c
Fix typo: maxThreadsPerMultiProcessor -> MaxSharedMemoryPerMultiprocessor
...
Device property MaxSharedMemoryPerMultiprocessor set equal to totalGlobalMem (HIP path).
Reason: MaxSharedMemoryPerMultiprocessor should be as the same as group memory size. Group memory will not be paged out, so, the physical memory size = total shared memory size = group region size. NVCC path remains untouched: CUDA's device property MaxSharedMemoryPerMultiprocessor is reported.
hipify is updated as well.
[ROCm/clr commit: 460b501cbb ]
2016-02-12 01:29:20 +03:00
Evgeny Mankov
a8b7647f8b
BDFID (BusID/DeviceID/FunctionID) support.
...
Except FunctionID (or DomainID in CUDA) support, because cudaDeviceProp::pciDomainID is not reported by CUDA.
[ROCm/clr commit: 658e9f0484 ]
2016-02-11 22:26:01 +03:00
Evgeny Mankov
2478fc078f
Formatting, no functional changes
...
[ROCm/clr commit: d9a94191f2 ]
2016-02-10 17:21:18 +03:00
Evgeny Mankov
9f596e0aab
Device property concurrentKernels is added to hipDeviceProp_t struct.
...
For HCC path concurrentKernels is set to true since all ROCR hardware supports this feature.
For NVCC path concurrentKernels is obtained from CUDA's device property cudaDeviceProp::concurrentKernels.
[ROCm/clr commit: 4d4ca3ef3f ]
2016-02-09 17:10:35 +03:00
Ben Sander
3b04ce4e81
minor doc touchup
...
[ROCm/clr commit: 2ecb345a67 ]
2016-02-08 22:11:11 -06:00
Sam Kolton
136baccbe5
Implementation of hipDeviceGetAttribute()
...
[ROCm/clr commit: afe45964ae ]
2016-02-04 17:39:27 +03:00
sunway513
90f385fc55
Fix some typos and incorrect namings in comments
...
[ROCm/clr commit: 04aa623569 ]
2016-01-28 13:17:44 -06:00
Ben Sander
28f87a0428
Initial commit for GPUOpen Launch
...
[ROCm/clr commit: 304171c1a2 ]
2016-01-26 20:14:33 -06:00