Граф коммитов

34 Коммитов

Автор SHA1 Сообщение Дата
Maneesh Gupta c99285c5fb Prefix HIP includes with hip/
[ROCm/clr commit: c4cc76398c]
2016-04-19 15:02:12 +05:30
pensun 84dbc09fe6 Add missing cudaMemsetAsync transformation
[ROCm/clr commit: 596e8e4e4e]
2016-04-14 09:02:02 -05:00
pensun b276e6b0d7 fix query of memoryClockRate and memoryBusWidth for both NV and HCC path
[ROCm/clr commit: a8ae62d399]
2016-03-25 09:24:08 -05:00
Aditya Atluri 399994788f Revert "Revert "fix nvcc for hipHostMalloc* flags.""
This reverts commit 0978c92dbc.


[ROCm/clr commit: bde1e6182d]
2016-03-21 10:39:49 -05:00
Aditya Atluri 0978c92dbc Revert "fix nvcc for hipHostMalloc* flags."
This reverts commit 849395ec02.


[ROCm/clr commit: 83fee90e83]
2016-03-21 10:36:14 -05:00
Ben Sander 849395ec02 fix nvcc for hipHostMalloc* flags.
[ROCm/clr commit: d495ffb1d3]
2016-03-21 09:33:46 -05:00
Ben Sander 37a02661a6 hipHostRegister and hipHostMalloc refactor.
Note hipHostMalloc (not hipHostAlloc or hipMallocHost).
 -  the hipHost* is used for all HIP APIs dealing with Host memory.
    (including hipHostMalloc, hipHostFree, hipHostRegister,
hipHostUnregister, hipHostGetFlags, hipHostGetDevicePointer).
  - hipMallocHost is consistent with "hipMalloc" for allocating device
    memory.  Enumerations hipHostMalloc* also used as optional
    flags parm to hipHostMalloc.


[ROCm/clr commit: 2d0fade1f7]
2016-03-22 02:30:10 -05:00
Ben Sander c2fd536c22 fix nvcc compiler
- MallocHost and FreeHost deprecation.
- Change tests to call new hipHost* equivs.
- Add missing StreamSynchronize.


[ROCm/clr commit: 6984f24d3d]
2016-03-19 04:20:15 -05:00
Ben Sander b1d3df6484 Deprecate hipMallocHost and hipFreeHost.
These will print compiler warnings if used, so we can weed them out
before removing.

Also add a default flags args for hipHostAlloc, in the C++ functioin
headers.  So you can replace hipMallocHost(&ptr, size( with hipHostAlloc(&ptr, size)


[ROCm/clr commit: 57365eb7a3]
2016-03-19 22:53:59 -05:00
Aditya Atluri cf92bfb8c7 Added hipHostRegister flags
[ROCm/clr commit: ffeba62a74]
2016-03-07 10:52:40 -06:00
Aditya Atluri 8a8836088b Added hipHostRegister feature for CUDA backend and its tests
[ROCm/clr commit: 496c549141]
2016-03-07 03:42:50 -06:00
Maneesh Gupta e0e67b0e0c Fix typo in nvcc_detail/hip_runtime_api.h
[ROCm/clr commit: b62040f6fd]
2016-03-07 09:40:15 +05:30
Aditya Atluri d7d689cf32 added feature for hipHostGetFlags for CUDA and HIP
[ROCm/clr commit: 45408db5dc]
2016-03-06 12:17:30 -06:00
Aditya Atluri e5657dec23 corrected hipDeviceGetProperties to hipGetDeviceProperties - not docs
[ROCm/clr commit: 8a21b42943]
2016-03-06 08:31:04 -06:00
Aditya Atluri 375d90a632 Added hipHostAlloc feature for CUDA
[ROCm/clr commit: 2212d35e2d]
2016-03-05 13:58:56 -06:00
Aditya Atluri 28aae1e6c6 v2 Added canHostMapMemory
[ROCm/clr commit: 6085d94f7b]
2016-03-05 13:15:07 -06:00
Aditya Atluri 8254ed4dfa Revert "Added canMapHostMemory feature"
This reverts commit 7eb8b2cc1d.


[ROCm/clr commit: a8d30da648]
2016-03-05 13:08:57 -06:00
Aditya Atluri 7eb8b2cc1d Added canMapHostMemory feature
[ROCm/clr commit: 8c3777d317]
2016-03-05 13:06:37 -06:00
Aditya Avinash Atluri a44710cd7a Merge pull request #4 from AMDComputeLibraries/memtracker
hipGetPointerAttrib behavioral changes

[ROCm/clr commit: 9c4819bc29]
2016-02-27 10:51:23 -06:00
Aditya Avinash Atluri e33fedcf9f Added CUDA support for hipPointerGetAttributes
[ROCm/clr commit: a31f878218]
2016-02-26 12:33:55 -06:00
Ben Sander e345f23846 Merge branch 'memtracker' into privatestaging
Conflicts:
	include/nvcc_detail/hip_runtime_api.h


[ROCm/clr commit: 7a1b4c3878]
2016-02-26 06:17:05 -06:00
Evgeny Mankov 8e6e28df60 Attribute hipDeviceAttributeIsMultiGpuBoard for obtaining Device property isMultiGpuBoard is added.
On HIP path property obtaining done through hsa_iterate_agents and counting the devices of HSA_DEVICE_TYPE_GPU type.

P.S.
On multi-boards systems it might be problems with detection what board a GPU plugged into (not tested).


[ROCm/clr commit: 7bb0f17656]
2016-02-25 23:44:39 +03:00
Evgeny Mankov c76791140d Guard #ifdef USE_ROCR_20 is added for ROCR_20 device properties (memoryClockRate, memoryBusWidth)
By default isn't defined.
To add ROCR_20 support HIP have to be compiled as follows: make CXX_DEFINES+=-DUSE_ROCR_20


[ROCm/clr commit: d4b15399f5]
2016-02-19 13:27:03 +03:00
Evgeny Mankov a17733dd80 Device property memoryBusWidth implementation.
+ Device property memoryBusWidth is added to hipDeviceProp_t struct.
+ Device attribute hipDeviceAttributeMemoryBusWidth is added to hipDeviceAttribute_t struct.
+ Tests update.


[ROCm/clr commit: da8169dd89]
2016-02-18 18:15:01 +03:00
Evgeny Mankov a47073f25d Device property memoryClockRate implementation.
+ Device property memoryClockRate is added to hipDeviceProp_t struct.
+ Device attribute hipDeviceAttributeMemoryClockRate is added to hipDeviceAttribute_t struct.
+ Tests update.
+ Rename hipDevAttrConcurrentKernels to hipDeviceAttributeConcurrentKernels.


[ROCm/clr commit: 8aace64dce]
2016-02-18 17:25:28 +03:00
Evgeny Mankov 3faa6fd86c Attribute hipDevAttrConcurrentKernels for obtaining Device property concurrentKernels is added.
[ROCm/clr commit: d4bd94e9a0]
2016-02-18 14:34:18 +03:00
Ben Sander a3ec0ae280 remove extra :
[ROCm/clr commit: 866e64f6e2]
2016-02-18 03:05:53 -06:00
Evgeny Mankov 6add51ef8c Fix typo: maxThreadsPerMultiProcessor -> MaxSharedMemoryPerMultiprocessor
Device property MaxSharedMemoryPerMultiprocessor set equal to totalGlobalMem (HIP path).
Reason: MaxSharedMemoryPerMultiprocessor should be as the same as group memory size. Group memory will not be paged out, so, the physical memory size = total shared memory size = group region size. NVCC path remains untouched: CUDA's device property MaxSharedMemoryPerMultiprocessor is reported.

hipify is updated as well.


[ROCm/clr commit: 460b501cbb]
2016-02-12 01:29:20 +03:00
Evgeny Mankov a8b7647f8b BDFID (BusID/DeviceID/FunctionID) support.
Except FunctionID (or DomainID in CUDA) support, because cudaDeviceProp::pciDomainID is not reported by CUDA.


[ROCm/clr commit: 658e9f0484]
2016-02-11 22:26:01 +03:00
Evgeny Mankov 9f596e0aab Device property concurrentKernels is added to hipDeviceProp_t struct.
For HCC path concurrentKernels is set to true since all ROCR hardware supports this feature.
For NVCC path concurrentKernels is obtained from CUDA's device property cudaDeviceProp::concurrentKernels.


[ROCm/clr commit: 4d4ca3ef3f]
2016-02-09 17:10:35 +03:00
Maneesh Gupta 8442259da0 Move HIP_DEVICE_COMPILE defines to hip_common.h
[ROCm/clr commit: f6e7abd710]
2016-02-09 10:57:20 +05:30
Ben Sander 0f1752e720 Fix getdeviceattr compilation for NVCC
[ROCm/clr commit: 9aec91a3b7]
2016-02-04 16:26:33 -06:00
Sam Kolton 136baccbe5 Implementation of hipDeviceGetAttribute()
[ROCm/clr commit: afe45964ae]
2016-02-04 17:39:27 +03:00
Ben Sander 28f87a0428 Initial commit for GPUOpen Launch
[ROCm/clr commit: 304171c1a2]
2016-01-26 20:14:33 -06:00