Allocate EOP queue local to GPU

On discrete GPUs place the EOP queue in VRAM. The reader/writer of this
queue is the CP and the size is small. Dispatch latency improves
through lower read latency in AQL completion phase.

Change-Id: Id8351dcddbd21fd7c7d699803c96434c9132db71
Signed-off-by: Jay Cornwall <Jay.Cornwall@amd.com>
This commit is contained in:
Jay Cornwall
2018-02-22 13:06:15 -06:00
parent 25170c3c57
commit e2c353dc0d
3 changed files with 13 additions and 10 deletions
+1 -1
View File
@@ -76,7 +76,7 @@ HSAKMT_STATUS HSAKMTAPI hsaKmtCreateEvent(HsaEventDescriptor *EventDesc,
if (is_dgpu && !events_page) {
events_page = allocate_exec_aligned_memory_gpu(
KFD_SIGNAL_EVENT_LIMIT * 8, PAGE_SIZE, 0, true);
KFD_SIGNAL_EVENT_LIMIT * 8, PAGE_SIZE, 0, true, false);
if (!events_page) {
pthread_mutex_unlock(&hsakmt_mutex);
return HSAKMT_STATUS_ERROR;