Allocate EOP queue local to GPU
On discrete GPUs place the EOP queue in VRAM. The reader/writer of this queue is the CP and the size is small. Dispatch latency improves through lower read latency in AQL completion phase. Change-Id: Id8351dcddbd21fd7c7d699803c96434c9132db71 Signed-off-by: Jay Cornwall <Jay.Cornwall@amd.com>
This commit is contained in:
+1
-1
@@ -76,7 +76,7 @@ HSAKMT_STATUS HSAKMTAPI hsaKmtCreateEvent(HsaEventDescriptor *EventDesc,
|
||||
|
||||
if (is_dgpu && !events_page) {
|
||||
events_page = allocate_exec_aligned_memory_gpu(
|
||||
KFD_SIGNAL_EVENT_LIMIT * 8, PAGE_SIZE, 0, true);
|
||||
KFD_SIGNAL_EVENT_LIMIT * 8, PAGE_SIZE, 0, true, false);
|
||||
if (!events_page) {
|
||||
pthread_mutex_unlock(&hsakmt_mutex);
|
||||
return HSAKMT_STATUS_ERROR;
|
||||
|
||||
Reference in New Issue
Block a user