MT doesn't use GPU waits, but CPU for sync between engines. Change the threshold values for CPU waits for direct dispatch. That will bring behavior closer to MT. Change-Id: Ia41c3cb812614962aff2746b6cf858f1bf77dda2 [ROCm/clr commit: ca2ea70a6c]
ca2ea70a6c