839c07c4aa
* [CI] Testing Stability
- CMake option ROCPROFILER_DISABLE_UNSTABLE_CTESTS
- used for tests which periodically fail around 1 out of every 10 runs
- set to ON while instability remains, this needs to set to OFF in ROCm 7.1 or, ideally, ROCm 7.0.1
- Use FIXTURES_SETUP and FIXTURES_REQUIRED for some tests
- replace "threw an exception" with "${ROCPROFILER_DEFAULT_FAIL_REGEX}" for misc FAIL_REGULAR_EXPRESSIONS
* Remove contents of all EXCLUDE_{TESTS,LABEL}_REGEX from CI workflow
* Disable patch git step in code-coverage run
* Tweak spin time of reproducible runtime
* Removed patch git step in code-coverage run
* Update ROCPROFILER_DEFAULT_FAIL_REGEX
* Mark test-counter-collection tests as unstable
- add fixtures setup/required
* Remove ATTACHED_FILES_ON_FAIL
- CDash doesn't store enable downloading these properly anyway
* Relax collection-period fuzzing window
* Disable unstable collection-period test
- too unstable
* formatting
* Disable unstable device_counting_service_test.async_counters
* Suppress perfetto internal data race errors
* Switch code-coverage CI jobs to mi300 runner
* Timeout increases
* rocprofv3-test-rocpd updates
- add fixtures
- switch executable
- redefine input/output paths
* Revert code-coverage job to mi300a runner
* Update rocprofv3-test-rocpd-execute-multiproc
- reduce problem size
* disable multiproc rocpd
* Split code-coverage into separate workflow
- network issues cause this job to fail frequently
- when in a separate workflow, it can be restarted easily
* Fixtures for rocprofv3-test-trace-hip-in-libraries
* Disable unstable device_counting_service_test.sync_counters
* Potential fix for code scanning alert no. 171: Workflow does not contain permissions
Co-authored-by: Copilot Autofix powered by AI <62310815+github-advanced-security[bot]@users.noreply.github.com>
* Switch code-coverage to run on rocprof-azure
- mi300a EMU runner set is unstable (network issues)
* tests/rocprofv3/pc-sampling SKIP_REGULAR_EXPRESSION
* Update rocprofv3-test-list-avail-trace-execute
- reduce log level and increase timeout
* rocprofv3: Prevent recursive call to rocprofv3_error_signal_handler + log chaining
* rocprofv3: Use ROCP_ERROR + std::exit instead of ROCP_FATAL
- should help with SKIP_REGULAR_EXPRESSION
---------
Co-authored-by: Copilot Autofix powered by AI <62310815+github-advanced-security[bot]@users.noreply.github.com>
[ROCm/rocprofiler-sdk commit: 640ca55ac0]
40 lines
937 B
Plaintext
40 lines
937 B
Plaintext
#
|
|
# ThreadSanitizer suppressions file for rocprofiler project.
|
|
#
|
|
|
|
# leaked thread
|
|
thread:libhsa-runtime64.so
|
|
|
|
# data race in operator delete(void*)
|
|
race:libamdhip64.so
|
|
|
|
# data race arising from hsa runtime
|
|
race:libhsa-runtime64.so
|
|
|
|
# data race(s) arising from OpenMP
|
|
race:__kmp_resume_template
|
|
race:__kmp_suspend_initialize_thread
|
|
|
|
# unlock of an unlocked mutex (or by a wrong thread)
|
|
mutex:libhsa-runtime64.so
|
|
|
|
# unlock of an unlocked mutex (or by a wrong thread)
|
|
mutex:librocm_smi64.so
|
|
|
|
# google logging
|
|
race:google::LogMessageTime::CalcGmtOffset
|
|
race:tzset_internal
|
|
|
|
# bug in libtsan.so.0 which thinks there is a
|
|
# double mutex lock (there isn't one)
|
|
mutex:external/ptl/source/PTL/TaskGroup.hh
|
|
|
|
# lock order inversion that cannot happen
|
|
mutex:source/lib/common/synchronized.hpp
|
|
|
|
# signal-unsafe function called from signal handler
|
|
signal:rocprofv3_error_signal_handler
|
|
|
|
# data race within perfetto internals
|
|
race:perfetto::internal::
|