This website requires JavaScript.
Tutki
Apua
Kirjaudu sisään
AI
/
amd-strix-halo-vllm-toolboxes
Tarkkaile
2
Tähti
0
Fork
0
You've already forked amd-strix-halo-vllm-toolboxes
Koodi
Ongelmat
Pull-pyynnöt
Actions
Paketit
Projektit
Julkaisut
Wiki
Toiminta
Files
4b0918877695d8ac4260ab2db098a99b026a1318
amd-strix-halo-vllm-toolboxes
/
benchmarks
T
Historia
Donato Capitella
a1105a0b96
feat: Enhance vLLM benchmarking to compare Triton and ROCm attention, introduce a new script for cluster configuration, and update Dockerfile for new tools and dependencies.
2026-02-01 19:36:07 +00:00
..
benchmark_results
updates
2025-12-20 11:37:06 +00:00
benchmark_results_rocm_attn
/benchmark_results
added ROCm/Triton attention comparison
2025-12-20 11:49:03 +00:00
find_max_context.py
updates
2025-12-20 11:37:06 +00:00
max_context_results.json
feat: Enhance vLLM benchmarking to compare Triton and ROCm attention, introduce a new script for cluster configuration, and update Dockerfile for new tools and dependencies.
2026-02-01 19:36:07 +00:00
run_vllm_bench.py
feat: Enhance vLLM benchmarking to compare Triton and ROCm attention, introduce a new script for cluster configuration, and update Dockerfile for new tools and dependencies.
2026-02-01 19:36:07 +00:00
vllm_cluster_bench.py
feat: Enhance vLLM benchmarking to compare Triton and ROCm attention, introduce a new script for cluster configuration, and update Dockerfile for new tools and dependencies.
2026-02-01 19:36:07 +00:00