This website requires JavaScript.
Explore
Help
Sign In
AI
/
amd-strix-halo-vllm-toolboxes
Watch
2
Star
0
Fork
0
You've already forked amd-strix-halo-vllm-toolboxes
Code
Issues
Pull Requests
Actions
Packages
Projects
Releases
Wiki
Activity
57
Commits
3
Branches
0
Tags
fde8f520d9bc6ed34d48ae959285bc64e8c8c693
Commit Graph
5 Commits
Author
SHA1
Message
Date
Donato Capitella
1ddcb9a202
feat: Configure ROCm attention via
--attention-backend
CLI argument, disable the Ray dashboard, and make eager mode configurable for cluster benchmarks.
2026-02-02 15:40:16 +00:00
Donato Capitella
ba503f6e61
feat: centralize model configurations and benchmark settings into a new
models.py
module and update Dockerfile and scripts to use it.
2026-02-01 21:17:15 +00:00
Donato Capitella
039484a41e
Updated name of card
2025-12-24 08:13:34 +00:00
Donato Capitella
3b0e736c94
feat: Implement dynamic model discovery from benchmark results, add benchmark notes, and include
dialog
dependency.
2025-12-20 12:31:20 +00:00
Donato Capitella
5e8b6bb545
updates
2025-12-20 11:37:06 +00:00