Απαιτείται JavaScript για να εμφανιστεί αυτή η ιστοσελίδα.
Εξερεύνηση
Βοήθεια
Είσοδος
AI
/
amd-strix-halo-vllm-toolboxes
Παρακολούθηση
2
Αστέρι
0
Fork
0
Έχετε ήδη κάνει fork το amd-strix-halo-vllm-toolboxes
Κώδικας
Ζητήματα
Pull Requests
Δράσεις
Πακέτα
Έργα
Κυκλοφορίες
Wiki
Δραστηριότητα
59
Υποβολές
3
Κλάδοι
0
Ετικέτες
backup-before-cleanup
Γράφημα Υποβολών
6 Υποβολές
Συγγραφέας
SHA1
Μήνυμα
Ημερομηνία
Donato Capitella
1f96c391fb
feat: Add comprehensive RDMA cluster setup guide, enforce eager mode in cluster benchmarks, and update documentation with cluster details.
2026-02-02 19:34:33 +00:00
Donato Capitella
c587981d73
refactor: Centralize Ray/vLLM cluster management into a new
cluster_manager.py
module and refactor
start_vllm_cluster.py
to use it.
2026-02-01 22:19:34 +00:00
Donato Capitella
128ddade14
fix: improve RDMA stability by configuring NCCL IB timeout and retry count.
2026-02-01 22:04:34 +00:00
Donato Capitella
0d8afba093
feat: Add
RAY_DISABLE_METRICS=1
to disable Ray metrics across cluster configurations and scripts.
2026-02-01 21:52:48 +00:00
Donato Capitella
ba503f6e61
feat: centralize model configurations and benchmark settings into a new
models.py
module and update Dockerfile and scripts to use it.
2026-02-01 21:17:15 +00:00
Donato Capitella
e5cc96bf48
feat: Introduce vLLM cluster benchmarking and setup scripts, and expand the list of models for local benchmarks.
2026-02-01 15:43:56 +00:00