Для этого сайта требуется поддержка JavaScript.
Обзор
Помощь
Вход
AI
/
amd-strix-halo-vllm-toolboxes
Следить
2
В избранное
0
Форкнуть
0
Вы уже форкнули amd-strix-halo-vllm-toolboxes
Код
Задачи
Запросы на слияние
Действия
Пакеты
Проекты
Релизы
Вики
Активность
59
коммитов
3
Ветки
0
Теги
backup-before-cleanup
Граф коммитов
6 Коммитов
Автор
SHA1
Сообщение
Дата
Donato Capitella
1f96c391fb
feat: Add comprehensive RDMA cluster setup guide, enforce eager mode in cluster benchmarks, and update documentation with cluster details.
2026-02-02 19:34:33 +00:00
Donato Capitella
c587981d73
refactor: Centralize Ray/vLLM cluster management into a new
cluster_manager.py
module and refactor
start_vllm_cluster.py
to use it.
2026-02-01 22:19:34 +00:00
Donato Capitella
128ddade14
fix: improve RDMA stability by configuring NCCL IB timeout and retry count.
2026-02-01 22:04:34 +00:00
Donato Capitella
0d8afba093
feat: Add
RAY_DISABLE_METRICS=1
to disable Ray metrics across cluster configurations and scripts.
2026-02-01 21:52:48 +00:00
Donato Capitella
ba503f6e61
feat: centralize model configurations and benchmark settings into a new
models.py
module and update Dockerfile and scripts to use it.
2026-02-01 21:17:15 +00:00
Donato Capitella
e5cc96bf48
feat: Introduce vLLM cluster benchmarking and setup scripts, and expand the list of models for local benchmarks.
2026-02-01 15:43:56 +00:00