Ce site Web nécessite JavaScript.
Explorateur
Aide
Connexion
AI
/
amd-strix-halo-vllm-toolboxes
Suivre
2
Ajouter aux favoris
0
Bifurcation
0
Vous avez déjà forké amd-strix-halo-vllm-toolboxes
Code
Tickets
Demandes d'ajout
Actions
Paquets
Projets
Publications
Wiki
Activité
54
Révisions
3
Branches
0
Étiquette
693757f5d945bb9f28949ea52ce76c7ea3cf1e42
Graphe des révisions
5 Révisions
Auteur
SHA1
Message
Date
Donato Capitella
4d3b046870
feat: Add new benchmark results for various models and configurations, and update documentation UI with filtering for attention and tensor parallelism.
2026-02-02 21:30:17 +00:00
Donato Capitella
1f96c391fb
feat: Add comprehensive RDMA cluster setup guide, enforce eager mode in cluster benchmarks, and update documentation with cluster details.
2026-02-02 19:34:33 +00:00
Donato Capitella
a1105a0b96
feat: Enhance vLLM benchmarking to compare Triton and ROCm attention, introduce a new script for cluster configuration, and update Dockerfile for new tools and dependencies.
2026-02-01 19:36:07 +00:00
Donato Capitella
711de530f6
added ROCm/Triton attention comparison
2025-12-20 11:49:03 +00:00
Donato Capitella
5e8b6bb545
updates
2025-12-20 11:37:06 +00:00