Files

biondizzle f8a9d372e5 Use PyPI vLLM wheel instead of building (QEMU cmake try_compile fails)

- vLLM 0.18.1 aarch64 wheel includes pre-compiled FA2, FA3, MoE kernels
- Original build-from-source code commented out for GH200 restoration
- CMake compiler ABI detection fails under QEMU emulation

2026-04-03 00:05:56 +00:00

.gitignore

Updated for CUDA 13

2025-10-21 19:21:13 +00:00

Dockerfile

Use PyPI vLLM wheel instead of building (QEMU cmake try_compile fails)

2026-04-03 00:05:56 +00:00

README.md

Updated to v0.11.1rc3

2025-10-23 18:11:41 +00:00

README.md

VLLM images for GH200

Hosted here

 docker login
# Alternative
# docker buildx build --platform linux/arm64 --memory=600g -t rajesh550/gh200-vllm:0.9.0.1 .
 docker build --memory=450g --platform linux/arm64 -t rajesh550/gh200-vllm:0.11.1rc2 . 2>&1 | tee build.log 
 docker push rajesh550/gh200-vllm:0.11.1rc2