This website requires JavaScript.
Explore
Help
Register
Sign In
biondizzle
/
vllm
Watch
1
Star
0
Fork
0
You've already forked vllm
Code
Issues
Pull Requests
Actions
2
Packages
Projects
Releases
Wiki
Activity
Files
c308501cb6a922af8c4183bd65be0094dd73de9a
vllm
/
tests
/
v1
/
tpu
History
Nick Hill
eeb135eb87
[Core] Use
CpuGpuBuffer
for block table tensors (
#24795
)
...
Signed-off-by: Nick Hill <
nhill@redhat.com
>
2025-09-16 19:18:06 -07:00
..
worker
[Core] Use
CpuGpuBuffer
for block table tensors (
#24795
)
2025-09-16 19:18:06 -07:00
__init__.py
…
test_basic.py
…
test_kv_cache_update_kernel.py
[TPU] kv cache update kernel doesn't need to be padded slices to multiple of num_slices_per_block (
#22394
)
2025-08-09 20:49:04 -07:00
test_mha_attn.py
…
test_multimodal.py
[CI/Build] Serve images used by multimodal tests through local HTTP Server (
#23907
)
2025-09-03 16:13:11 +08:00
test_pallas.py
[Attention][FlashInfer] Enable FP8 FlashInfer (TRTLLM) MLA decode (
#24705
)
2025-09-12 15:45:53 -06:00
test_perf.py
…
test_sampler.py
…
test_spmd_model_weight_loading.py
…
test_topk_topp_sampler.py
[TPU] Remove TopKTopPSampler dependency for TPU sampler (
#24391
)
2025-09-07 01:12:36 -07:00
test_tpu_int8.py
…
test_tpu_qkv_linear.py
…