This website requires JavaScript.
Explore
Help
Sign In
xinyun
/
vllm
Watch
1
Star
0
Fork
0
You've already forked vllm
mirror of
https://git.datalinker.icu/vllm-project/vllm.git
synced
2026-04-02 23:17:04 +08:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
vllm
/
benchmarks
/
kernels
History
22quinn
4671ac6e2a
[Bugfix][Benchmark] Fix Marlin benchmark (
#19929
)
2025-06-24 07:25:12 +09:00
..
deepgemm
…
bench_fp8_gemm.py
[Benchmark] Refactor benchmark script for fp8 & int8 (
#19627
)
2025-06-15 15:15:37 +08:00
bench_int8_gemm.py
[Benchmark] Refactor benchmark script for fp8 & int8 (
#19627
)
2025-06-15 15:15:37 +08:00
benchmark_aqlm.py
…
benchmark_bitblas.py
…
benchmark_cutlass_fp4_moe.py
[Hardware][NVIDIA] FP4 MoE kernel optimization (
#19110
)
2025-06-05 09:48:26 -07:00
benchmark_grouped_gemm_cutlass.py
[Kernel] Integrate CUTLASS MoE kernel with PPLX (
#18762
)
2025-06-06 18:26:11 -07:00
benchmark_layernorm.py
…
benchmark_lora.py
…
benchmark_machete.py
…
benchmark_marlin.py
[Bugfix][Benchmark] Fix Marlin benchmark (
#19929
)
2025-06-24 07:25:12 +09:00
benchmark_moe_align_block_size.py
[Frontend] Expose custom args in OpenAI APIs (
#16862
)
2025-06-18 17:41:11 -07:00
benchmark_moe_permute_unpermute.py
…
benchmark_moe.py
[Bugfix] Fix benchmark_moe.py (
#19016
)
2025-06-09 18:04:36 -07:00
benchmark_paged_attention.py
…
benchmark_quant.py
…
benchmark_rmsnorm.py
…
benchmark_rope.py
…
benchmark_shapes.py
…
benchmark_w8a8_block_fp8.py
…
graph_machete_bench.py
…
requirements.txt
…
utils.py
…
weight_shapes.py
…