xinyun/vllm - vllm - 丝路新云-代码仓

mirror of https://git.datalinker.icu/vllm-project/vllm.git synced 2026-05-30 22:17:14 +08:00

Author	SHA1	Message	Date
Cyrus Leung	e83b7e379c	Revert "[Renderer] Separate out `RendererConfig` from `ModelConfig` (#30145 )" (#30199 )	2025-12-07 00:00:22 -08:00
Cyrus Leung	27f4c2fd46	[Renderer] Separate out `RendererConfig` from `ModelConfig` (#30145 ) Signed-off-by: DarkLight1337 <tlleungac@connect.ust.hk>	2025-12-06 23:15:42 -08:00
Shengqi Chen	990f806473	[Doc] clarify nightly builds in developer docs (#30019 ) Signed-off-by: Shengqi Chen <harry-chen@outlook.com>	2025-12-05 00:28:37 +08:00
Finbarr Timbers	38caf7fa1a	Update FAQ on interleaving sliding windows support (#29796 ) Signed-off-by: Finbarr Timbers <finbarrtimbers@gmail.com>	2025-12-01 19:15:19 +00:00
Yifei Zhang	1ab8fc8197	Make PyTorch profiler gzip and CUDA time dump configurable (#29568 ) Signed-off-by: Yifei Zhang <yifei.zhang1992@outlook.com>	2025-12-01 04:30:46 +00:00
Cyrus Leung	ccbdf51bd5	[Doc] Reorganize benchmark docs (#29658 ) Signed-off-by: DarkLight1337 <tlleungac@connect.ust.hk>	2025-11-28 17:19:25 +08:00
rongfu.leng	480598958e	[Feature][Bench] Add pareto visualization (#29477 ) Signed-off-by: rongfu.leng <rongfu.leng@daocloud.io>	2025-11-27 23:53:20 -08:00
Matthew Bonanni	430dd4d9eb	[Attention] Remove imports from `vllm/attention/__init__.py` (#29342 ) Signed-off-by: Matthew Bonanni <mbonanni@redhat.com>	2025-11-26 10:53:15 -07:00
Roger Wang	0ff70821c9	[Core] Deprecate `xformers` (#29262 ) Signed-off-by: Roger Wang <hey@rogerw.io>	2025-11-24 04:18:55 +00:00
Cyrus Leung	aab0102a26	[V0 deprecation] Remove more V0 references (#29088 ) Signed-off-by: DarkLight1337 <tlleungac@connect.ust.hk>	2025-11-21 11:56:59 +00:00
Shanshan Shen	d44e9df7d4	[Model][Mamba] Add selector for mamba attention backend and make it pluggable for other device (#26487 ) Signed-off-by: shen-shanshan <467638484@qq.com>	2025-11-19 16:24:55 +00:00
Didier Durand	083cf326dc	[Doc]: fix typos in various files (#28863 ) Signed-off-by: Didier Durand <durand.didier@gmail.com>	2025-11-17 20:32:14 -08:00
Didier Durand	63fed55506	[Doc]: fix typos in various files (#28811 ) Signed-off-by: Didier Durand <durand.didier@gmail.com>	2025-11-16 14:30:06 +00:00
Harry Mellor	67187554dd	[Docs] Enable some more markdown lint rules for the docs (#28731 ) Signed-off-by: Harry Mellor <19981378+hmellor@users.noreply.github.com>	2025-11-14 18:39:19 +00:00
Julien Denize	085424808e	Remove audio optional dependency for mistral-common (#28722 ) Signed-off-by: Julien Denize <julien.denize@mistral.ai> Signed-off-by: Julien Denize <40604584+juliendenize@users.noreply.github.com> Co-authored-by: Cyrus Leung <cyrus.tl.leung@gmail.com> Co-authored-by: Cyrus Leung <tlleungac@connect.ust.hk>	2025-11-14 09:54:38 -08:00
Harry Mellor	5f3cd7f7f2	[Docs] Update the name of `Transformers backend` -> `Transformers modeling backend` (#28725 ) Signed-off-by: Harry Mellor <19981378+hmellor@users.noreply.github.com>	2025-11-14 16:34:14 +00:00
Harry Mellor	97d1c99302	Rename clashing method names for vLLM model protocol (#27583 ) Signed-off-by: Harry Mellor <19981378+hmellor@users.noreply.github.com>	2025-11-12 19:14:33 -08:00
Benjamin Chislett	975676d174	[Feat] Drop-in Torch CUDA Profiler (#27841 ) Signed-off-by: Benjamin Chislett <bchislett@nvidia.com>	2025-11-08 14:07:37 -08:00
Kuntai Du	8bff831f0a	[Benchmark] Cleanup deprecated nightly benchmark and adjust the docstring for performance benchmark (#25786 ) Signed-off-by: KuntaiDu <kuntai@uchicago.edu>	2025-10-30 04:43:37 +00:00
Cyrus Leung	ecca3fee76	[Frontend] Add `vllm bench sweep` to CLI (#27639 ) Signed-off-by: DarkLight1337 <tlleungac@connect.ust.hk> Signed-off-by: Harry Mellor <19981378+hmellor@users.noreply.github.com> Co-authored-by: Harry Mellor <19981378+hmellor@users.noreply.github.com>	2025-10-29 05:59:48 -07:00
Matvei Pashkovskii	130aa8cbcf	Add load pattern configuration guide to benchmarks (#26886 ) Signed-off-by: Matvei Pashkovskii <mpashkov@amd.com> Signed-off-by: Matvei Pashkovskii <matvei.pashkovskii@amd.com> Co-authored-by: Harry Mellor <19981378+hmellor@users.noreply.github.com>	2025-10-28 10:49:15 -07:00
Cyrus Leung	8fb7b2fab9	[Doc] Fix links to GH projects (#27530 ) Signed-off-by: DarkLight1337 <tlleungac@connect.ust.hk>	2025-10-26 17:55:51 +08:00
Cyrus Leung	ceacedc1f9	[Benchmark] Add plot utility for parameter sweep (#27168 ) Signed-off-by: DarkLight1337 <tlleungac@connect.ust.hk>	2025-10-21 20:30:03 -07:00
Huy Do	becb7de40b	Update PyTorch to 2.9.0+cu129 (#24994 ) Co-authored-by: Luka Govedič <ProExpertProg@users.noreply.github.com>	2025-10-21 17:20:18 -04:00
Cyrus Leung	b3aba04e5a	[Benchmark] Convenience script for multiple parameter combinations (#27085 ) Signed-off-by: DarkLight1337 <tlleungac@connect.ust.hk>	2025-10-18 23:57:01 -07:00
dongbo910220	a1946c9f61	[Chore] Separate out profiling utilities from vllm.utils (#27150 ) Signed-off-by: dongbo910220 <1275604947@qq.com>	2025-10-18 19:12:01 +00:00
Harry Mellor	483ea64611	[Docs] Replace all explicit anchors with real links (#27087 ) Signed-off-by: Harry Mellor <19981378+hmellor@users.noreply.github.com>	2025-10-17 02:22:06 -07:00
Harry Mellor	4ffd6e8942	[Docs] Reduce custom syntax used in docs (#27009 ) Signed-off-by: Harry Mellor <19981378+hmellor@users.noreply.github.com>	2025-10-16 20:05:34 -07:00
Cyrus Leung	ef9676a1f1	[Doc] ruff format some Python examples (#26767 ) Signed-off-by: DarkLight1337 <tlleungac@connect.ust.hk>	2025-10-14 03:21:53 -07:00
Maximilien de Bayser	fe3edb4cf0	Add support for the /rerank endpoint in vllm bench serve (#26602 ) Signed-off-by: Max de Bayser <mbayser@br.ibm.com>	2025-10-14 04:25:43 +00:00
Harry Mellor	8fcaaf6a16	Update `Optional[x]` -> `x \| None` and `Union[x, y]` to `x \| y` (#26633 ) Signed-off-by: Harry Mellor <19981378+hmellor@users.noreply.github.com>	2025-10-12 09:51:31 -07:00
Harry Mellor	e09d1753ec	Remove Python 3.9 support ahead of PyTorch 2.9 in v0.11.1 (#26416 ) Signed-off-by: Harry Mellor <19981378+hmellor@users.noreply.github.com>	2025-10-08 10:40:42 -07:00
Cyrus Leung	44b9af5bb2	[Benchmark] Enable MM Embedding benchmarks (#26310 ) Signed-off-by: DarkLight1337 <tlleungac@connect.ust.hk>	2025-10-06 19:51:58 +00:00
Wenlong Wang	79aa244678	[Multi Modal] Configurable MM Profiling (#25631 ) Signed-off-by: wwl2755 <wangwenlong2755@gmail.com> Signed-off-by: Harry Mellor <19981378+hmellor@users.noreply.github.com> Co-authored-by: Harry Mellor <19981378+hmellor@users.noreply.github.com>	2025-10-03 03:59:10 -07:00
Cyrus Leung	d00d652998	[CI/Build] Replace `vllm.entrypoints.openai.api_server` entrypoint with `vllm serve` command (#25967 ) Signed-off-by: DarkLight1337 <tlleungac@connect.ust.hk>	2025-10-02 10:04:57 -07:00
Naman Lalit	9bedac9623	[Doc] Add documentation for vLLM continuous benchmarking and profiling (#25819 ) Signed-off-by: Naman Lalit <nl2688@nyu.edu>	2025-09-29 20:49:49 +00:00
Jialin Ouyang	c216119d64	[Core] GC Debug callback (#24829 ) Signed-off-by: Jialin Ouyang <jialino@meta.com> Signed-off-by: Jialin Ouyang <Jialin.Ouyang@gmail.com> Co-authored-by: Jialin Ouyang <jialino@meta.com>	2025-09-27 17:53:31 +00:00
Cyrus Leung	27d7638b94	[Bugfix] Merge MM embeddings by index instead of token IDs (#16229 ) Signed-off-by: DarkLight1337 <tlleungac@connect.ust.hk> Signed-off-by: NickLucche <nlucches@redhat.com> Signed-off-by: Roger Wang <hey@rogerw.io> Co-authored-by: NickLucche <nlucches@redhat.com> Co-authored-by: Roger Wang <hey@rogerw.io>	2025-09-27 08:15:12 +00:00
vllmellm	0d9fe260dd	[docs] Benchmark Serving Incorrect Arg (#25474 ) Signed-off-by: vllmellm <vllm.ellm@embeddedllm.com>	2025-09-23 06:05:11 -07:00
Roger Wang	21da73343a	[Misc] Clean up flags in `vllm bench serve` (#25138 ) Signed-off-by: Roger Wang <hey@rogerw.io>	2025-09-18 12:43:33 +00:00
Harry Mellor	32baf1d036	[Docs] Clean up the contributing README (#25099 ) Signed-off-by: Harry Mellor <19981378+hmellor@users.noreply.github.com>	2025-09-17 21:05:18 -07:00
yyzxw	5672ba90bd	[Docs] fix invalid doc link (#25017 ) Signed-off-by: zxw <1020938856@qq.com>	2025-09-16 20:53:23 -07:00
Isotr0py	5a411ef6c4	[Benchmarks] Add MMVU video dataset support and clean up deprecated datasets (#24719 ) Signed-off-by: Isotr0py <mozf@mail2.sysu.edu.cn>	2025-09-17 03:29:43 +00:00
elvischenv	3059b9cc6b	[Doc] Add --force-overwrite option to generate_cmake_presets.py (#24375 ) Signed-off-by: elvischenv <219235043+elvischenv@users.noreply.github.com>	2025-09-16 18:45:29 -07:00
Ye (Charlotte) Qi	85e0df1392	[Docs] move benchmarks README to contributing guides (#24820 )	2025-09-16 05:52:57 -07:00
Woosuk Kwon	759ef49b15	Remove V0 Encoder-Decoder Support (#24907 ) Signed-off-by: Woosuk Kwon <woosuk@thinkingmachines.ai>	2025-09-15 21:17:14 -07:00
Harry Mellor	361ae27f8a	[Docs] Fix formatting of transcription doc (#24676 ) Signed-off-by: Harry Mellor <19981378+hmellor@users.noreply.github.com>	2025-09-11 11:18:06 -07:00
Wentao Ye	4984a291d5	[Doc] Fix Markdown Pre-commit Error (#24670 ) Signed-off-by: yewentao256 <zhyanwentao@126.com>	2025-09-11 09:05:59 -07:00
Nicolò Lucchesi	404c85ca72	[Docs] Add transcription support to model (#24664 ) Signed-off-by: NickLucche <nlucches@redhat.com>	2025-09-11 07:39:01 -07:00
Michael Yao	2f0b833a05	[Docs] Fix a tip indentation and typo (#24419 ) Signed-off-by: windsonsea <haifeng.yao@daocloud.io>	2025-09-08 00:19:40 -07:00

1 2 3

102 Commits