mirror of
https://git.datalinker.icu/vllm-project/vllm.git
synced 2026-08-12 09:47:52 +08:00
[Doc] Add vllm-metal to hardware plugin documentation (#31174)
Signed-off-by: mgoin <mgoin64@gmail.com>
This commit is contained in:
parent
b10f41c894
commit
9586354053
@ -28,3 +28,4 @@ The backends below live **outside** the main `vllm` repository and follow the
|
|||||||
| Cambricon MLU | `vllm-mlu` | <https://github.com/Cambricon/vllm-mlu> |
|
| Cambricon MLU | `vllm-mlu` | <https://github.com/Cambricon/vllm-mlu> |
|
||||||
| Baidu Kunlun XPU | N/A, install from source | <https://github.com/baidu/vLLM-Kunlun> |
|
| Baidu Kunlun XPU | N/A, install from source | <https://github.com/baidu/vLLM-Kunlun> |
|
||||||
| Sophgo TPU | N/A, install from source | <https://github.com/sophgo/vllm-tpu> |
|
| Sophgo TPU | N/A, install from source | <https://github.com/sophgo/vllm-tpu> |
|
||||||
|
| Apple Silicon (Metal) | N/A, install from source | <https://github.com/vllm-project/vllm-metal> |
|
||||||
|
|||||||
@ -4,6 +4,9 @@ vLLM has experimental support for macOS with Apple Silicon. For now, users must
|
|||||||
|
|
||||||
Currently the CPU implementation for macOS supports FP32 and FP16 datatypes.
|
Currently the CPU implementation for macOS supports FP32 and FP16 datatypes.
|
||||||
|
|
||||||
|
!!! tip "GPU-Accelerated Inference with vLLM-Metal"
|
||||||
|
For GPU-accelerated inference on Apple Silicon using Metal, check out [vllm-metal](https://github.com/vllm-project/vllm-metal), a community-maintained hardware plugin that uses MLX as the compute backend.
|
||||||
|
|
||||||
# --8<-- [end:installation]
|
# --8<-- [end:installation]
|
||||||
# --8<-- [start:requirements]
|
# --8<-- [start:requirements]
|
||||||
|
|
||||||
|
|||||||
Loading…
x
Reference in New Issue
Block a user