r/LocalLLaMA · · 1 min read

is switching from llama cpp to vllm worth it

Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.

I have hp z8 g4 with 512 ram and 1x3090 1x5060 16gb. has anyone made the transition from llama cpp to vllm recently? is it worth it? docker under windows or full linux install?

I am mainly interested in the model support, it seems that many new local models are supported day 0 in official vllm, while for llama cpp it takes months sometimes.

submitted by /u/Exciting-Engine882
[link] [comments]

Discussion (0)

Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.

Sign in →

No comments yet. Sign in and be the first to say something.

More from r/LocalLLaMA