r/MachineLearning · · 1 min read

Cloud-vLLM Benchmark Differences [R]

Mirrored from r/MachineLearning for archival readability. Support the source by reading on the original site.

Does anyone know of any evidence/forum/paper analyzing benchmark result differences between cloud inference platforms (togetherai) and running models locally with vLLM under greedy decoding?

submitted by /u/No_Cardiologist7609
[link] [comments]

Discussion (0)

Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.

Sign in →

No comments yet. Sign in and be the first to say something.

More from r/MachineLearning