Perplexity open-sourced their Mac inference server for Qwen 3.6
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
| Here is link to repo: https://github.com/perplexityai/pplx-garden/tree/main/lily It's optimized for just one model to get best perf on apple silicon [link] [comments] |
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.