r/LocalLLaMA · · 1 min read

The best model is the one you can actually run

Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.

The best model is the one you can actually run

Don't get me wrong, all the big models are amazing, and every contribution to open source models is great. But I'm GPU poor and I can't use them locally.

I'm currently running gemma-4-12b-it-qat-GGUF:UD-Q4_K_XL as my personal chat assistant, and I am so so happy with it! I still can't believe I can talk to my computer.

submitted by /u/OneFanFare
[link] [comments]

Discussion (0)

Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.

Sign in →

No comments yet. Sign in and be the first to say something.

More from r/LocalLLaMA