r/LocalLLaMA · · 1 min read

PR for running Ternary-Bonsai-8B-Q2_0.gguf in llama.cpp with CUDA support just got merged

Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.

Time to see what it's capable of

Join the most real place on the internet

Continue with Phone Number
Continue with Email

By continuing, you agree to our User Agreement and acknowledge that you understand the Privacy Policy.


Comments Section

Why just 8B? 27B should also be supported right?

39m ago

Mb, I meant 27B, idk what happened to my brain when posting. Doesn't matter which one though, it's Q2_0 being supported rn

Discussion (0)

Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.

Sign in →

No comments yet. Sign in and be the first to say something.

More from r/LocalLLaMA