r/LocalLLaMA
500 articles archived · Visit source ↗ · RSS
-
-
r/LocalLLaMA community 1d ago
Qwen3.8-27B IQ3_XXS vs Qwen3.6-35B-A3B Q4_K_M
Which one is better for difficult tasks like web scrapping, coding, using tools? Looking for any benchmarks because i couldn't actually find one after quite some digging   submitted by   /u/Loose_Doubt367 [link]   [comments]
30 -
-
-
r/LocalLLaMA community 2d ago
Swift 1.5 27b: Swift Qwen just got faster
Enjoy! Fucking loving it.   submitted by   /u/sleight42 [link]   [comments]
11 -
-
-
r/LocalLLaMA community 2d ago
Custom Models in Oh My Pi: vLLM, llama.cpp, SGLang and More
  submitted by   /u/bolts98 [link]   [comments]
27 -
-
-
-
-
r/LocalLLaMA community 2d ago
Is there a lightweight version of Hermes agent?
I have limited Context (usually around 64k) For local use I don’t only do coding But also want like a personal assistant with memory and such. What is the best option?   submitted by   /u/Adventurous-Gold6413 [link]   [comments]
35 -
-
r/LocalLLaMA community 2d ago
Ion v0.2.0 — No install. No backend. Just one HTML file.
The new Ion is available, a harness that run directly from a single HTML file, no install or backend required: is now more capable, more customizable, and has better tools!   submitted by   /u/fredconex [link]   [comments]
26 -
-
-
-
r/LocalLLaMA community 2d ago
muse charm but local and free
  submitted by   /u/llo7d [link]   [comments]
38 -
-
-
-