Swift 1.5 Qwen3.8 27b (A must-have for low thinking!)
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
Just made this post for those who missed it : https://huggingface.co/ukisai/Swift-1.5-Qwen3.8-27b
UkisAI released their updated Qwen 27B (tuned for token efficiency). I grabbed the IQ4_XS quant to test against Unsloth's Q4_K_S:
Low-thinking: UkisAI consistently beat Unsloth in most of my tests.
High-thinking: Unsloth still pulled ahead here.
I was struggling with a custom script in Directory Opus. I gave it to Gemini Flash (medium thinking on Antigravity free tier) it looped for 40 minutes, tried many things, burnt all the weekly limit-tokens, and failed to solve it.
Fed the exact same problem to this 27B model: Fixed it completely in 6 minutes on an old 3090 (67 t/s)
Honestly i was kinda impressed, didn't expect an IQ4_XS quant of a 27B model in low thinking to beat a major cloud model.
[link] [comments]
More from r/LocalLLaMA
-
NVIDIA shipped OpenShell, an open source sandbox that gives local and open agents real runtime limits instead of prompt rules. Over 100 firms joined the safety stack. OpenAI did not.
Sep 28
-
3090 for $1500???
Sep 28
-
modified qwen 3.8 27b modifies windows credential dumper to bypass EDR detection
Sep 28
-
Minisforum MS-S1 MAX-P495 @ €7.799,00
Sep 28
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.