Inkling-Small-276B-12B, effort "max" VS Qwen3.6-27B
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
| I saw u/danielhanchen's 1-bit Kimi K3 post: https://huggingface.co/unsloth/Kimi-K3-GGUF/discussions/12#6a6a4a90ec74ef13d85d7cf6 and decided to test Inkling-Small and Qwen3.6-27B myself, based on the full shared prompt: https://huggingface.co/unsloth/Kimi-K3-GGUF/discussions/12#6a6aba4da9b88c3996c80fa6 Inkling-Small-276B-12B, UD-Q2_K_XL, effort "max"(above "xhigh"), result:https://i.redd.it/v3hoyuo98ggh1.gif On DGX Spark GB10, it thought for 6 minutes and then started writing lots of hacky code: It then dumped the file and wrote a short summary: --- Qwen3.6-27B result:https://i.redd.it/rfcertf39ggh1.gif
Checked -> the feature was there under a different name.
[link] [comments] |
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.