Kimi Linear 48B A3B?
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
| Just noticed this exists, 1M context MOE with 48B par seems just like what Ive been looking for - it runs pretty damn fast too compared to Qwen 3.6 35B. after some testing it seems capable of producing *not terrible* results but it always tries to go for the minimun possible output that could be presented as what was asked. first image is the default meme prompt it generated - second is after replying with 'detail' Frontend results aren't too bad either, it seems very capable of generating well structured animated pages fairly well. but it seems like there is just something off with its brain wondering if a fine tune would be possible and if at all capable of tightning its screws, has anyone tried messing around with this one? [link] [comments] |
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.