r/LocalLLaMA · · 1 min read

I compared local models and different quants / config on a subset of swe-verified bench

Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.

I compared local models and different quants / config on a subset of swe-verified bench

And gathered a lot of data. you can see them for yourself

And For the most curious, there are additional details here

In this graph, I regrouped the finetunes under their base models. but you can see the details in the page.

The python code to generate those pages is obviously vibecoded. I find the output kinda pretty and somewhat useful for me. maybe it's useful for someone else.

Heading for a vacation for a few weeks, but if you have any suggestion, I will consider each of them.

submitted by /u/WonderRico
[link] [comments]

Discussion (0)

Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.

Sign in →

No comments yet. Sign in and be the first to say something.

More from r/LocalLLaMA