<15B local models to test on medical benchmark?
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
HI, in a few weeks the medical exam for italian residents will drop, and I would like to compare small models that run on my laptop against it. It would not be a standard benchmark for llms but it would give a sense on how much better does it perform compared to the average medical resident. Do you have any reccomendation? Already last year some 32B models maxed it so i think it is sufficient to test <15B.
[link] [comments]
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.