A community project to build open source models
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
This is probably a dumb question, but I am going to ask it anyway.
In light of open source models getting so popular, the chance any of us have to train a large frontier sized model on our own, at least for most of us, just isn't possible.
But what if we created a skill, plug-in, some sort of tool, built into maybe in Pi, aider or OMP, that allowed for community data to be gained and stored after each session? Not sure if anyone is doing this already but if not I think it would be a fun project for the community to look into.
I was thinking maybe we could build a bunch of small models, 3b-14b. From reliability models to testing models, tracing, critic models. Just workflow-specific verification models. We have benchmarks, trajectory datasets, maybe we build on top of those and build reliability models that are more focused on things like:
semantic drift
architecture/spec drift
traceback disregards
dependency drift
bad rollback behavior
etc.
So we just create a more specific taxonomy on top of the already existing work like SWE-Bench, SWE-gym, Openhands and other trajectories and in combination with our own more defined schema, a more specific taxonomy
Just freestyling, let me know why it wouldn't work, or if there are already initiatives out there that exist.
[link] [comments]
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.