OpenLLM-France/Luciole-23B-Instruct-1.1 (Apache 2.0 license, 8B and 1B also available)
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
| Luciole-23B-Instruct-1.1 is a fine-tuned and aligned version of Luciole-23B-Base, an open-source, multilingual causal language model created by OpenLLM-France. Luciole-23B-Instruct-1.1 was developed by LINAGORA and the OpenLLM-Franceconsortium as a part of the OpenLLM France project, funded by BPI France through the France 2030program. Training of Luciole-23B-Instruct-1.1 was conducted on Jean Zay in three phases: (i) a supervised fine-tuning (SFT) phase on instruction data with thinking traces, (ii) an SFT phase on instruction data without thinking traces, and (iii) a final preference alignment phase using Direct Preference Optimization (DPO). The training data covers topics in math, science, coding, general chat, RAG and translation. 8B and 1B also available. Collection: https://huggingface.co/collections/OpenLLM-France/luciole-llm [link] [comments] |
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.