News / #open-source Tag Open source 500 articles archived under #open-source · RSS Sign in to follow r/LocalLLaMA community 17d ago Anthropic is calling for a ban on open-weights models by proposing mandatory requirements they will probably never be able to meet   submitted by   /u/realmvp77 [link]   [comments] 9 Simon Willison community 17d ago moonshotai/Kimi-K3 moonshotai/Kimi-K3 As promised earlier this month , Moonshot have released the weights for their excellent 2.8 trillion parameter Kimi K3. They're a hefty 1.56TB on Hugging Face. Kimi introduced their own janky modified version of the MIT license with K2 back in July 2025. That… 28 r/LocalLLaMA community 17d ago Our position on open-weights models   submitted by   /u/RhubarbSimilar1683 [link]   [comments] 8 Hacker News — AI on Front Page community 17d ago Our position on open-weights models Article URL: https://www.anthropic.com/news/position-open-weights-models Comments URL: https://news.ycombinator.com/item?id=49076057 Points: 313 # Comments: 373 26 Hugging Face Daily Papers research 17d ago Interactive Training 2: Auditable Control Plane for Live Model Training Abstract Experiment trackers show how training is progressing, but changing a live run still usually requires trainer-specific code. We present Interactive Training 2, an open-source control plane for steering training through a shared protocol. Training applications declare… 26 NVIDIA Developer Blog official-blog 17d ago NVIDIA Ising Enables Fully Automated Quantum Computer Calibration with Enhanced In-Context Learning NVIDIA Ising Calibration is an open source vision language model (VLM) designed to interpret diagnostic outputs from quantum processors and determine how they... 9 r/LocalLLaMA community 17d ago Nvidia CEO Jensen Huang defends Open Source AI by saying distillation is fundamental to learning Nvidia CEO Jensen Huang “Distillation - learning from AI, learning from other people, and learning from other sources of knowledge, is fundamental to intelligence. We are constantly learning from one another. AI also has to learn from something.” Since using AI Desktop 98 , I… 23 r/LocalLLaMA community 18d ago The entire tech industry (save for Anthropic) has come out in favor of open source AI. So what happens next? Will Anthropic change its lobbying efforts? Not likely. Now the gaslighting begins: “Nobody is trying to ban open source.”   submitted by   /u/SignificantLegs [link]   [comments] 34 r/LocalLLaMA community 18d ago Meta has confirmed that it will release an open source model in the future https://preview.redd.it/k97l56d8ypfh1.png?width=606&format=png&auto=webp&s=2a2e2156bea56b25f5709e8f1df2bf82525eb089 https://x.com/alexandr_wang/status/2081501627836661928?s=20   submitted by   /u/External_Mood4719 [link]   [comments] 38 Smol AI News news-outlet 18d ago not much happened today **Moonshot** released the **Kimi K3** open-weights model, a **2.8T-parameter MoE** with **104B active parameters**, **896 experts**, and **1M-token context** featuring native visual understanding. The release includes open-source infrastructure like **FlashKDA**, **MoonEP**, and… 33 r/LocalLLaMA community 18d ago [OSS] Use case only possible with local inference at its core: an on-device LLM understands your entire life, then proactively offers to get your work done through computer use! Open-source & free :D Hey r/LocalLLaMA ! :D I wanna share a really cool fully OSS thing I've been building that's only possible with local models: truly proactive AI! All your existing LLM systems waits for a prompt. Truly proactive AI has to read your entire life, every single day (every file,… 19 r/LocalLLaMA community 18d ago MiniMax (official) on X: "Open weights. Open research. Open innovation.🫶 Marching for an open future.🤍   submitted by   /u/RhubarbSimilar1683 [link]   [comments] 26 r/LocalLLaMA community 18d ago Sources: OpenAI and Anthropic quietly lobby Washington regulators to restrict open-source AI models, even as Sam Altman publicly says he supports open source AI   submitted by   /u/pscoutou [link]   [comments] 15 r/LocalLLaMA community 18d ago Kimi K3 gets open weighted tomorrow! Kimi K3 is supposed to get open weighted tomorrow! Can't run it or even a model a hundred times smaller lol, but its still a great win for open source. For me, personally im more awaited for the new inference providers that will open up hopefully.… 14 r/LocalLLaMA community 19d ago We open-sourced Logue — a privacy-first macOS meeting-notes + writing app that runs on-device (MLX, Apple Silicon) entirely At Bitwize, we've been building Logue, a native macOS app for AI meeting notes and writing, and we just open-sourced it (MIT). We're sharing it here because the whole point is that it runs 100% on-device — we wanted something that could transcribe and summarize meetings without… 27 r/LocalLLaMA community 19d ago Karparthy removed Anthropic from his bio Andrej Karpathy, a prominent advocate for open-source AI and a co-founder of OpenAI, appears to have removed Anthropic from his X bio, suggesting he may have left the company. Karpathy joined Anthropic only a few months ago, making the apparent departure somewhat surprising.… 36 r/LocalLLaMA community 19d ago Mobile Offline LLMs: What do you use them for? I've spent the last year or so playing around with open source MLX and GGUF models on iPhone hardware. Given the limitations in memory, GPU/CPU/ANE, and in turn the context window I've been trying to figure out the best use cases for them. I've also done a lot of testing with… 13 r/LocalLLaMA community 19d ago Benchmarks: TensorSharp vs. llama.cpp Cuda and Vulkan Benchmark: TensorSharp vs. llama.cpp I would like to share my latest open source local Unsloth (GGUF) LLM inference engine and applications. It supports many models from Unsloth, like Gemma4, DiffusionGemma, Qwen3.6 with multi-modal (image, vision, audio), Qwen… 38 r/LocalLLaMA community 20d ago DKV: Open-source KV-cache compression framework for local LLM inference (CLI + technical report) Hi everyone! Over the past five months I've been working on DKV (DifferentialKV), an open-source project exploring KV-cache compression for long-context local LLM inference. The goal is to reduce KV-cache memory requirements through anchor-based representations, joint low-rank… 21 r/LocalLLaMA community 20d ago AMD Instella-MoE-16B-A3B https://huggingface.co/amd/Instella-MoE-16B-A3B-Think I was browsing HuggingFace and came across this model apparently uploaded a day ago, and thought to share it here. I've not tried it out yet, but it's good to see AMD joining the open source model game.   submitted by… 18 r/LocalLLaMA community 20d ago What i think the forseeable for open models will be The realization hit when Qwen announced they'll be releasing 3.8 as open weights at >2t weights. I think China strategically released open models in multiple phases, and they're now in their final phase. Phase 1: Dump small but very capable open models (especially during the… 26 r/LocalLLaMA community 20d ago Is corruption the lobbying against Open weights? Like, reading things like Anthropic "donated" to some people with the condition of lobbying against Chinese LLMs.. it's that right? It feels nothing like freedom but at the same time it's said "out loud"? I'm not from USA so I'm not very familiar with that..it's normal? allowed?… 6 r/LocalLLaMA community 20d ago Zagreus-0.4B-por a small open source language model for Portuguese mii-llm , an open source AI lab, released Zagreus-0.4B-por , a compact bilingual Portuguese–English language model pretrained entirely from scratch. The model has approximately 400 million parameters and is part of the Zagreus family, an ongoing experiment in building small open… 19 r/LocalLLaMA community 20d ago More than 20 companies including NVIDIA, Meta, Microsoft, Palantir, and Hugging Face have signed a letter urging policymakers to avoid premature restrictions on open weight models. The Open Letter was initiated by Microsoft and published today: “ Open Weights and American AI Leadership ”. It argues against broad or premature restrictions on open-weight models and explicitly says policymakers should distinguish legitimate model distillation from… 5 Hacker News — AI on Front Page community 20d ago Nvidia, Microsoft, Meta warn against overregulating open-weight models Letter: https://images.nvidia.com/pdf/Open-Weights-and-American-AI-L... [pdf] https://x.com/JensenHuang/status/2080643682408321103 , https://xcancel.com/JensenHuang/status/2080643682408321103 https://www.wired.com/story/silicon-valley-is-completely-div... ,… 13 r/MachineLearning community 20d ago I built an open-source multi-agent SDLC harness that beats a cold Claude Code run on large repos, by learning the repo once. Real benchmarks (incl. where it loses) inside. [P] Built an open-source AI coding agent that was 7%–75% cheaper than a cold "claude -p" run on 6/6 well-localized tasks across repositories up to ~82k LOC. The biggest difference: Cold agent: $6.83, 207 turns AutoDev Studio: ~$1.70 for the same bug The full benchmark (including… 14 arXiv — NLP / Computation & Language research 21d ago Making Open-Source Text LLM Watermarks Durable Against Merging arXiv:2607.20435v1 Announce Type: new Abstract: Open-source LLMs (OSMs)arereaching near state-of-the-art performance, prompting prior works to trace the text they generate by embedding text watermarking algorithms directly into their weights. Yet, OSMs are subject to… 5 arXiv — NLP / Computation & Language research 21d ago Naver-News-KO: A Korean News Summarization Dataset for Open-Source Fine-Tuning of Summarization Models arXiv:2607.20442v1 Announce Type: new Abstract: We release Naver-News-KO, a Korean news summarization dataset of 27,400 (document, summary) pairs collected from Naver News over a ten-day window in July 2022 across two categories (Economy and IT/Science; 77/23 split), with… 8 arXiv — NLP / Computation & Language research 21d ago Domyn-Small: A European 10B Reasoning Language Model arXiv:2607.20448v1 Announce Type: new Abstract: We introduce Domyn-Small, a 10-billion-parameter open-weight reasoning language model released under the MIT license. Domyn-Small is the product of an initial pre-training phase on 9 trillion tokens multilingual data, followed by a… 32 Hacker News — AI on Front Page community 21d ago The arguments against open source AI are bad Article URL: https://tombedor.dev/arguments-against-open-source-ai-are-very-bad/ Comments URL: https://news.ycombinator.com/item?id=49024643 Points: 203 # Comments: 144 20 r/LocalLLaMA community 22d ago Arcee AI has spoken out against the ban on open Chinese models in US This is rather counterintuitive, since banning Chinese models would benefit them the most. Jensen Huang is also against the ban , although the interests here are more obvious. Do you think that if Arcee, Cohere or Mistral release an open source GPT/Claude level model, they will… 20 r/LocalLLaMA community 22d ago AI9Stars released G9v3-3B AI9Stars has released G9v3-3B an open weights language model designed to deliver strong reasoning capabilities within a lightweight 3 billion parameter size. It is released under the Apache 2.0 license making it fully open for personal and commercial use The best use case is a… 17 Hugging Face Daily Papers research 22d ago G-MAD: A Game-Based Data Generation Framework for Multi-View RGB-T Aerial Object Detection Abstract This work introduces G-MAD, an open-source framework that uses Arma3 to generate synchronized multi-view RGB-T data for aerial object detection. G-MAD addresses key limitations of real-world aerial dataset construction, including limited viewpoint control, imperfect… 20 Latent.Space news-outlet 22d ago Inside the Model Factory — Eiso Kant, Poolside AI Poolside's co-CEO on how his small team of top researchers built a model factory capable of training Laguna S - a 118B MOE beating Thinky's ~1T open weights model... and this is just the beginning. 24 arXiv — Machine Learning research 22d ago SCPP: A Unified Python Library for Soft Clustering arXiv:2607.19620v1 Announce Type: new Abstract: In this paper, we present SCPP (Soft Clustering Python Package), an open-source Python framework for soft clustering. SCPP establishes a canonical, scikit-learn-compatible estimator interface that standardizes model training,… 8 arXiv — NLP / Computation & Language research 22d ago Enhancing LLMs for Identifying and Prioritizing Important Medical Jargons from Electronic Health Record Notes Utilizing Data Augmentation: A Comparative Study arXiv:2502.16022v3 Announce Type: replace Abstract: OpenNotes gives patients access to their EHR notes, but dense medical jargon limits comprehension. We evaluate closed-source and open-source LLMs for extracting and prioritizing the jargon terms most relevant to individual… 14 Simon Willison community 22d ago Quoting Thomas Ptacek I genuinely believe that if you took an open weights model from 2025 and built a pentest harness for it, it could do this kind of sandbox escape and scan/hack in most networks. This is only surprising because you assume OpenAI has sounder sandboxes. — Thomas Ptacek ,… 12 r/LocalLLaMA community 22d ago Sanctions on Open Source. hope they don’t do anything stupid here.   submitted by   /u/MLExpert000 [link]   [comments] 34 TechCrunch — AI news-outlet 22d ago Arcee, a US open source AI lab, says Chinese models are not inherently dangerous As Chinese AI models grow in capability and popularity among US companies, the arguing over what should be done about them has reached a fever pitch. 23 r/LocalLLaMA community 22d ago We built NeuTTS-2E, an open-source on-device TTS model with 7 controllable emotions We’re open sourcing an alpha release of NeuTTS-2E : an on-device TTS model with 125M active parameters and 7 controllable emotions. The goal was simple: when you select “angry,” “fearful,” or “happy,” the delivery should follow that instruction rather than whatever emotion the… 14 Hugging Face Daily Papers research 23d ago AgentDebugX: An Open-Source Toolkit for Failure Observability, Attribution, and Recovery in LLM Agents Abstract LLM agent failures are difficult to debug because the step where an error surfaces is often not the one that caused it. Existing observability tools replay execution traces but provide little support for identifying the root cause or translating diagnosis into recovery.… 16 arXiv — Machine Learning research 23d ago Uncertainty Quantification for AI-Driven Crash Simulation Surrogates: A Comparative Study of Monte Carlo Dropout and Deep Ensemble on Open-Source Bumper Beam Benchmark arXiv:2607.18294v1 Announce Type: new Abstract: Machine learning surrogate models are increasingly being explored in engineering product development to augment simulation-driven design, offering near-instantaneous predictions that complement computationally expensive… 11 arXiv — Machine Learning research 23d ago Interactive Training 2: Auditable Control Plane for Live Model Training arXiv:2607.18314v1 Announce Type: new Abstract: Experiment trackers show how training is progressing, but changing a live run still usually requires trainer-specific code. We present Interactive Training 2, an open-source control plane for steering training through a shared… 31 arXiv — NLP / Computation & Language research 23d ago AgentDebugX: An Open-Source Toolkit for Failure Observability, Attribution, and Recovery in LLM Agents arXiv:2607.18754v1 Announce Type: cross Abstract: LLM agent failures are difficult to debug because the step where an error surfaces is often not the one that caused it. Existing observability tools replay execution traces but provide little support for identifying the root… 4 r/LocalLLaMA community 23d ago Gigatoken: A new open source tokenizer ~100x faster than Tiktoken, -500-1000x faster than Huggingface   submitted by   /u/Thrumpwart [link]   [comments] 31 r/LocalLLaMA community 23d ago There is no need to worry about Trump banning China‘s open source model at all. I’ve seen many people worrying that if Trump moves to block Chinese AI models, aggregators like OpenRouter will no longer be able to host them. As someone from China, let me reassure you: there is really no need for such concern. China has been locked in trade conflicts with… 7 r/LocalLLaMA community 23d ago CEO of Hugging Face: Banning open-source AI would hurt defenders 10x more than attackers, which would make the world 10x more dangerous and this is a good example why! From clem 🤗 on 𝕏: https://x.com/ClementDelangue/status/2079301434357456931 Fortune: Hugging Face says it resorted to a Chinese AI model to battle a fully autonomous cyberattack because U.S. model guardrails stymied its defense:… 19 r/LocalLLaMA community 24d ago US gov't lobbied by major US labs is about to ban open source models.   submitted by   /u/FlowCritikal [link]   [comments] 8 arXiv — NLP / Computation & Language research 24d ago OpenLanguageModel: Readable and Composable Small-Language-Model Pretraining for Education and Research arXiv:2607.16669v1 Announce Type: new Abstract: OpenLanguageModel (OLM) is an open-source PyTorch library for building and pretraining small language models while keeping their machinery visible. In OLM, model code reads like the architecture: components are ordinary modules,… 29 r/MachineLearning community 24d ago Tri-Net v2: Open-source implementation of our Scientific Reports paper on unified skin lesion and symptom-based monkeypox detection [R] Hi everyone, We've open-sourced Tri-Net v2, the official implementation accompanying our recently published Scientific Reports (Nature Portfolio) paper: "Tri-Net: Unified Deep Learning for Skin Lesion and Symptom-Based Monkeypox Detection" Rather than releasing only training… 5 Page 3 of 10 · 500 articles ← Newer Older →