News / #model-release Tag Model releases 500 articles archived under #model-release · RSS Sign in to follow r/LocalLLaMA community 3d ago I built an open-weight alternative to Jev / TypeSafe - introducing OpenJudgement-4B (early preview) I’m releasing OpenJudgement-4B-Preview, an experimental Qwen-based model fine-tuned on custom datasets for classification, scoring and true/false judgments. It scores answer options directly, and Python formats the results into JSON with probabilities. It still uses an LLM… 30 r/LocalLLaMA community 3d ago R9V Update: Created and adopted KVA projections based on Deepseek V4.1 Flash + HySparse2/MiMo-V3 for Qwen3.8 Flash Next. This is a game changer for models that don't natively implement it. 1.45-1.85x speedup in prefill to 3k+ at a small deficit to perplexity. [2x R9700, 128GB… Here's my *first* implementation of KVA projectors on QFN (just the uncensored model for now) the highlights are basically as follows for using the projectors at each different layer: Starting at layer 12, prompt processing speeds up 1.85x [1700 t/s -> 3150 t/s] at the tradeoff… 24 The Information — AI news-outlet 3d ago Google, OpenAI and Anthropic AI Safety Group Takes Shape Google , OpenAI and Anthropic are pushing forward with a plan to create a new AI safety-focused standards body on their own, without government oversight, in hopes of launching it by the end of the year or early in 2027, according to people familiar with the matter. The three… 14 r/LocalLLaMA community 3d ago Qwen-3.8-27B is good enough that I stopped using API Many a praise have been sung on Qwen-3.8, but here is mine. Qwen-3.8 and I had a rocky start, because it thinks so much. Watching it working is painful, so you have to stop doing that. You have to let it work unsupervised. And that's okay, because it really is able to complete… 20 OpenAI official-blog 3d ago Ringg’s AI agents resolve up to 65% of customer calls with OpenAI Using GPT-5.6, Ringg powers multilingual agents across voice, chat, WhatsApp, and web for 90% less cost vs. GPT-4.1. 35 LangChain releases dev-tools 4d ago langchain-openai==1.6.6 Changes since langchain-openai==1.6.5 release(openai): 1.6.6 ( #40800 ) fix(openai): raise on error events in stream path ( #40791 ) 8 The Information — AI news-outlet 4d ago DeepSeek Aims to Close $7.5 Billion Funding Round by End-October DeepSeek is aiming to complete its second funding round by the end of October as it prepares to go public on the Shanghai Stock Exchange, The Information reported on Thursday. The company’s fundraising efforts may be fuelled by a boost in revenue, which has hit $1 billion on an… 27 r/MachineLearning community 4d ago arXiv receives Multiyear Philanthropic Commitments to Support Its Launch as an Independent Nonprofit [N] $17.2 million investment, spanning three to five years from Simons Foundation International , XTX Markets , and Siegel Family Endowment #Philanthropy) https://blog.arxiv.org/2026/09/23/arxiv-receives-multiyear-investment/   submitted by   /u/Nunki08 [link]  … 16 Vercel — AI dev-tools 4d ago The Vercel Bug Bounty Program is now publicly available In 2022, we launched a private bug bounty program through HackerOne. Over the last several years, we've worked with HackerOne's VIP program to refine our process, targets, and scope, onboarding thousands of their best researchers and hardening security across our platform. We've… 12 r/LocalLLaMA community 4d ago JEV almost dead: CLM vs JEV Original post: https://www.reddit.com/r/LocalLLaMA/comments/1woscea/contrastive_language_models/ (sorry I felt it wasn't giving CLM the highlight it deserves) What it is: a new projection head for Qwen3-8B. github: https://github.com/Contrastive-LM/CLM hf:… 28 r/LocalLLaMA community 4d ago My foray into local ai. Two BC-250 ex mining apus running Qwen3.6-35B-A3B Q4_K_M at 60 tok/s with 64k context These boards cost me $115 each and I have them connected using llama.cpp with Vulkan and RPC on Bazzite. The boards have roughly 27GB of combined GPU memory and communicate over 1gb Ethernet. For around $300 including psu I’m loving the performance. I have a few more and want to… 22 The Information — AI news-outlet 4d ago DeepSeek’s Annualized Revenue Hits $1 Billion as Startup Finalizes $7.5 Billion Fundraising DeepSeek’s annualized revenue run rate has hit $1 billion, more than double from less than $500 million a few months ago, buoyed by a recent price hike as well as continued popularity of its models, according to two people with direct knowledge of the matter. The latest revenue… 37 The Information — AI news-outlet 4d ago Open Source, Model Price Cuts Keep AI Costs Under Control Open-source models as well as cheaper model releases from Anthropic and OpenAI are helping business customers rein in costs while getting as much or more out of AI, said speakers at The Information’s AI Agenda Live conference in San Francisco Wednesday. “The existence of open… 34 arXiv — NLP / Computation & Language research 4d ago Cross-Scale Transfer Learning for Depression Severity Prediction: From PHQ-8 to HAMD-17 Across Languages and Clinical Paradigms arXiv:2609.28430v1 Announce Type: new Abstract: This work addresses continuous depression-severity score prediction from clinical interview transcripts under data scarcity. We propose a sequential low-rank adaptation (LoRA) protocol for cross-scale transfer: a Qwen3 backbone… 27 arXiv — NLP / Computation & Language research 4d ago Text Scores Can Miss Waveform Use: A Qwen2-Audio Quantization Case Study arXiv:2609.26823v1 Announce Type: cross Abstract: Post-training quantization of speech language models is often summarized with text-output scores and nominal bit widths. Those numbers alone do not establish behavior that depends on information missing from a transcript, or… 38 The Information — AI news-outlet 4d ago Google Nears Release of Flagship Gemini 4 AI Model Google is nearing the release of its newest flagship model, Gemini 4, the head of its DeepMind division said Wednesday, a long-awaited development after the tech giant fell behind rivals Anthropic and OpenAI in the AI model race. Speaking at The Information’s AI Agenda Live… 13 The Information — AI news-outlet 4d ago Amazon’s New AI Offer Reflects Discounting Surge Have we gone full circle on AI pricing? A bunch of tech firms are throwing around discounts and free offers for AI products, with Amazon on Wednesday announcing that merchants selling on its shopping site will get a year’s free use of its Quick Plus AI assistant product.… 34 r/LocalLLaMA community 4d ago Qwen 3.8 Flash Next q4_k_m, 130k context, q8 cache on 16GB VRAM ann 64GB RAM, 15-20 t/s on 4080 Thought it's about time to share after testing for a week. You need four things most people miss: the right quant, the right model, the right branch, and the right cache flags. https://github.com/dtm-beep/qwen38-flash-next-mtp-16gb TLDR: AtomicChat AD-4.27bpw Q4_K_M target + the… 25 r/LocalLLaMA community 4d ago Using uncensored models makes working less of a headache I have a lot of projects with my friends and team at work that I copy to use for my personal projects, whether it's a plugin I borrow with their consent or a script. I always find that Qwen 3.8 and Muse Spark 1.3 straight up refuse to do anything, as they see it as a steal, so I… 17 NVIDIA Developer Blog official-blog 4d ago Introducing NV-Reason-CT Open 3D CT VLM for Radiologist Chain-of-Thought Reasoning Radiology AI has made remarkable strides in detecting abnormalities across chest X-rays, pathology slides, and 2D scans. Yet one of the most clinically rich and... 17 TechCrunch — AI news-outlet 4d ago Anthropic says its biology lab has already found something big But maybe the biggest reveal is that Anthropic has not let Claude run loose in its biology lab. Humans are still, so far, in the loop. 17 Don't Worry About the Vase community 4d ago Claude Opus 5.5: The System Card Introducing the world’s most powerful model, at least by some measures like Artificial Analysis or any standard benchmark list, which is now Claude Opus 5.5. 7 r/LocalLLaMA community 4d ago Qwen FN vs 27B --- Think I'm saturated. Got QFN up and running on our Strix box this past weekend and have been running on it for a few days now. Big thank you/shout out to the Halogen team, it's running fantastic on the Strix, this is clearly "the setup" right now for this hardware with this model, really impressive… 34 llama.cpp releases dev-tools 4d ago v0.5.0 Overview This release focuses on backend performance and correctness, broader model coverage, and more robust server/router operation. It adds HRM-Text (DFM Mimir 1B) support, MiMo-V2.6 and HunyuanOCR conversion support, ggml 0.25.0 backend improvements, multi-address HTTP… 26 The Information — AI news-outlet 4d ago Anthropic Says Its AI Helped Discover a Possible New Gene-Editing Tool Anthropic said that Claude helped discover a previously unknown molecular system that the company said could represent a new gene editing tool akin to CRISPR, which has transformed the creation of new gene therapies. In a post on X , Anthropic CEO Dario Amodei cautioned that the… 5 r/LocalLLaMA community 4d ago Introducing Support for Local AI Models in the Antigravity SDK   submitted by   /u/dryadofelysium [link]   [comments] 31 r/LocalLLaMA community 4d ago M5U base 96GB inference numbers for Q3.8FN after 112M tokens TLDR; Base M5 Ultra 96 GB ran Q3.8 FN aggregate 3.2k PP and ~170 TG in 4 concurrency Alert: Numbers and custom server details at end are AI assisted So the good news is that I got the base model on launch day with only 64 core GPU. All benchmarks are for current maxed out model,… 34 The Information — AI news-outlet 4d ago Former Lone Pine Capital Director Launches $500 Million-Plus Tech Fund Kevin Salimian, former managing director at investment firm Lone Pine Capital, has launched a new firm with more than $500 million in committed capital, according to a person familiar with the firm’s operations. The firm, Voxel Capital Partners, will invest in 12 to 15 public… 23 Hacker News — AI on Front Page community 4d ago Claude discovers a novel enzyme system with CRISPR-like repeats Article URL: https://www.anthropic.com/news/claude-discovers-novel-enzyme-system Comments URL: https://news.ycombinator.com/item?id=49820134 Points: 253 # Comments: 250 14 r/LocalLLaMA community 4d ago BFL releases FLUX 3 Action: a 7B robot model read more: https://bfl.ai/models/flux-3-action   submitted by   /u/paf1138 [link]   [comments] 36 LangChain releases dev-tools 4d ago langchain-anthropic==1.7.4 Changes since langchain-anthropic==1.7.3 chore(anthropic): fix integration test cassette ( #40790 ) release(anthropic): 1.7.4 ( #40786 ) fix(anthropic): add Opus 5.5 and GPT-6 profile augmentations ( #40785 ) feat(anthropic,openai): mid-conversation tool changes on SystemMessage… 17 Simon Willison community 4d ago Gemini 3.8 TTS Playground Tool: Gemini 3.8 TTS Playground Google released two new Gemini text-to-speech models today - gemini-3.8-flash-tts and gemini-3.8-flash-lite-tts . They come with a library of over 2,000 voices, plus the ability to create a custom voice with "just a 30-second audio sample of your… 16 r/LocalLLaMA community 4d ago Pi agent qwen 3.8 flash next plays Baldur's Gate 2 If like me you enjoyed classics like Baldur's Gate 2, this is a small fun experiment. For anyone interested: https://www.youtube.com/live/8FhPfRKTucw?si=PMIVdHbZ-31PNwZD Qwen 3.8 Flash Next has amazing agentic capabilities but what about playing video games? There is some… 13 Google DeepMind official-blog 4d ago Advancing Private AI Compute with secure, server-side memory Introducing private, server-side memory to Private AI Compute for personal AI. 17 LangChain releases dev-tools 4d ago langchain-openai==1.6.5 Changes since langchain-openai==1.6.4 release(openai): 1.6.5 ( #40787 ) fix(anthropic): add Opus 5.5 and GPT-6 profile augmentations ( #40785 ) feat(anthropic,openai): mid-conversation tool changes on SystemMessage ( #40758 ) 29 Hacker News — AI on Front Page community 4d ago Gemini 3.8 text-to-speech Article URL: https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-8-text-to-speech/ Comments URL: https://news.ycombinator.com/item?id=49817615 Points: 214 # Comments: 111 35 Google DeepMind official-blog 4d ago Gemini 3.8 text-to-speech says hello Gemini 3.8 text-to-speech says hello Sep 23, 2026 | x.com Facebook LinkedIn Mail Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS are our most expressive audio generation models yet. Generate custom character voices and direct scene dialogue across Google AI Studio, Gemini… 26 Hacker News — AI on Front Page community 4d ago GPT-6 Astra has gained the ability to drive a car Article URL: https://drivingbench.com/ Comments URL: https://news.ycombinator.com/item?id=49817404 Points: 213 # Comments: 181 29 r/LocalLLaMA community 4d ago Space Bunny is the new stealth model in OpenCode. Free to try, multimodal I’m really curious to know which company released it.   submitted by   /u/Greney_Yunan [link]   [comments] 31 r/LocalLLaMA community 4d ago Perhaps the highest quality mainline quants of Qwen3.8 27B? I am proud to release these quants of Qwen 3.8 27B. They beat the excellent ISTA and Unsloth quants byte-for-byte on three corpora. Both KLD and top 1% were tested 3x. It took a week of continuous GPU and CPU time to generate these, all done on a single Strix Halo.… 4 TechCrunch — AI news-outlet 4d ago YouTube will let you build your own algorithm with AI ouTube’s new custom feeds let users describe the videos they want to see in their own words, then use Gemini to build a personalized feed around the request. 16 TechCrunch — AI news-outlet 4d ago YouTube releases new AI features for creators within its Studio app YouTube is adding new features to generate ideas and monitor the performance of thumbnails. 23 Hacker News — AI on Front Page community 4d ago Seattle City Council votes to ban surveillance pricing in sale of groceries Article URL: https://advocacy.consumerreports.org/press_release/seattle-city-council-votes-to-ban-surveillance-pricing-in-sale-of-groceries/ Comments URL: https://news.ycombinator.com/item?id=49816374 Points: 219 # Comments: 120 9 TechCrunch — AI news-outlet 4d ago Spotify’s is giving you the keys to its recommendation algorithm with US launch of ‘Taste Profile’ Spotify is rolling out Taste Profile to Premium users in the U.S., letting listeners see how the streamer understands their tastes and use natural language to reshape their recommendations 18 llama.cpp releases dev-tools 4d ago b11130 make-release : update summary prompt Website: https://llama.app Attestations: https://github.com/ggml-org/llama.cpp/attestations/49525233 macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework Linux: Ubuntu… 30 Hacker News — AI on Front Page community 4d ago Claude Code reads AGENTS.md only when telemetry is on Article URL: https://blog.szypowi.cz/p/claude-code-reads-agents.md-only-when-telemetry-is-on/ Comments URL: https://news.ycombinator.com/item?id=49814947 Points: 283 # Comments: 127 20 OpenAI official-blog 4d ago Harvey turns legal context into stronger drafts with GPT-6 Astra GPT-6 Astra produces more structured, context-aware legal documents, freeing lawyers to focus on strategy. 33 OpenAI official-blog 5d ago Introducing MentalHealthBench MentalHealthBench is an expert-informed benchmark for evaluating helpful and safe AI responses across realistic mental health conversations. 7 The Information — AI news-outlet 5d ago OpenAI Partners With Grab in Southeast Asia for AI Skills Program OpenAI and Grab, a Singapore-based transport and food delivery company, on Wednesday launched a program to train Grab’s drivers, merchants and delivery riders to use ChatGPT for work. The program covers Singapore, Thailand, Indonesia and the Philippines and aims to train around… 29 r/LocalLLaMA community 5d ago Most powerful harness for Qwen 3.8? I hear Qwen code unlocks the model better. I also think it has more power user features than open code? It’s nice open code can work with multiple models easier though Thoughts?   submitted by   /u/FactoryReboot [link]   [comments] 5 Page 3 of 10 · 500 articles ← Newer Older →