News / #security Tag Security 500 articles archived under #security · RSS Sign in to follow r/LocalLLaMA community 2h ago NVIDIA shipped OpenShell, an open source sandbox that gives local and open agents real runtime limits instead of prompt rules. Over 100 firms joined the safety stack. OpenAI did not. https://x.com/JensenHuang/status/2104499465055023424   submitted by   /u/InternationalGap3698 [link]   [comments] 14 r/MachineLearning community 5h ago Free, open-source AI engineering course where you build each algorithm by hand: 523 lessons, now as EPUB/PDF books [P] AI Engineering from Scratch is an MIT-licensed curriculum: 523 lessons across 20 phases, from linear algebra and backprop to transformers, LLMs, agents, and production serving. The code is stdlib-first, so you see every step instead of calling a library. This month's edition: -… 37 arXiv — NLP / Computation & Language research 7h ago All In Good Time: Causality-Aware Framework for LLM-Based Simultaneous Speech-to-Speech Translation arXiv:2609.30416v1 Announce Type: new Abstract: Large Language Models (LLMs) have shown strong performance in low-resource offline translation; however, extending them to simultaneous speech-to-speech translation (Simul-S2ST) remains challenging due to the scarcity of causally… 30 arXiv — NLP / Computation & Language research 7h ago REALMS: An AI-Assistant Conversational System for Real-Time Exact Audience Sizing over High-Dimensional Nested Profiles arXiv:2609.30547v1 Announce Type: new Abstract: Audience sizing is a critical component of digital marketing. It enables precise resource allocation, campaign planning, and performance optimization. Traditional approaches using skeleton audiences, sampling, or predictive… 9 arXiv — NLP / Computation & Language research 7h ago THA: Weighted Finite-State Text Normalization and Inverse Text Normalization for Khmer arXiv:2609.30984v1 Announce Type: new Abstract: Text-to-speech needs written text in spoken form, and speech recognition output needs the reverse. For Khmer, neither direction has a maintained open-source tool, and the script makes both harder: words are not separated by spaces,… 8 arXiv — NLP / Computation & Language research 7h ago ZooWork-ShopRanker: An Open, Preference-Aligned E-Commerce Reranker arXiv:2609.31002v1 Announce Type: new Abstract: Open rerankers trained for general web retrieval transfer imperfectly to e-commerce, where ranking decisions depend not only on topical relevance but also on user preferences, product constraints, and comparative product fit. These… 18 arXiv — NLP / Computation & Language research 7h ago Evaluating Cultural Awareness of LLMs for Haitian Creole arXiv:2609.31506v1 Announce Type: new Abstract: Large language models (LLMs) exhibit substantial performance disparities between high- and low-resource languages. Beyond lower task performance, they often fail to capture the cultural norms and values of underrepresented… 22 arXiv — NLP / Computation & Language research 7h ago Prompt Injection Detection for Email Agents Through Attack Chain Modeling arXiv:2609.30657v1 Announce Type: cross Abstract: Large language model email assistants are particularly vulnerable to indirect prompt injection because untrusted email content can be retrieved into the model context and influence subsequent tool use. Existing prompt injection… 6 arXiv — NLP / Computation & Language research 7h ago NaijaNLP: A Survey of Nigerian Low-Resource Languages arXiv:2502.19784v3 Announce Type: replace Abstract: With over 500 languages in Nigeria, three languages - Hausa, Yor\`ub\'a and Igbo spoken by more than 175 million people, account for about 65% of the languages. However, these languages are classed as low-resource due to… 10 arXiv — NLP / Computation & Language research 7h ago Towards Automated Lexicography: Generating and Evaluating Definitions for Learner's Dictionaries arXiv:2601.01842v2 Announce Type: replace Abstract: Dictionary definitions are an essential resource for learning word senses, but manually creating them is costly. We thus study dictionary definition generation (DDG), i.e., the generation of non-contextualized definitions for… 12 r/LocalLLaMA community 20h ago We released VeriLoop E2 (27B, Apache-2.0). The design question behind it: should an LLM be allowed to commit its own state? Disclosure: I’m one of the authors of VeriLoop E2, a 27B model post-trained from Qwen3.8-27B. The weights are Apache-2.0; the harness we evaluate it in is not open source (details at the bottom). This is a release post, but rather than a benchmark dump I want to talk about the… 28 r/LocalLLaMA community 21h ago What if open-source AI focused less on giant models and more on reusable capabilities? Instead of everyone building another general-purpose model, the community could distill open models into domain specialists—biology, Python, accounting, OCR, and more. Developers could combine these capabilities into local tools: small model + OCR + accounting → local accounting… 13 r/LocalLLaMA community 1d ago SupersonicLabs/Julia-1 · Hugging Face New open source Jev like model for running on local devices from a group called Supersonic Labs. It's a 144M param local non-generative local classifier. From their site : Julia 1 opens our research into compact decision models. It builds on mmBERT-small , a multilingual… 15 llama.cpp releases dev-tools 1d ago b11203 cuda: add F16 input to the FWHT ( #29096 ) cuda: add F16 input to the FWHT The CUDA FWHT accepts F32 input only. This makes the source type a template parameter, so the kernel reads an F16 source directly instead of requiring a converted copy. The F32 path is unchanged.… 33 r/LocalLLaMA community 1d ago I added Qwen-Image 2.1 + LoRA support to TensorSharp (GGUF, local inference) I maintain TensorSharp , an open-source inference engine. It can now run Qwen-Image 2.1 locally for text-to-image generation and image editing, with support for its LoRA adapters. I’ve added configs for regular style and editing LoRAs, plus accelerated adapters with their own… 14 r/LocalLLaMA community 1d ago 2400cc Inference Racer: Dual RTX 3090 motors, NVLink turbo, naked 7840U ThinkPad ECU, VW Golf radiator Today I present a fine piece of engineering, carefully assembled inside a custom chipboard chassis: the 2400cc Inference Racer , a.k.a. my winter heater. Power comes from two second-hand AORUS RTX 3090 XTREME WATERFORCE cards. One glows a beautiful teal, the other red. I have no… 37 The Information — AI news-outlet 2d ago OpenAI Found ‘Dozens’ of New Instances of AI Misbehavior OpenAI said Friday it has notified dozens of entities, including governments and universities, that its agents may have breached or spammed their websites or services. The agents took actions such as obtaining unauthorized access to websites or using the organizations’ public… 11 The Information — AI news-outlet 2d ago Exclusive: Meta Bolsters Muse Safety Warning After Security Vulnerability Found Meta Platforms is adding a clearer safety warning within Muse after a security researcher discovered a vulnerability in the AI agent that could let an attacker access a user’s sensitive personal information. The security flaw, which was flagged by an outside researcher through… 18 r/LocalLLaMA community 2d ago Jev vs. Kev: open-source Jev alternative tested side by side We hosted Kev 4B (Jared Palmer's Apache-2.0 fine-tune of Qwen3.5-4B) and ran it side by side with Jev on the same endpoint to see how it compares. We built a fresh set of 362 items published after both models shipped (new arXiv papers, Stack Exchange questions, GitHub issues),… 9 arXiv — Machine Learning research 3d ago Language Specificity vs. Domain Diversity: Benchmarking Transformers for Bangla Medical NER arXiv:2609.29101v1 Announce Type: new Abstract: Medical Named Entity Recognition (NER) for low-resource languages remains a challenging task due to high linguistic variability and a scarcity of domain-specific annotated corpora. This work presents a comprehensive empirical… 31 arXiv — Machine Learning research 3d ago On the second-order optimization for spiking neural networks arXiv:2609.29379v1 Announce Type: new Abstract: Spiking Neural Networks (SNNs) offer an energy-efficient alternative to conventional neural networks by exploiting sparse, binary spikes, and event-driven computation. However, the training of SNNs remains challenging, as spiking… 35 arXiv — Machine Learning research 3d ago When Temporal Perturbations Act Like Sensor Biases: Label-Free Auditing of Wearable Activity Recognizers arXiv:2609.29937v1 Announce Type: new Abstract: Wearable human-activity recognition (HAR) models operate across sensors, subjects, and backbones, yet a smooth waveform may appear temporal while exploiting a persistent sensor offset primarily. We introduce SpectrumAudit, a… 34 arXiv — NLP / Computation & Language research 3d ago PTC-Bias: Phoneme-Level Temporal Competition for Bias Retrieval and Post-Decoding Correction in Speech LLMs arXiv:2609.28727v1 Announce Type: new Abstract: Contextual biasing improves rare-word recognition in speech large language models (SpeechLLMs), but efficiently exploiting large bias lists remains challenging. We propose PTC-Bias, a two-stage framework based on phoneme-level… 35 arXiv — NLP / Computation & Language research 3d ago EnSiTa - A Trilingual Multi-Domain Parallel Dataset and Benchmark for Domain-Specific Machine Translation arXiv:2609.29511v1 Announce Type: new Abstract: Machine Translation (MT) for low-resource languages remains far behind that of high-resource languages, and the gap is widest in specialised domains, where parallel data is scarce or entirely absent. We present EnSiTa, a trilingual… 27 arXiv — NLP / Computation & Language research 3d ago Confident but Wrong: A Constrained Decoding Diagnostic for Low-Resource Automatic Post-Editing arXiv:2609.29680v1 Announce Type: new Abstract: Automatic Post-Editing (APE) for low-resource languages (LRLs) often fails to improve Machine Translation (MT), and the score alone cannot say why: whether more training would help, or whether the training data is too inconsistent… 6 arXiv — NLP / Computation & Language research 3d ago Adaptive Fisher-Whitened Cross-Covariance for Low-Resource Speech Recognition arXiv:2609.29800v1 Announce Type: new Abstract: Adapting multilingual speech foundation models to low-resource languages remains difficult, especially for languages that are poorly represented during pre-training. While parameter-efficient fine-tuning (PEFT) reduces the cost of… 29 arXiv — NLP / Computation & Language research 3d ago ChunkRank: Model-Aware Text Chunking and Abstention-Aware Answer Selection for LLM Pipelines arXiv:2609.29828v1 Announce Type: new Abstract: We present ChunkRank, an open-source Python library that derives chunk boundaries from a target model's tokenizer and context window, and selects an answer among candidates produced independently per chunk. It ships a validated… 38 arXiv — NLP / Computation & Language research 3d ago Scoring Both Directions: LLMs realize the MRS they cannot reliably parse arXiv:2609.30071v1 Announce Type: new Abstract: The English Resource Grammar (ERG) is a hand-written computational grammar of English. Given a sentence, its processor, ACE, produces a formal meaning representation called Minimal Recursion Semantics (MRS): a graph of the… 22 arXiv — NLP / Computation & Language research 3d ago Reward-Tilted On-Policy Distillation for Acoustic Grounding in Audio-Language Models arXiv:2609.28778v1 Announce Type: cross Abstract: Audio-language models (ALMs) can exploit textual shortcuts to answer questions while overlooking acoustic evidence, weakening audio understanding. On-policy distillation (OPD) trains compact ALMs by supervising student-generated… 28 TechCrunch — AI news-outlet 3d ago Oracle sends force majeure notice on its New Mexico Stargate data center The notice would allow Oracle to delay payments should the facility miss its 2028 target to come online. 13 The Information — AI news-outlet 3d ago One of Tether’s Bank Partners Caught Up in Assets Seizure On the surface, Tether, the largest stablecoin issuer in the world, might appear to be a formidable force in global finance. It owns more than $100 billion in U.S. Treasury bills, mostly held at Cantor Fitzgerald, the investment bank formerly led by U.S. Secretary of Commerce… 32 Hacker News — AI on Front Page community 3d ago Show HN: Whiteboard (YC W26) – An open-source IDE for thoughtful software design Hello! We’re Sid, Alex, Ketan, and Milan. We’re building Whiteboard ( https://whiteboard.dev.fast/ ), an open-source desktop app where humans and agents can architect software together in a common workspace. Here’s our repo: https://github.com/devdotfast/whiteboard . We were… 4 Ars Technica — AI news-outlet 3d ago OpenAI agent “didn’t accept no for an answer” in Australian government breach "There will obviously be legal consequences," prime minister promises. 18 The Information — AI news-outlet 3d ago Oracle Invokes Force Majeure in Aim to Protect Itself From Data Center Cost Overruns Oracle, which is set to lease a New Mexico data center on behalf of OpenAI, sent a notice to the project’s developer asserting its contractual rights to withhold payments if project delays persist, according to Bloomberg . Oracle sent a notice that cited “force majeure” to a… 11 TechCrunch — AI news-outlet 3d ago Australia to investigate if OpenAI hack of government health website broke the law The incident is the first known breach to affect a government agency, and Australia's prime minister has vowed to hold OpenAI accountable. 9 The Information — AI news-outlet 4d ago Open Source, Model Price Cuts Keep AI Costs Under Control Open-source models as well as cheaper model releases from Anthropic and OpenAI are helping business customers rein in costs while getting as much or more out of AI, said speakers at The Information’s AI Agenda Live conference in San Francisco Wednesday. “The existence of open… 34 arXiv — Machine Learning research 4d ago Learning Risk Scores Robust to Unobserved Confounders arXiv:2609.27144v1 Announce Type: new Abstract: We consider the problem of learning risk scores to prioritize individuals for scarce resources or interventions, from historical observational data affected by unobserved confounding. Decisions about who receives scarce resources… 16 arXiv — Machine Learning research 4d ago A Systematic Benchmark of Explainable Methods for Temporal Attribution in Sequential Recommendation Systems arXiv:2609.27201v1 Announce Type: new Abstract: Sequential RecSys are central to modern personalization, exploiting user's historical interaction sequences to drive next-step decisions. Deep learning models, particularly CNN and Transformer-based architectures, have proven… 32 arXiv — Machine Learning research 4d ago Discover, Falsify, Revise: Auditing Input-Use Claims from Source Code to Predictive Contribution in Agent-Discovered Cell Models arXiv:2609.27234v1 Announce Type: new Abstract: AI virtual cells aim to predict cellular responses to specified interventions, yet held-out predictive performance alone does not establish use of the supplied perturbation information. This prediction-claim gap matters in agentic… 6 arXiv — Machine Learning research 4d ago Graph Learning with Spectral Connectivity Priors for Scarce Data arXiv:2609.27278v1 Announce Type: new Abstract: Learning a sparse graph from scarce data is practically important but challenging. Motivated by the desirable combination of local sparsity and strong global connectivity exhibited by expander-like graphs, we propose spectral… 5 arXiv — Machine Learning research 4d ago When Labels Are Scarce: An Oscillatory State Space Model for Vibration Diagnosis arXiv:2609.27411v1 Announce Type: new Abstract: Machine fault diagnosis from vibration requires learning from scarce labelled fault recordings while meeting the computational constraints of edge devices for local inference. We introduce DualRes, a compact oscillatory state-space… 30 arXiv — Machine Learning research 4d ago EBRL: Asynchronous Embodied RL by Multi-Grained Resource Management arXiv:2609.27547v1 Announce Type: new Abstract: Embodied reinforcement learning (RL) improves model capabilities with a pipeline of environment simulation, action generation, and model updates. These stages show heterogeneous CPU and GPU demands, making efficient resource… 7 arXiv — Machine Learning research 4d ago TNLearn: An Open Source Python Package for Task-based Neurons arXiv:2609.27564v1 Announce Type: new Abstract: The brain does not rely on a single type of neuron to perform all kinds of tasks; instead, it designs different neurons for different tasks. The concept of task-based neurons represents a paradigm shift compared to task-based… 36 arXiv — Machine Learning research 4d ago When Adaptation Hurts: Split Sensitivity and Person-Level Negative Transfer in Federated Wearable Onboarding arXiv:2609.27819v1 Announce Type: new Abstract: Federated wearable models eventually serve people absent from source training, but favorable average accuracy does not establish that unlabeled onboarding helps each person. We evaluate six core onboarding strategies on five… 20 arXiv — Machine Learning research 4d ago Exact Minimax One-Bit Unbiased Compression: Heavy-Tail Necessity and Finite-Randomness Approximation arXiv:2609.27860v1 Announce Type: new Abstract: A pointwise-unbiased one-bit compressor reconstructs every real input in expectation while transmitting one bit. For a scalar source $P$ with CDF $F$, mean $m$, and $\mathcal J(P)=\int_{\mathbb R}\sqrt{F(r)(1-F(r))}\,dr$, we prove… 19 arXiv — NLP / Computation & Language research 4d ago AraGenre 2026: A Hierarchical Definition-Guided Arabic Genre Classification Shared Task arXiv:2609.27387v1 Announce Type: new Abstract: AraGenre is a shared task on hierarchical, definition-guided Arabic genre classification, motivated by the limited availability of annotated data in Arabic and other low-resource languages. Systems assign each Arabic text segment… 38 arXiv — NLP / Computation & Language research 4d ago Hard Negatives Reveal What Easy Negatives Hide: Cross-Lingual Harmfulness Representations Degrade with Resource Tier Under Hard Negatives arXiv:2609.27758v1 Announce Type: new Abstract: Safety alignment in large language models is trained primarily in English, and recent work reports that the underlying harmfulness representation survives translation: English-trained probes separate harmful from harmless prompts… 38 Hacker News — AI on Front Page community 4d ago Ideas on modernizing the open-source desktop Article URL: https://lwn.net/SubscriberLink/1095425/2d9f411252325784/ Comments URL: https://news.ycombinator.com/item?id=49825642 Points: 301 # Comments: 349 30 The Information — AI news-outlet 4d ago Meta Announces New Retail Partnerships for Agent Muse Meta Platforms announced partnerships with Walmart and other retailers for its AI agent Muse days after e-commerce giant Amazon blocked the personalized agent from accessing its shopping site. At Meta’s annual Connect conference on Wednesday, Chief AI officer Alexandr Wang said… 10 r/LocalLLaMA community 4d ago RSI ACHIEVED at Reef(Rumours again lol) and it's open source   submitted by   /u/Muhlwa_Sholanke [link]   [comments] 29 Page 1 of 10 · 500 articles Older →