News / #security Tag Security 500 articles archived under #security · RSS Sign in to follow r/LocalLLaMA community 4d ago MiMo-V2.6 (both Pro and Flash) is a benchmaxxed scam MiMo-V2.6-Pro has an insanely high score of 46 on AA, putting it at the head of the opensource models available. It also costs pennies. Flash is not out on AA yet, but it costs less than half on datacenter and is slightly below on Xiaomi's own benchmarks. It also fits in 192GB,… 12 The Information — AI news-outlet 5d ago Shares of Chinese AI Model Firms Fall After Report of Regulatory Probe Hong Kong-listed shares of Z.ai and MiniMax plunged on Wednesday, after The Information reported that China’s internet regulator is probing potential data leaks to Anthropic. Z.ai, also known as Zhipu, dropped 12.4%, MiniMax fell 4% and Alibaba declined 4.4%. The benchmark Hang… 37 arXiv — Machine Learning research 5d ago MT-ProtBERT: Multi-task Learning ProtBERT for Intrinsically Disordered Proteins Classification with Scarce Data arXiv:2609.25334v1 Announce Type: new Abstract: Intrinsically disordered proteins (IDPs) differ from folded proteins in that they are dynamic, lack a stable three-dimensional conformation, and have low sequence similarity between similar proteins. The conformational… 8 arXiv — Machine Learning research 5d ago Deep Reinforcement Learning on Item-Compatibility Graphs for One-Dimensional Bin Packing arXiv:2609.25397v1 Announce Type: new Abstract: The one-dimensional bin packing problem (1D-BPP) is a classical NP-hard combinatorial optimization problem with applications ranging from logistics and manufacturing to cloud resource management. Although deep reinforcement… 13 arXiv — Machine Learning research 5d ago EMGBlend: Heterogeneity-Aware Self-Supervised Pretraining for Gesture and Force Decoding arXiv:2609.25582v1 Announce Type: new Abstract: Public surface electromyography (EMG) datasets vary widely in electrode layout, channel count, frequency support, and size. Simply mixing them for pretraining can misalign channel semantics, introduce spectral targets that some… 21 arXiv — Machine Learning research 5d ago Graph Domain Adaptation Does Not End with Representation Learning arXiv:2609.25692v1 Announce Type: new Abstract: Graph domain adaptation (GDA) transfers knowledge from a labeled source graph to an unlabeled target graph under shifts in both node attributes and graph structure. Existing methods primarily adapt graph representations through… 7 arXiv — Machine Learning research 5d ago Multi-View Fair Clustering Guided by Cross-View Sensitive Information Discrepancy arXiv:2609.25811v1 Announce Type: new Abstract: Multi-view clustering (MVC) aims to uncover latent cluster structures by exploiting complementary information from multiple views. Despite substantial progress in clustering performance, fairness remains an important concern when… 38 arXiv — Machine Learning research 5d ago Protocol before progress: leakage-aware evaluation of AIS trajectory prediction arXiv:2609.25827v1 Announce Type: new Abstract: Reported gains in vessel-trajectory prediction from Automatic Identification System (AIS) data are credited to new architectures, but the evaluation protocol is rarely measured as a source of error reduction. We build a… 30 arXiv — Machine Learning research 5d ago Quantifying Protocol-Induced Uncertainty in Comparative Predictive-Model Evaluation: Evidence from Large-Scale Daily PM10 Forecasting arXiv:2609.26288v1 Announce Type: new Abstract: Comparative studies of predictive models often end by ranking candidate models, yet these rankings depend on evaluation protocols whose influence is rarely treated as a source of uncertainty. We formalize this problem as… 29 arXiv — NLP / Computation & Language research 5d ago LLM-Driven Training-free Location-Attribute Synergic Fusion: A Closed-Loop Paradigm for Dual-source Encrypted POIs and LULC Mapping arXiv:2609.25051v1 Announce Type: new Abstract: Dual-source encrypted points of interest (DSEP), POIs from two encrypted coordinate systems, suffer from intertwined location and attribute uncertainties, including nonlinear systematic misalignment and naming inconsistency,… 20 arXiv — NLP / Computation & Language research 5d ago FinFIRST: Benchmarking Search Agents for Financial Information Retrieval, Sourcing and Traceability arXiv:2609.25192v1 Announce Type: new Abstract: Financial search is a highly demanding task for LLM agents, requiring not only a correct final answer but also temporally valid information retrieval, authoritative source selection, entity and period alignment, unit and definition… 35 arXiv — NLP / Computation & Language research 5d ago TelecomGPT-R1: Unified Post-Training for Reasoning Across Heterogeneous Telecom Tasks arXiv:2609.25356v1 Announce Type: new Abstract: Large language models (LLMs) offer great potential to automate a broad range of telecom engineering tasks by reasoning over standards, network configurations, mathematical models, source code, and operational logs. However,… 8 arXiv — NLP / Computation & Language research 5d ago Isolated Sign Language Recognition for Icelandic Sign Language: Experiments in a Low-resource Setting arXiv:2609.25862v1 Announce Type: new Abstract: We present the first experiments on isolated sign language recognition (ISLR) for Icelandic Sign Language (\'ITM). We use \'ITM SignWiki, a dataset derived from a bilingual Icelandic--\'ITM online dictionary. It is genuinely… 22 arXiv — NLP / Computation & Language research 5d ago Capable yet Parsimonious: Extracting and Characterizing Hidden Chain-of-Thought in Frontier Models arXiv:2609.26637v1 Announce Type: new Abstract: The rapid capability gains of frontier language models are widely attributed to improved reasoning abilities, yet this cannot be verified as raw CoT traces in closed-source systems are hidden. By registering a simple custom tool… 35 r/LocalLLaMA community 5d ago DeepSeek and Moonshot AI face Beijing's probe over potential data leaks to Anthropic The rumor about Kimi execs getting arrested finally has some legs. I believe the reality is more like under investigation for potential arrests or fine.   submitted by   /u/Ok_Warning2146 [link]   [comments] 20 r/LocalLLaMA community 5d ago New 6B image model coming, AntLing just open sourced the Ming-Image-0.1-Design family • Ming-Image-0.1-Design, 6B • Ming-Image-0.1-Design-Layer, 6B • Two open-source Agent Skills: the Ling UI Design Skill and the Image-to-Editable-PPT Skill Ming-Image-0.1-Design ranks #1 among open-weight models on Artificial Analysis’s UI/UX Design leaderboard.… 12 r/LocalLLaMA community 5d ago AntLing open sourced the Ming-Image-0.1-Design family AntLing open sourced the Ming-Image-0.1-Design family: • Ming-Image-0.1-Design, 6B • Ming-Image-0.1-Design-Layer, 6B • Two open-source Agent Skills: the Ling UI Design Skill and the Image-to-Editable-PPT Skill Ming-Image-0.1-Design ranks #1 among open-weight models on Artificial… 9 Hacker News — AI on Front Page community 5d ago WordPress: Unauthenticated path traversal leading to conditional RCE Article URL: https://github.com/WordPress/wordpress-develop/security/advisories/GHSA-7hp8-65ch-5whp Comments URL: https://news.ycombinator.com/item?id=49803959 Points: 212 # Comments: 111 4 The Information — AI news-outlet 6d ago China Probes DeepSeek, Moonshot Over Potential Data Leaks to Anthropic China’s internet regulator is investigating DeepSeek and Moonshot AI after Anthropic alleged both companies had been routing sensitive user data to Claude models, according to people with knowledge of the matter. In a 154-page report published on Sept. 10, Anthropic detailed how… 16 arXiv — Machine Learning research 6d ago Gaussian Process Decorrelation for Spatiotemporal Deep Learning-Based Snow Water Equivalent Prediction arXiv:2609.22182v1 Announce Type: new Abstract: In the Western United States, snowmelt is essential to the agricultural industry in addition to being a key source of municipal drinking water. Consequently, accurate snowpack forecasting is critical for water policy and… 31 arXiv — Machine Learning research 6d ago EvoRank: LLM-Guided Evolution of Multi-Objective Learning-to-Rank Pipelines arXiv:2609.22196v1 Announce Type: new Abstract: We present EvoRank, an open autonomous ranking engineer: an LLM-guided evolutionary loop that discovers complete Learning-to-Rank pipelines (features, models, losses, ensembles) for multi-objective e-commerce search. On the Expedia… 25 arXiv — Machine Learning research 6d ago UniGIO: Unified Generative Global In-situ Weather Modeling from Spatiotemporal Incomplete Observations arXiv:2609.22217v1 Announce Type: new Abstract: Global In-situ Observation (GIO) provides fine-scale, direct records of the global weather system from sparse point stations, making it an indispensable source for capturing localized and transient dynamics beyond the reach of… 36 arXiv — Machine Learning research 6d ago Resist, Update, Reject: Preference Optimization Installs a Prior-Dependent Reliability Switch arXiv:2609.22359v1 Announce Type: new Abstract: An aligned model asked to hold its answer against a manipulative source must still update on a reliable one and reject an unreliable one: resistance, reliable-update, and unreliable-source rejection are one three-way contract, not… 25 arXiv — Machine Learning research 6d ago COREM: Cosine-Relation Momentum Reshaping with Stateful Writeback arXiv:2609.22487v1 Announce Type: new Abstract: Matrix-valued optimizer states may contain relational structure that is not captured by treating their entries independently. We study whether relations within matrix-valued optimizer states can be exploited to improve… 25 arXiv — NLP / Computation & Language research 6d ago Evaluating Fine-Tuned and Base Language Models in Maternal and Vaccination Healthcare for African Settings arXiv:2609.22110v1 Announce Type: new Abstract: Background: Large language models (LLMs) can improve healthcare information delivery in low-resource settings but may produce inaccurate or culturally inappropriate advice. This study evaluated domain-specific fine-tuning for… 20 arXiv — NLP / Computation & Language research 6d ago Attributable Post-Rationalization in RAG Citations: A Controlled Reproduction and an RLVR Comparison arXiv:2609.23053v1 Announce Type: new Abstract: A RAG system can hand you the right answer and cite a source it did not actually use. Models output these unfaithful citations via post-rationalization: they write the answer first and then attach a citation to whatever passage… 24 r/LocalLLaMA community 6d ago 50+ Hours and 100M+ Tokens Later, Open Source Autonomous Agent is GETTING CLOSER at Solving an Open Math problem This experiment is live, you can inspect all the internal reasoning, memories, attempts here: https://artificium-covering-experiment.gr.bio/ The problem that the agent is trying to solve is a covering design problem: https://en.wikipedia.org/wiki/Covering_design Known as… 19 The Information — AI news-outlet 6d ago OpenAI Releases Proposal for International AI Safety Coordination OpenAI released a proposal on Monday that would create international coordination around AI safety. In a blog post, OpenAI called for national AI safety institutes, such as the U.S. Commerce Department’s Center for AI Standards and Innovation, to set standards around areas… 30 Ars Technica — AI news-outlet 6d ago Trump rejects AI slowdown calls, launches "AI Force" instead The president offered few details on what his proposed new AI Force would do. 8 arXiv — Machine Learning research 7d ago Bio-MF: Low-Latency and High-Fidelity EEG-to-fNIRS Cross-Modal Generation for Hybrid Motor-Imagery Brain--Computer Interfaces arXiv:2609.20904v1 Announce Type: new Abstract: Hybrid motor-imagery brain-computer interfaces (MI-BCIs) combining EEG and fNIRS can outperform EEG-only systems by exploiting complementary electrophysiological and hemodynamic information. To obtain such hybrid information when… 7 arXiv — Machine Learning research 7d ago Talk to Me, Jarvis: An Open-Source Edge-Deployable Voice Assistant Framework for Autonomous Racecars arXiv:2609.21109v1 Announce Type: new Abstract: Recent advances in large language models have improved their effectiveness as back-end components for voice assistants, particularly in intent understanding and context-aware input classification. However, online-hosted models… 13 arXiv — Machine Learning research 7d ago GEM-MPC: Balancing Exploration and Exploitation through Expert-Guided Planning arXiv:2609.21735v1 Announce Type: new Abstract: Effective exploration in high-dimensional continuous control remains a central challenge in reinforcement learning. Planning-based methods address this by combining online planning with learned policies and value functions, but… 25 arXiv — Machine Learning research 7d ago Matrix AdaGrad: Row-wise and Column-wise Adaptive Subgradient Methods arXiv:2609.21815v1 Announce Type: new Abstract: Adaptive optimization methods such as AdaGrad and Adam are widely used in modern neural-network training, but their adaptive scaling is primarily designed for vector-valued parameters and does not explicitly exploit matrix… 31 arXiv — Machine Learning research 7d ago $\lambda$-Controlled GRPO: Turning Flow-Matching Ratio Instability into a Budgeted Resource arXiv:2609.22041v1 Announce Type: new Abstract: Reinforcement learning is increasingly used to align image generators with reward signals, and Flow-GRPO recently extended this paradigm to flow-matching models by treating the denoising sampler as a stochastic policy that can be… 38 arXiv — Machine Learning research 7d ago BrainWideBench: Benchmarking large-scale pretraining and across-animal transfer in multi-region neural recordings arXiv:2609.22064v1 Announce Type: new Abstract: Advances in large-scale neural recording have made it possible to collect data across many animals and distributed brain regions, raising the question of whether this scale can be exploited to learn general-purpose neural… 14 arXiv — Machine Learning research 7d ago Fragment-Aware Vision Transformers for Fresco-Fragment Style Classification arXiv:2609.21012v1 Announce Type: cross Abstract: Artistic style classification is usually studied on complete artworks, where models can exploit global composition, spatial organisation, and iconographic structure. In archaeological settings, however, artworks often survive… 15 arXiv — NLP / Computation & Language research 7d ago COAL-SQL: Coverage-Guided Augmentation and Failure-Driven Learning for Text-to-SQL Post-Training arXiv:2609.20842v1 Announce Type: new Abstract: Text-to-SQL translates natural-language questions into executable SQL queries, but open-source large language models still require task-specific post-training for complex, real-world SQL generation. Effective post-training requires… 28 arXiv — NLP / Computation & Language research 7d ago Benchmarking Gender Bias in Machine Translation Evaluation Metrics across Occupations arXiv:2609.21490v1 Announce Type: new Abstract: Gender bias remains a persistent concern in machine translation (MT), affecting both generated translations and their automatic evaluation. When a source text leaves a person's gender unspecified, translations may realize that… 19 arXiv — NLP / Computation & Language research 7d ago Trustworthy FinAInce: Unpacking How AI-Mediated Financial Advice is Judged arXiv:2609.20989v1 Announce Type: cross Abstract: As generative AI is increasingly used as a source of personal financial guidance, understanding how people appraise such advice is important for supporting appropriate reliance. We conducted a randomized vignette experiment with… 20 r/LocalLLaMA community 7d ago ZCode is now open source ZCode is now open source , and the reported security issues have been addressed. Source code: https://github.com/zai-org/ZCode The repo includes its desktop app, web workspace, backend, Agent CLI, and runtime. Official announcement: In response to the ZCode product security… 24 r/LocalLLaMA community 7d ago laya.cpp: Optimized laya near-instant decision making After seeing u/Nandakishor_ml’s post introducing Laya , I wanted to see how fast it could run in a standalone C++ implementation. Credit to u/Nandakishor_ml for the architecture, training and open-source release. My contribution is the inference implementation: laya.cpp , built… 25 llama.cpp releases dev-tools 8d ago b11059 metal: add F16 input to the FWHT ( #29094 ) metal: add F16 input to the FWHT The Metal FWHT kernel accepts F32 input only. This change makes the source type a template parameter, so the kernel reads an F16 source directly instead of requiring a converted copy. The F32… 11 r/MachineLearning community 8d ago ProgramAsWeights: compile English function descriptions into neural programs that run locally [R] Given the recent interest in tools like Jev, I wanted to share ProgramAsWeights (PAW), an open-source research project I'm working on at the University of Waterloo. You describe a text function in English, compile it into a reusable neural program, and run it locally, including… 13 TechCrunch — AI news-outlet 8d ago Flock reportedly tries to shrink workforce with employee buyouts Without buyouts, Flock would "almost certainly" need to lay off staff. 14 r/LocalLLaMA community 8d ago “DeadGrid” now open source exclusively made with qwen 3.8 27b Q4KM Play it right in your browser, no download. Some of the GLB’s are messy, but overall I’ve enjoyed playing with it, and the last iteration made was the gun/weapon placement. Contributors are welcome, would like to see what can be made of this from local inference only.… 4 TechCrunch — AI news-outlet 8d ago Trump suggests rebranding AI with a new name, says he’s also creating an AI Force Trump claimed, without evidence, that the AI backlash is a Democratic hoax. 20 r/LocalLLaMA community 8d ago Von: Open-source 395M "System One" model Took me a while since I'm on a family trip and have limited hardware, but here it is! Von: Open-source "System One" drop-in replacement for TypeSafe's JEV. https://github.com/wfzyx/von https://huggingface.co/wfzyx/von-1.0 It runs entirely on a CPU with 1–2 GB of memory (I… 21 llama.cpp releases dev-tools 8d ago b11053 server : improve startup log messages ( #29125 ) server-models : show source per model in log Show [source] tag (preset/models_dir/cache) per model instead of cryptic * marker Show HF hub cache path in the 'Loaded cached model presets' log Add hf_cache::get_cache_dir() public… 14 TechCrunch — AI news-outlet 8d ago Vals, backed by Andreessen Horowitz, is looking to become the gold standard for AI benchmarking Vals AI is hoping to make AI benchmarking a more neutral and trustworthy resource in a world increasingly inundated by AI models. 32 Hacker News — AI on Front Page community 9d ago Laya the open source version of Jev Article URL: https://laya.convaiinnovations.com/ Comments URL: https://news.ycombinator.com/item?id=49765348 Points: 330 # Comments: 64 15 Page 2 of 10 · 500 articles ← Newer Older →