News / #security Tag Security 500 articles archived under #security · RSS Sign in to follow r/MachineLearning community 9d ago ACM TAPS moved my camera-ready to support, deadline is in 2 days. Anyone been through this? [D] My paper got accepted at an ACM conference and I'm doing the camera-ready through TAPS. The PDF compiled fine, but TAPS flagged a "source to HTML conversion issue" and moved the file to TAPS Support. Now the upload option is gone from my dashboard, so I can't resubmit myself.… 35 r/LocalLLaMA community 9d ago Qwen3.8-27B at 144 tok/s on an M5 Max MacBook Pro Meet Inco Splash, open-source inference engine, built around the model and around Apple silicon. Up to 3× the decode speed of Ollama, 2× oMLX, and almost 4× when an agent fans out into sub-agents. Requirements: M3 or newer, macOS 26.4+, 36 GB Get started with a single command:… 32 r/LocalLLaMA community 9d ago I truly think every major AI lab is purposefully making fear-mongering headlines to get regulations that hurt open-source models   submitted by   /u/Fusseldieb [link]   [comments] 35 The Information — AI news-outlet 9d ago Google’s Gemini Model Hacks Companies During Test Google acknowledged that its Gemini AI model unexpectedly breached the networks of three outside companies during safety evaluations conducted by third-party testing firm Irregular last May, The Wall Street Journal reported. During the exercise, Gemini gained entry to the… 33 r/LocalLLaMA community 9d ago ProgramAsWeights: describe an AI function in English, compile it once, and run it locally on CPU With the recent interest in Jev, especially in open-source and locally executable alternatives, I wanted to share a related project we've been building at the University of Waterloo: ProgramAsWeights (PAW). The idea is simple: describe a function in English, compile it once,… 9 Hacker News — AI on Front Page community 9d ago Korea raises data breach fines to 10% of revenue Article URL: https://www.koreajoongangdaily.com/business/korea-raises-data-breach-fines-to-10-of-revenue/12869899 Comments URL: https://news.ycombinator.com/item?id=49759466 Points: 219 # Comments: 68 22 Ars Technica — AI news-outlet 9d ago US government website used Chinese model the FBI called "malicious" The Federal Register website briefly used an open source Chinese AI search tool. 12 Stratechery (Ben Thompson) community 9d ago 2026.38: Doomforce The best Stratechery content from the week of September 14, 2026, including the view from anywhere but San Francisco, the limited potential for a pacing deal, and the Salesforce zag. 33 r/LocalLLaMA community 9d ago MiniMax Code is now open source. Maybe M3.1 is the next thing to watch. MiniMax has released the source code for MiniMax Code’s terminal agent: https://github.com/MiniMax-AI/minimax-code https://preview.redd.it/st9izmm8zaqh1.png?width=3016&format=png&auto=webp&s=5853bc316ea0b1823e42fd5574a4e713b47e3441 For those unfamiliar with it, MiniMax Code is… 35 r/LocalLLaMA community 9d ago MiniMax Code goes open source MiniMax has open-sourced the terminal version of MiniMax Code: https://github.com/MiniMax-AI/minimax-code How can developers verify the content that encoding proxies read, send, and store? This is a topic that has been widely discussed recently. Open sourcing the agent doesn’t… 4 r/LocalLLaMA community 9d ago China’s mysterious AI company Naive AI is valued at over $1.4 billion and could release its first open-source llm model as early as this month. According to people familiar with the matter, Naive AI, an AI startup founded in February this year by Tsinghua University professor Dai Jifeng, has completed three funding rounds totaling $400 million, bringing its valuation to more than $1.4 billion and earning it unicorn… 5 TechCrunch — AI news-outlet 9d ago Researchers used Anthropic’s Claude to hack into OpenAI Security researchers used Anthropic’s Claude to exploit vulnerabilities in OpenAI’s systems, taking over employee accounts and gaining access to an internal code repository before reporting the flaws. 17 r/MachineLearning community 10d ago What studies isolate back-and-forth LLM interaction from one-way sharing and self-refinement [D] I'm trying to find out whether back-and-forth between two different LLMs improves task success beyond strong alternatives under a controlled resource budget — and I'd rather hear that it's already settled than spend money finding out. AI-assisted throughout: the protocol was… 7 r/LocalLLaMA community 10d ago Flyweight: open-source C++/CUDA engine for running MoE models bigger than your VRAM on one GPU + system RAM. First PyPI release, looking for contributors. Been building this for a few months, mostly for myself, and it just got a proper release so figured I'd post it. It's a native GGUF inference runtime with OpenAI/Anthropic-compatible APIs and a chat UI. The whole point is one consumer NVIDIA card + lots of RAM: MoE models that… 21 r/LocalLLaMA community 10d ago Made the horizontal open-source model for Jev with RLCD, and it surpasses all the Jev benchmarks. HF space, benchmark, model, repo Thanks for the exceptional support ( https://www.reddit.com/r/LocalLLaMA/comments/1wijo3e/i_literally_built_the_jev_architecture_one_year/ ) and for the dozens of requests to make a generic model, run benchmarks, and create an HF space so anyone can test it. So here you go,… 30 arXiv — Machine Learning research 10d ago OceanMoE: Structured Conditional Sparse Computation for Long-Horizon Multivariate Ocean Forecasting arXiv:2609.19768v1 Announce Type: new Abstract: Multivariate ocean forecasting must exploit shared evolution in a coupled ocean system while adapting to the heterogeneous statistical and dynamical characteristics of different prediction variables and locations. Fully shared… 36 arXiv — Machine Learning research 10d ago AURA: Adaptive Uncertainty-Routed Analysis for Email Threat Detection arXiv:2609.19873v1 Announce Type: new Abstract: Email spam and phishing attacks remain a critical security threat. Adversaries increasingly exploit large language models to craft contextually convincing malicious messages, and existing spam detection systems often struggle to… 10 arXiv — Machine Learning research 10d ago Intact-to-Amputee Transfer in Surface-EMG Gesture Decoding: Training Source and Calibration Budget arXiv:2609.20297v1 Announce Type: new Abstract: A recogniser trained on one person rarely transfers to the next, and useful performance usually demands a fresh round of labelled calibration from the end user. A systematic review of 1077 studies quantifies where the evidence is… 9 arXiv — Machine Learning research 10d ago Distributionally Robust Federated Learning with Multi-Source Data arXiv:2609.20501v1 Announce Type: new Abstract: Federated learning trains a shared model from private client data. In practice, data-generating distributions may differ, and the true mixture across clients is often unknown, making the underlying group distribution difficult to… 36 arXiv — NLP / Computation & Language research 10d ago From Parameters to Behaviors: A Survey of Model Fusion for Large Language Models arXiv:2609.19553v1 Announce Type: new Abstract: Model fusion integrates the capabilities from source models into a single target model. As of June 2026, Hugging Face hosts more than 2M models. This growing pool provides a rich base for model reuse and capability integration. Yet… 28 arXiv — NLP / Computation & Language research 10d ago V\={a}kQA: A Benchmark and Evaluation Study for Telugu Spoken Factoid Question Answering arXiv:2609.19879v1 Announce Type: new Abstract: Question answering has advanced rapidly with large language models, but predominantly for high-resource languages, in both text and spoken settings. Spoken question answering (SQA) benchmark for Telugu remains unexplored, and the… 11 arXiv — NLP / Computation & Language research 10d ago Viveka-Insight: a cross-lingual concept graph and citation-grounded retrieval resource over the complete works of Swami Vivekananda in English and Bengali arXiv:2609.20303v1 Announce Type: new Abstract: Classical philosophical corpora pose three compounding challenges for language resources: they exist in several languages without parallel alignment, their vocabulary is remote from that of contemporary readers, and generated text… 14 arXiv — NLP / Computation & Language research 10d ago Riemannian--Lorentz Fusion of Vision Transformers and State-Space Models arXiv:2609.19384v1 Announce Type: cross Abstract: Scaling deep learning faces critical bottlenecks: data exhaustion, exponential training costs, and resource concentration. Model merging combines pre-trained checkpoints without gradient descent, offering orders-of-magnitude… 5 arXiv — NLP / Computation & Language research 10d ago BurnRiSc: Toward Non-Invasive Burnout Screening in Open Source from Public Repository Signals arXiv:2609.19422v1 Announce Type: cross Abstract: Burnout is a chronic occupational syndrome, and open source is close to a worst case for it: maintainers absorb unbounded demand with no manager to reallocate work and no organization to notice decline. The cost is not only… 7 Vercel — AI dev-tools 10d ago Reproducing, disclosing, and fixing the libheif vulnerability with Hacktron and the maintainers In August 2026, Hacktron reported what looked like a remote code execution (RCE) vulnerability in Next.js image optimization. Their investigation found that the vulnerable code was not in Next.js itself, but upstream in libheif, an AVIF image decoder used by Next.js, ImageMagick… 23 Simon Willison community 10d ago Self-generated prompt injections in compaction summaries Self-generated prompt injections in compaction summaries In Our framework for reporting model misalignment OpenAI provide "six reports on unexpected or concerning model behavior we’ve observed in the last six months". This one here is my favorite: they caught some of their… 38 Vercel — AI dev-tools 10d ago Run Terminal-Bench and other Harbor evals on Vercel Sandbox You can now run Harbor evals on Vercel Sandbox. Harbor is the open-source harness behind Terminal-Bench , whose registry includes many other benchmarks such as SWE-bench, tau3-bench and OSWorld. Pass --env vercel to harbor run and each trial executes in its own isolated… 18 Vercel — AI dev-tools 10d ago The skills CLI now supports Notion hosted skills [email protected] adds Notion skills databases as an install source for agent skills . Notion skills are reusable agent skills written as Notion pages. Teams author, review, and update them in the workspace they already use, then install them into any agent the skills CLI supports.… 27 r/LocalLLaMA community 10d ago shots fired at dario from glm https://preview.redd.it/bqjrvgknt3qh1.png?width=2810&format=png&auto=webp&s=ec9f25709d6f2951a530228d430e0b8ecdbd8738 source super interesting read from glm as usual   submitted by   /u/Elux91 [link]   [comments] 12 The Information — AI news-outlet 10d ago Introducing The Information’s 50 Most Influential People in Tech Money, technical savvy and disruptive ideas all matter within the technology industry. But the most seismic force is influence. Those who have it build, fund and steer the companies that alter the world—again and again. Those without it can only rise so far. So who’s got… 37 r/LocalLLaMA community 10d ago We built an open-source GPU profiler you point an AI agent at, instead of reading traces yourself We've been tuning vLLM/SGLang/llama.cpp setups for a long time and got tired of the profiling part: nsys trace, open the GUI, squint, change a flag, repeat. The profilers assume a human is looking at the timeline. These days the thing doing our tuning is usually an agent, and it… 26 arXiv — Machine Learning research 11d ago FedPGT: Progressive Gradient Transmission for Vehicular Federated Learning over Time-Varying Channels arXiv:2609.18089v1 Announce Type: new Abstract: Vehicular federated learning (VFL) enables privacy-preserving collaborative model training for intelligent transportation systems, where communication resource allocation and gradient sparsification techniques have been explored to… 32 arXiv — Machine Learning research 11d ago Multi-Appliance Non-Intrusive Load Monitoring via Label-Preserving Aggregate Recomposition and Prediction Consistency arXiv:2609.18315v1 Announce Type: new Abstract: Non-intrusive load monitoring (NILM) estimates appliance power sequences from aggregate power, but models trained on source households commonly lose accuracy in unseen households. Aggregate power also contains loads from other… 23 arXiv — NLP / Computation & Language research 11d ago MudawanSn: A Gold-Standard Wolof-Arabic Parallel Corpus for Machine Translation arXiv:2609.17539v1 Announce Type: new Abstract: We present MudawanSn, a gold-standard resource of 1,271 sentence-aligned pairs manually translated from Wolof into Modern Standard Arabic (MSA). The source texts are drawn from the MasakhaNER corpus and cover politics, society,… 30 arXiv — NLP / Computation & Language research 11d ago T-SANDHI: Tone Sandhi-aware Adaptive Network with Decoupled Hybrid Injection for Low-resource Taiwanese Hokkien Speech Recognition arXiv:2609.18194v1 Announce Type: new Abstract: In Taiwanese Hokkien automatic speech recognition (ASR), prior studies often treat tone sandhi as a major challenge under the assumption that models fail to process implicit phonological variations. However, our experiments on… 36 arXiv — NLP / Computation & Language research 11d ago Behavior2Value: Benchmarking and Empowering LLMs for Consumer Value Measurement from E-commerce Behaviors arXiv:2609.18203v1 Announce Type: new Abstract: Human values are deep motivational orientations that shape human behaviors. In e-commerce, they reveal the stable drivers behind users' purchase decisions. Compared with short-term interests, consumer values better explain how… 26 arXiv — NLP / Computation & Language research 11d ago A Scalable Framework for Automated NER Annotation Correction in Low-Resource Languages arXiv:2609.18739v1 Announce Type: new Abstract: Poor quality or noisy annotations in Named Entity Recognition (NER), as in any other NLP task, make it challenging to achieve state-of-the-art performance. In this paper, we present a multi-step framework to enhance the annotation… 29 arXiv — NLP / Computation & Language research 11d ago Zero-Shot Cross-Lingual Recognition of Sign Language Handshapes arXiv:2609.18772v1 Announce Type: new Abstract: Sign language processing advances rapidly for high-resource languages such as American Sign Language (ASL), yet most of the world's sign languages lack the phonological annotations new methods require. We present the first… 29 arXiv — NLP / Computation & Language research 11d ago When Audit Quality Fails to Predict Downstream Utility: A Counterfactual Study of Synthetic-Data Selectors for Low-Resource African NLP arXiv:2609.18960v1 Announce Type: new Abstract: Quality-aware synthetic-data selection rests on a proxy: examples that an LLM judge rates as good should also help a downstream model learn. In a controlled replay in low-resource African-language classification, we show that this… 15 Vercel — AI dev-tools 11d ago Native Marketplace integrations now support custom environments You can now connect native Marketplace resources to custom environments . Previously, resource connections could only target production, preview, and development environments. Choose custom environments when connecting a resource from the Vercel dashboard, Vercel CLI, or REST… 4 Hacker News — AI on Front Page community 12d ago Salesforce Global Outage Article URL: https://status.salesforce.com/products/all Comments URL: https://news.ycombinator.com/item?id=49724488 Points: 221 # Comments: 126 18 Stratechery (Ben Thompson) community 12d ago Salesforce AI Force, Agents as UI, The Race to Headless Salesforce is abandoning UI as a moat, which is a very smart move because it's disappearing for everyone. 24 llama.cpp releases dev-tools 12d ago b10996 chat : force \n</think> on reasoning budget end for qwen3-coder ( #28869 ) Website: https://llama.app Attestations: https://github.com/ggml-org/llama.cpp/attestations/47858837 macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS… 29 arXiv — Machine Learning research 12d ago Evaluating Open-Weight E-Commerce Agents with Environment-Grounded Verification arXiv:2609.16093v1 Announce Type: new Abstract: A shopping conversation has many routes to the same cart, and a task-success rate reduces all of them to one score. We build a deterministic and reproducible e-commerce environment that precommits each trial's customer and… 8 arXiv — Machine Learning research 12d ago Multi-Label Proportion Learning for Sea-Ice Type Prediction arXiv:2609.16347v1 Announce Type: new Abstract: Sea-ice type prediction is important for climate monitoring, maritime navigation, and decision-making in polar regions. The main source of label data for this task is the ice chart, produced manually by ice analysts who interpret… 26 arXiv — Machine Learning research 12d ago OPEN-1B: A Fully Auditable Training Run arXiv:2609.17380v1 Announce Type: new Abstract: Open-source language models have a reproducibility problem. Despite releasing weights, training data, and recipes, none of them are provably reproducible due to the non-associativity of floating-point arithmetic. Deep learning… 19 arXiv — Machine Learning research 12d ago Nonsmooth Optimization via Orthogonalized Momentum arXiv:2609.13677v1 Announce Type: cross Abstract: Modern real application problems involve matrix-valued parameters, yet conventional optimizers treat them as vectors, thereby motivating matrix-aware methods that exploit input-output geometry, such as Muon which orthogonalizes… 4 arXiv — NLP / Computation & Language research 12d ago Style-Debiased DPO: Updating LLM Knowledge with Factuality-Aware Synthetic Preference Data arXiv:2609.16532v1 Announce Type: new Abstract: Continued pretraining (CPT) with data augmentation such as paraphrasing can store inside a large language model (LLM) the knowledge of a small source corpus. The stored knowledge, however, is not always retrieved correctly. We… 10 arXiv — NLP / Computation & Language research 12d ago PunGraph: Retrieval-Enhanced Phonetic-Semantic Graph Reasoning for Pun Understanding arXiv:2609.16557v1 Announce Type: new Abstract: Puns are a challenging form of figurative language that exploit phonetic similarity and semantic ambiguity to convey multiple meanings. Although large language models (LLMs) demonstrate strong language understanding capabilities,… 6 arXiv — NLP / Computation & Language research 12d ago Benchmarking Factual Robustness of LLMs via Multi-conversation Persuasion arXiv:2609.16777v1 Announce Type: new Abstract: As Large Language Models (LLMs) increasingly serve as primary knowledge retrieval interfaces, their robustness against \textit{persuasion attacks}---attempts to inject misinformation or enforce counterfactuals---has become a… 13 Page 3 of 10 · 500 articles ← Newer Older →