News / #security Tag Security 500 articles archived under #security · RSS Sign in to follow arXiv — NLP / Computation & Language research 15d ago Symphony of Bias: Exploring Gender Associations with Musical Instruments in Multimodal LLMs arXiv:2607.26355v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly embedded in everyday life and widely used for information seeking, raising concerns about their potential to perpetuate social biases and reinforce stereotypes. In this study, we… 38 arXiv — NLP / Computation & Language research 15d ago MediaWiki Code2Code Search: Neural Retrieval for the Semantic Discovery of Open-Source Software Entities arXiv:2607.26766v1 Announce Type: cross Abstract: Code search in large-scale ecosystems is often hindered by the lexical gap between user queries and implementation details, alongside the trade-off between the low latency of traditional Information Retrieval (IR) and the… 5 arXiv — NLP / Computation & Language research 15d ago On-Policy Distillation for LLM Safety: A Routing Approach to Template-Robust Realignment arXiv:2607.27081v1 Announce Type: cross Abstract: Fine-tuning is the dominant paradigm for specializing large language models (LLMs), yet it exposes a critical vulnerability: malicious data providers can embed harmful behaviors into downstream corpora, creating models that… 5 Vercel — AI dev-tools 15d ago Shopify and Vercel are rebuilding Hydrogen for faster storefronts Shopify and Vercel Open source and runtime agnostic, runs anywhere JavaScript does Standard Actions brings agentic commerce to every storefront Feature development cut from months to a week for retailers like Global Retail Brands Shopify powers commerce for millions of merchants… 30 r/MachineLearning community 15d ago Open-source tabular model validation toolkit TanML needs feedback [D] We’re developing TanML, an MIT-licensed automated model-validation toolkit for tabular machine-learning models. TanML runs locally and provides an end-to-end workflow covering data profiling, preprocessing, feature-power ranking, model development, evaluation, drift analysis,… 5 Simon Willison community 15d ago AI Worming through Word AI Worming through Word Neat new prompt injection variant by Håkon Måløy, who found a way to upgrade prompt injection attacks against Microsoft Word to full self-replicating worms: An attacker places hidden instructions in a document that is later used as source material in… 36 Hacker News — AI on Front Page community 15d ago Keychron announces first open-source firmware for gaming mice Article URL: https://www.digitalfoundry.net/news/2026/07/keychron-announces-first-open-source-firmware-for-gaming-mice Comments URL: https://news.ycombinator.com/item?id=49099715 Points: 213 # Comments: 87 4 Ars Technica — AI news-outlet 15d ago Anthropic is finding bugs faster than Microsoft can fix them Microsoft is on a mad dash behind the scenes to patch exploits before hackers find them. 12 Hacker News — AI on Front Page community 15d ago Show HN: Open-source engine running Gemma 4 26B in 2 GB RAM on any M-series Mac Hi HN, I built a specialized inference engine for running 4-bit Gemma 4 26B-A4B-IT on any M-series Mac using about 2 GB of RAM. It is called TurboFieldfare and is written in Swift and Metal. I have always adored on-device AI. It feels like magic that you can run a powerful NN on… 29 Hugging Face Daily Papers research 16d ago GLI-AL: A Multi-Modal Glioma MRI Label Resource with Unified Anatomy-Lesion Labels Abstract Existing BraTS-GLI datasets provide a widely used benchmark for adult glioma MRI segmentation, but their task definition focuses on tumor subregions and does not systematically represent coexisting white matter hyperintensities (WMH). In joint segmentation settings,… 38 Hugging Face Daily Papers research 16d ago CodeNib: A Multi-View Data System for Serving Repository Context to Coding Agents Abstract Coding agents repeatedly search, navigate, and retain context from evolving repositories, but disconnected indexes, language servers, and task-local histories force repeated discovery and obscure lifecycle costs. CodeNib builds reusable lexical, dense, and structural… 37 Hugging Face Daily Papers research 16d ago Mapping CVEs to MITRE ATT&CK Techniques: A Curated Gold-Set Classifier and the Limits of LLM-Assisted Label Expansion Abstract We present a reproducible pipeline for mapping Common Vulnerabilities and Exposures (CVEs) to MITRE ATT&CK Enterprise techniques from free-text vulnerability descriptions. Rather than relying on the CWE->CAPEC->ATT&CK derivation chain, whose table-expansion artifacts we… 14 arXiv — NLP / Computation & Language research 16d ago Where Steering Signals Come From: Activation Source Selection in Activation Steering arXiv:2607.25270v1 Announce Type: new Abstract: Activation steering controls language models by adding vectors or features to hidden states at inference time, but the upstream source of these steering signals is often treated as a secondary detail. We study this source choice as… 25 arXiv — NLP / Computation & Language research 16d ago IRIS: Reusable Identity Representations from Frozen LLMs for Entity Alignment arXiv:2607.25579v1 Announce Type: new Abstract: Entity alignment (EA) identifies entities across knowledge graphs (KGs) that refer to the same real-world object. Conventional EA methods mainly exploit explicit graph structures and textual fields, which often provide insufficient… 7 arXiv — NLP / Computation & Language research 16d ago The Effect of Text Chunk Size on Retrieval-Augmented Generation Performance arXiv:2607.24767v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) systems have emerged as a powerful process for allowing large language models (LLMs) to retrieve relevant information to use as source material during text generation. A critical yet… 11 arXiv — NLP / Computation & Language research 16d ago On the Use of LLMs for Specialised Terminology: A Good Alternative to Corpora? arXiv:2607.24784v1 Announce Type: cross Abstract: Specialised translation relies on the use of documentary and terminological resources, including corpora. These resources are particularly useful for terminology. However, their compilation and exploitation have several… 6 arXiv — NLP / Computation & Language research 16d ago Ranked by Position: Order Sensitivity as an Exploitable Attack Surface in LLM Listwise Recommenders arXiv:2607.24869v1 Announce Type: cross Abstract: Large language models (LLMs) used as listwise rerankers in recommendation systems suffer from position bias when serializing candidate sets into prompts. We show this order sensitivity creates an exploitable attack surface: an… 15 arXiv — NLP / Computation & Language research 16d ago Building Large-Scale English-Romanian Literary Translation Resources with Open Models arXiv:2509.07829v4 Announce Type: replace Abstract: Literary translation has recently gained attention as a distinct and complex task in machine translation research, yet translation by small open models remains an open problem, particularly for low-resource languages such as… 12 r/LocalLLaMA community 16d ago Nvidia is expected to raise GeForce RTX GPU prices again by up to 30%   submitted by   /u/ab2377 [link]   [comments] 21 r/LocalLLaMA community 16d ago Agenta: an open-source Claude Cowork alternative where you can use self-hosted models (and any harness) Hey r/LocalLLaMA, I’m Mahmoud from Agenta. We built a self-hosted, more flexible, alternative to Claude Cowork . This short video shows how it works. I use it to build AI coworkers for my startup, like the marketing agent in the video. Or AI automations, like a daily report to… 33 Ars Technica — AI news-outlet 16d ago We now have a better understanding how OpenAI hacked into Hugging Face 10 days passed from OpenAI models exploiting JFrog Artifactory 0-day to release of a patch. 33 r/MachineLearning community 16d ago NeurIPS-side prompt injection triggering ethics reviewers? [D] Does anyone experience a similar story that some reviewers reporting ethical issue due to NeurIPS-side prompt injection for catching LLM-reviewers? Even ethics reviewers were not informed about this conference-side manipulation…   submitted by   /u/dontknowwhattoplay… 19 TechCrunch — AI news-outlet 16d ago Fish Audio raises $50M seed to build AI voice models for creators and enterprises Since launching last year, the startup today has more than 8 million people using the open-source or hosted version of its models, and now generates annual recurring revenue of $21 million. 4 r/MachineLearning community 17d ago NeurIPS 2026 AI-generated reviews [D] I'm really confused about what the point of the prompt injection was (speaking as an author). Is it just a study? I would really prefer that they took action against the AI-generated reviews. Obviously, we cannot assume that the reviewers were copy-pasting the output from the… 5 arXiv — NLP / Computation & Language research 17d ago Semalith v1.4: A Calibrated 184M Safety Classifier Achieving State-of-the-Art Prompt-Injection Detection at 44x Fewer Parameters than Llama-Guard-3-8B arXiv:2607.22545v1 Announce Type: cross Abstract: Deploying large language models in financial-services and agentic settings requires safety classifiers that simultaneously handle prompt injection, regulatory compliance, and general harm, a combination no existing open guardrail… 21 arXiv — Machine Learning research 17d ago Directional Influence Function: Estimating Training Data Influence in Constrained Learning arXiv:2607.23388v1 Announce Type: new Abstract: As constrained learning becomes increasingly common, models are trained under explicit feasibility requirements to enforce fairness, safety, robustness, regulariza- tion, and physics or logic constraints. Understanding how training… 25 arXiv — Machine Learning research 17d ago When Can Depth Replace Precision? A Resource Theory of Quantized Neural Computation arXiv:2607.23390v1 Announce Type: new Abstract: When can additional low-bit residual computation replace missing numerical precision for a fixed input-output map? We model a quantized residual system over a fixed horizon as a pure schedule selecting fields from a declared… 4 arXiv — Machine Learning research 17d ago Generalization bounds and sample complexity for remaining useful life prediction from complete degradation trajectories arXiv:2607.23454v1 Announce Type: new Abstract: Data-driven remaining useful life (RUL) prediction requires complete degradation trajectories for training, yet such run-to-failure data are scarce and expensive. Practitioners currently lack principled guidance on how many failure… 24 arXiv — Machine Learning research 17d ago Extreme Volatility Warning under Label Scarcity via Multi-Source Anomaly Fusion arXiv:2607.23682v1 Announce Type: new Abstract: Early warning of extreme market volatility is central to financial risk management, but actionable events are rare, nonstationary, and often triggered by exogenous information shocks. In our CSI~300 setting, only $\sim$80 positive… 38 arXiv — NLP / Computation & Language research 17d ago Explaining GAND: A Resource on Gender-Ambiguous Natural Data & Contrastive Attribution arXiv:2607.22546v1 Announce Type: new Abstract: Machine translation (MT) systems continue to produce gender-biased translations. In a time where self-expression is paramount, mistranslations based on default behaviour and stereotyping can lead to harm for users of these systems.… 11 arXiv — NLP / Computation & Language research 17d ago PatiGonit22K: A Comprehensive Dataset for Solving Complex Bengali MWPs arXiv:2607.22859v1 Announce Type: new Abstract: Mathematical Word Problems (MWPs) are an important benchmark for evaluating natural language understanding and quantitative reasoning. Despite recent progress in high resource languages, Bengali remains underexplored due to the… 24 arXiv — NLP / Computation & Language research 17d ago Simple Language Normalization Wins: Cross-Lingual Speaker Verification for the TidyVoice 2026 Challenge arXiv:2607.22923v1 Announce Type: new Abstract: Cross-lingual mismatch remains a key source of overall degradation in modern speaker verification. The TidyVoice2026 Challenge targets this setting with text-independent verification, comprising 3,666 training and 808 development… 29 arXiv — NLP / Computation & Language research 17d ago BERT-based Models vs. Large Language Models for Low-Resource Named Entity Recognition: A Comparative Study on Marathi arXiv:2607.23344v1 Announce Type: new Abstract: Named Entity Recognition (NER) for low-resource languages such as Marathi remains a challenging task due to limited annotated resources and linguistic complexity. Although recent Large Language Models (LLMs) have demonstrated… 31 arXiv — NLP / Computation & Language research 17d ago Guiding Language Models to Be More Empathetic: Culturally Sensitive Mental Health Advice Generation Through Human-LLM Collaboration arXiv:2607.23538v1 Announce Type: new Abstract: Despite recent advances in large language models (LLMs), their ability to generate empathetic mental health counseling responses in low-resource languages remains largely unexplored. To address this gap, we curate 625 authentic… 9 arXiv — NLP / Computation & Language research 17d ago Pointer-Augmented Autoregressive Generation of Patent Claims with Joint Topology and Content Decoding arXiv:2607.24040v1 Announce Type: new Abstract: Autoregressive decoders emit flat token sequences and cannot enforce hierarchical constraints across output segments, a limitation that becomes acute in patent claim generation, where a claim set forms a dependency forest whose… 38 arXiv — NLP / Computation & Language research 17d ago BioSentinel at EXIST 2026: Soft-Label Optimization with XLM-RoBERTa for Sexism Intent Classification in Memes arXiv:2607.24137v1 Announce Type: new Abstract: This paper describes the BioSentinel team's participation in EXIST 2026 Task 2.2: Source Intention in Memes, part of the CLEF 2026 evaluation campaign. The task requires classifying the communicative intent behind memes as direct,… 8 arXiv — NLP / Computation & Language research 17d ago CAGE: Cognitive Attribution Graphs for Faithful Inline Citation Generation in Long-Form Question Answering arXiv:2607.24236v1 Announce Type: new Abstract: Long-form question answering increasingly relies on retrieved evidence to make LLM outputs verifiable, with inline citations tracing claims to source documents. However, existing systems often attach citations that are topically… 12 arXiv — NLP / Computation & Language research 17d ago Systematic Analysis of Large Language Models and Transformer-Based Machine Translation for English-Tamil and Tamil-English Across Diverse Datasets arXiv:2607.24515v1 Announce Type: new Abstract: The challenge of Machine Translation for low resource languages such as Tamil is primarily caused by the restricted amount of parallel data for these languages, as well as their substantial amount of domain variation and… 14 arXiv — NLP / Computation & Language research 17d ago Source-Aware Reranking for Retrieval-Augmented Generation: A Reliability Prior Approach arXiv:2607.22584v1 Announce Type: cross Abstract: Standard Retrieval-Augmented Generation pipelines rank retrieved documents by semantic similarity alone, without accounting for source provenance or credibility. This work evaluates a simple and interpretable modification to RAG… 5 Vercel — AI dev-tools 17d ago Vercel Sandbox supports forking Vercel Sandbox now supports forking with Sandbox.fork() . The fork starts from the source's current snapshot and inherits its config and environment variables. If the source is running, it forks the latest saved state, not the live in-memory state. If the source has no snapshot,… 14 Hugging Face Daily Papers research 17d ago Interactive Training 2: Auditable Control Plane for Live Model Training Abstract Experiment trackers show how training is progressing, but changing a live run still usually requires trainer-specific code. We present Interactive Training 2, an open-source control plane for steering training through a shared protocol. Training applications declare… 26 TechCrunch — AI news-outlet 17d ago OpenAI’s Hugging Face breach has reignited the debate over alignment and control OpenAI's Hugging Face breach has reignited debate over AI alignment and control, exposing competing views on whether increasingly capable AI should be better aligned, better contained, or both. 14 NVIDIA Developer Blog official-blog 17d ago NVIDIA Ising Enables Fully Automated Quantum Computer Calibration with Enhanced In-Context Learning NVIDIA Ising Calibration is an open source vision language model (VLM) designed to interpret diagnostic outputs from quantum processors and determine how they... 9 r/LocalLLaMA community 17d ago Nvidia CEO Jensen Huang defends Open Source AI by saying distillation is fundamental to learning Nvidia CEO Jensen Huang “Distillation - learning from AI, learning from other people, and learning from other sources of knowledge, is fundamental to intelligence. We are constantly learning from one another. AI also has to learn from something.” Since using AI Desktop 98 , I… 23 r/LocalLLaMA community 18d ago The entire tech industry (save for Anthropic) has come out in favor of open source AI. So what happens next? Will Anthropic change its lobbying efforts? Not likely. Now the gaslighting begins: “Nobody is trying to ban open source.”   submitted by   /u/SignificantLegs [link]   [comments] 34 r/LocalLLaMA community 18d ago Meta has confirmed that it will release an open source model in the future https://preview.redd.it/k97l56d8ypfh1.png?width=606&format=png&auto=webp&s=2a2e2156bea56b25f5709e8f1df2bf82525eb089 https://x.com/alexandr_wang/status/2081501627836661928?s=20   submitted by   /u/External_Mood4719 [link]   [comments] 38 Smol AI News news-outlet 18d ago not much happened today **Moonshot** released the **Kimi K3** open-weights model, a **2.8T-parameter MoE** with **104B active parameters**, **896 experts**, and **1M-token context** featuring native visual understanding. The release includes open-source infrastructure like **FlashKDA**, **MoonEP**, and… 33 arXiv — Machine Learning research 18d ago Physiological Signals as a Forensic Modality for Talking-Face Deepfake Detection arXiv:2607.21776v1 Announce Type: new Abstract: Talking-face (TF) deepfake generation synthesizes photore- alistic facial video from a static source image and an au- dio signal, producing forgeries that current image-based detectors consistently fail to identify. Unlike… 9 arXiv — NLP / Computation & Language research 18d ago LeAct: Learning to Reason from Expert Actions arXiv:2607.21856v1 Announce Type: cross Abstract: Modern reasoning models depend on reasoning data, today sourced from human annotations or distilled from stronger LLMs. However, a rich and largely untapped source of supervision lies in expert systems (e.g., game engines,… 17 arXiv — Machine Learning research 18d ago Optimization of time-consuming experimental conditions using pseudo-experimental data guided by adaptive polynomial regression arXiv:2607.22238v1 Announce Type: new Abstract: Bayesian optimization (BO) is an optimization method that sequentially proposes the next candidate explainable variables for optimizing target variables by balancing exploration and exploitation. BO is often used under a limited… 27 Page 5 of 10 · 500 articles ← Newer Older →