News / #security Tag Security 500 articles archived under #security · RSS Sign in to follow Hugging Face Daily Papers research 24d ago Beyond Success Rate: Cost-Aware Evaluation of Offensive and Defensive Security Agents Abstract Security-agent evaluations commonly measure peak offensive capability under generous inference budgets, emphasizing vulnerability discovery, exploit development, penetration testing, and CTF completion. Such measurements are useful but incomplete: in operational… 32 r/MachineLearning community 24d ago Exploring continual learning without replay buffers: Our findings using dynamic task-similarity routing [P] Hi, I’ve been doing some work in the continual learning space and wanted to share an open-source framework we put together called Coincidex, along with some architectural insights and failure modes we found along the way. Most conventional approaches to sequential task learning… 28 r/LocalLLaMA community 25d ago openbmb released MiniCPM5-2B, not yet available at huggingface According to source, it is the locally ranked AI model, the best among 4b models Source : https://x.com/i/status/2079088670804767114   submitted by   /u/Illustrious-Swim9663 [link]   [comments] 35 r/LocalLLaMA community 25d ago Is it possible to run a local model focused solely on "intelligence" and outsource its "knowledge" to web searches? I'm looking to run a very lightweight local model that acts as the brain, handling the logic and comprehension, while hooking it up to a web search tool to act as its memory and knowledge base.   submitted by   /u/chucrutcito [link]   [comments] 26 r/LocalLLaMA community 25d ago Sources: parts of the Trump administration are reigniting efforts to implement de facto bans on foreign open-source models, as Chinese AI models gain momentum   submitted by   /u/pscoutou [link]   [comments] 21 r/LocalLLaMA community 25d ago MiniCPM-Robot model series - MiniCPM-RobotManip & MiniCPM-RobotTrack 🚀 MiniCPM enters the physical world — enabling robots to understand, remember, and act. We open-source MiniCPM-Robot, our first embodied AI model series, including: 🤖 MiniCPM-RobotManip — a 1.5B general-purpose Vision-Language-Action (VLA) model for robotic manipulation. 🐕… 23 Hacker News — AI on Front Page community 25d ago Exploit brokers pay $500k for WordPress RCEs. I found one with GPT5.6 and $25 Article URL: https://slcyber.io/research-center/exploit-brokers-pay-500000-for-a-wordpress-rce-i-found-one-with-gpt5-6/ Comments URL: https://news.ycombinator.com/item?id=48975665 Points: 212 # Comments: 117 17 Hugging Face Daily Papers research 25d ago RAGU: A Multi-Step GraphRAG Engine with a Compact Domain-Adapted LLM Abstract Graph retrieval-augmented generation (GraphRAG) enhances large language models with structured knowledge, yet existing systems construct knowledge graphs in a single extraction pass, producing noisy entities and brittle retrieval. RAGU, an open-source modular GraphRAG… 4 arXiv — Machine Learning research 25d ago Who Became Financially Vulnerable After COVID-19? A Population-Level Machine Learning Analysis Using MEPS Data arXiv:2607.15446v1 Announce Type: new Abstract: The cost of healthcare remains a concern in the United States and may have been influenced by disruptions associated with the COVID-19 pandemic. This study examines healthcare financial vulnerability before and after the pandemic… 38 arXiv — Machine Learning research 25d ago Deep Learning Approaches for Sleep Apnea Classification from Polysomnographic EEG Signals arXiv:2607.15477v1 Announce Type: new Abstract: Sleep apnea diagnosis via polysomnography remains resource intensive and relies on time consuming manual data analysis and scoring. Recent work has demonstrated that central nervous system effects of sleep apnea events can be… 26 arXiv — Machine Learning research 25d ago An Auto-Scaling Approach for Serverless Environments Based on a Multi-Expert Consensus Mechanism arXiv:2607.15511v1 Announce Type: new Abstract: Serverless computing provides automatic resource management and pay-per-use execution, but effective autoscaling remains challenging because of dynamic workloads, cold-start latency, and dependencies among functions. We present a… 32 arXiv — Machine Learning research 25d ago Information-Directed Sampling for Causal Bandits arXiv:2607.15577v1 Announce Type: new Abstract: Causal bandits exploit structural relationships among variables to share information across interventions and accelerate the identification of high-reward decisions. In many applications, however, some variables cannot be directly… 28 arXiv — Machine Learning research 25d ago AquaAugmentor: A Novel Feature Augmentation Algorithm for Water Potability Prediction arXiv:2607.15775v1 Announce Type: new Abstract: Access to potable water is crucial for health, economic development, and sustainability. However, accurately classifying water quality remains a significant challenge due to the complexity and variability of water source data. This… 22 arXiv — Machine Learning research 25d ago DADiff: Diffusion-Driven Cross-Domain Policy Adaptation for Reinforcement Learning arXiv:2607.16090v1 Announce Type: new Abstract: Transferring policies across domains poses a vital challenge in reinforcement learning, due to the dynamics mismatch between the source and target domains. In this paper, we consider the setting of online dynamics adaptation, where… 18 arXiv — Machine Learning research 25d ago Closed-Loop Bayesian Bandit Encoder with GRAND Receiver for a Bursty Interference Channel arXiv:2607.15404v1 Announce Type: cross Abstract: Interleaving mitigates burst errors but introduces decoding delay and removes temporal error structure that a channel-aware decoder could exploit. We consider packet-level selection between a random linear code and the same code… 32 arXiv — Machine Learning research 25d ago Do Agents Dream of False Memories? Black-box Visual Attacks on Long-term Memory in Multimodal AI Agents arXiv:2607.15657v1 Announce Type: cross Abstract: Multimodal AI agents increasingly rely on persistent long-term memory to ground generation in past visual and textual episodes. We show that unconditional trust in visual data creates a critical vulnerability. We propose Lucid, a… 36 arXiv — NLP / Computation & Language research 25d ago From Articles to Premises: Building PrimeFacts, an Extraction Methodology and Resource for Fact-Checking Evidence arXiv:2605.06006v2 Announce Type: replace Abstract: Fact-checking articles encode rich supporting evidence and reasoning, yet this evidence remains largely inaccessible to automated verification systems due to unstructured presentation. We introduce PrimeFacts, a methodology and… 8 arXiv — NLP / Computation & Language research 25d ago AgentRedBench: Dynamic Redteaming and Integration-Aware Defense for LLM Agents over SaaS Integrations arXiv:2606.02240v3 Announce Type: replace-cross Abstract: Indirect prompt injection in tool-use agents is a concrete production threat: LLM agents read from integrations (third-party services such as Gmail, Salesforce, or Jira accessed through tool calls) whose response content… 36 arXiv — NLP / Computation & Language research 25d ago Crayotter: Traceable Multi-Agent Workflows for Long-Form Video Editing arXiv:2606.07636v2 Announce Type: replace-cross Abstract: Long-form video editing over heterogeneous footage requires agents to coordinate source selection, multimodal analysis, timeline construction, narration and subtitle alignment, rendering, and revision while exposing… 37 Simon Willison community 25d ago Quoting Sam Altman We have been having extensive discussions around open source strategy. We will discuss it more at our next board meeting, but one thing we’d like to do soon is to create a language model with the approximate capability of GPT-3 that can run locally on consumer hardware and… 16 r/LocalLLaMA community 25d ago given the increasing likelihood of an open source AI ban, what are the alternative channels for downloading models? the open ai exec in his "ai communism" post suggested a fraudulent FUD campaign and trump executice order against open source / chinese models. unfortunately that seems pretty likely to happen. i have always use huggingface for downloading models and they would be forced to… 34 r/LocalLLaMA community 25d ago I don't see how open-source AI models in the U.S. can successfully compete with those from China. Chinese startups benefit heavily from local government subsidies, state-backed banks offering ultra-low-interest loans, and long-term capital (without collateral)—a playbook China has successfully used across other tech sectors. In contrast, neither the U.S. federal nor state… 4 r/LocalLLaMA community 26d ago It could have been Meta Imma write some fan fiction for a second here if you indulge me. What we are seeing from the Chinese open models could have been Meta. As you can see from the name of this sub, they were the stars of open-source models, and the consistently poor decisions by their senior… 29 r/LocalLLaMA community 26d ago German SooFi team launches Soofi S 30B-A3B , an open-source Mixture-of-Experts (MoE) hybrid Mamba–Transformer foundation model for German and English.   submitted by   /u/epSos-DE [link]   [comments] 16 r/LocalLLaMA community 26d ago Arandu v0.6.5 available Repository: https://github.com/fredconex/Arandu Hello Guys, This is Arandu a open source app to easily launch models using llama.cpp, it tries to integrate all in one place with easy multi version management of llama.cpp, HuggingFace models search/download, intuitive arguments… 14 Hacker News — AI on Front Page community 27d ago TP-Link Kasa cameras leaked home GPS via unauthenticated UDP for 6 years Article URL: https://github.com/BadChemical/IoT-Vulnerability-Research-Public/blob/main/TP-Link_Kasa_EC71/Kasa_EC71.md Comments URL: https://news.ycombinator.com/item?id=48952565 Points: 203 # Comments: 83 22 r/LocalLLaMA community 27d ago A year ago you told me my open-source screen-watching app was flaky. You were right, so I spent the year fixing it with your feedback. Thank you r/LocalLLaMA c: !! TL;DR: This post is part update, mostly thank you for your support :)) Observer is an open-source app that lets local LLMs watch your screen and notify you (WhatsApp/SMS/email/Discord) when something happens. A year of your feedback later , setup went from "flaky and very… 31 Hacker News — AI on Front Page community 27d ago Mozilla: The state of open source AI Article URL: https://stateofopensource.ai/ Comments URL: https://news.ycombinator.com/item?id=48947825 Points: 204 # Comments: 141 26 r/LocalLLaMA community 28d ago Soofi S - 30B-A3B European Open Source Model I just saw that these had dropped. Still very much early days, but nice to see a new locally runnable foundation model, along with a couple of thinking preview versions. Anyone taken a look at this yet? Am intrigued to see how it holds up compared to Qwen 3.6 and Gemma 4 (my… 20 r/LocalLLaMA community 28d ago Chinese President Xi Jinping speaks at World AI Conference and reaffirms commitment to open source to promote"openness and win-win" From Vincent Chow on 𝕏: https://x.com/vince_chow1/status/2077947375964791028   submitted by   /u/Nunki08 [link]   [comments] 17 r/LocalLLaMA community 28d ago Will we get accessible open-source models again? Past April of 2026, all open-source LLMs have been in the 0.5T+ terrirory: MiniMax M3, Kimi-2.7-Code - now Kimi-3 (2.8T), GLM-5.2, Inkling by ThinkingMachines is also a 1T model and perhaps some more models I forget now. If chinese labs find it more profitable (I would not… 20 r/LocalLLaMA community 28d ago China’s Xi Touts Open-Source AI and Takes a Swipe at U.S. Dominance   submitted by   /u/pscoutou [link]   [comments] 19 Vercel — AI dev-tools 28d ago Data downloaded by Vercel Sandbox is now free Vercel Sandbox no longer bills for data it downloads from the internet. Installing packages, cloning a Git repository, or pulling artifacts and datasets from an external source does not count toward Sandbox Data Transfer usage. Traffic received on a Sandbox's exposed ports is… 16 Hugging Face Daily Papers research 28d ago VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding Abstract Recent advances in video understanding have spanned motion, long video, and streaming interaction, driving this field toward real-world applications. Despite this progress, current open-source models remain limited in several ways. They often struggle to generalize… 36 arXiv — Machine Learning research 28d ago RENEW: Towards Learning World Models and Repairing Model Exploitation from Preferences arXiv:2607.14180v1 Announce Type: new Abstract: World models are widely used in offline reinforcement learning (RL) to improve sample efficiency and generate experience beyond a fixed dataset. However, they are vulnerable to model exploitation where data coverage is thin. Prior… 9 arXiv — Machine Learning research 28d ago Supervised Fine-Tuning vs. In-Context Learning: An Equilibrium Analysis of LLM Personalization under Congestion arXiv:2607.14371v1 Announce Type: new Abstract: Large Language Models (LLMs) have revolutionized AI services, but a critical tension emerges: while personalization improves model performance, it consumes scarce computational resources that users must share. When should a user… 26 arXiv — Machine Learning research 28d ago Integration Matters: Rollout-Based Training for Constrained Diffusion Models arXiv:2607.14398v1 Announce Type: new Abstract: Constrained generative models aim to produce samples that satisfy complex feasibility constraints while remaining faithful to the data distribution. Existing constrained generation methods typically enforce constraints either… 13 arXiv — Machine Learning research 28d ago Deep-learning Causal Retrieval Optimization for Efficient e-commerce Distribution in Pinterest arXiv:2607.14161v1 Announce Type: cross Abstract: Pinterest is where people turn inspiration into action as users browse ideas, then take steps toward realization, often by discovering shoppable content. To support this journey, we must distribute commerce content when it helps,… 29 arXiv — Machine Learning research 28d ago Never Too Late for Force: Accelerating VLA Post-Training with Reactive Force Injection arXiv:2607.14236v1 Announce Type: cross Abstract: Pretrained vision-language-action (VLA) policies provide strong language-conditioned manipulation knowledge, but they remain largely vision-driven and can struggle once manipulation enters contact states where the scene is… 27 arXiv — NLP / Computation & Language research 28d ago MAPS: Modeling Co-Existing Subjective Perspectives and Shared Meaning in Multi-Agent Cognitive Dialogue arXiv:2607.14110v1 Announce Type: new Abstract: Human dialogue involves more than exchanging information; it also expresses beliefs, emotions, and subjective cognitive styles. Yet current AI dialogue systems often enforce semantic uniformity, sacrificing diversity and… 38 arXiv — NLP / Computation & Language research 28d ago Routing Ceilings Are Domain-Independent: Structural Prior Injection in Code Security Vulnerability Detection arXiv:2607.14628v1 Announce Type: new Abstract: Large language models (LLMs) exhibit a well-documented gap between latent capability and consistent activation: the router hypothesis posits that models possess the knowledge to solve a task but lack reliable internal routing to… 20 arXiv — NLP / Computation & Language research 28d ago Gold-Guided Programmatic Distillation for Financial Reasoning over Hybrid Tables and Text arXiv:2607.14709v1 Announce Type: new Abstract: Financial question answering over hybrid tabular and textual data may require multi-source reasoning and precise numerical computation. While large language models (LLMs) can generate intermediate reasoning steps, natural-language… 28 arXiv — NLP / Computation & Language research 28d ago Measuring How Students Rely on Generative AI in Academic Writing: Development and Multi-Source Validation of the Generative AI Reliance Types Scale (GenAI-RTS) arXiv:2607.14301v1 Announce Type: cross Abstract: As generative AI (GenAI) becomes increasingly embedded in undergraduate academic writing, how students rely on these tools, rather than simply whether they use them, has become a central question for learning, academic integrity,… 31 arXiv — NLP / Computation & Language research 28d ago MedFailBench: A Clinician-Built Open-Source Benchmark for Medical AI Safety Boundary Inspection arXiv:2607.15166v1 Announce Type: cross Abstract: Most medical AI benchmarks measure whether a model knows the correct answer. MedFailBench asks a different question: which safety boundary failed? We present a clinician-built synthetic benchmark and failure atlas that labels… 27 arXiv — NLP / Computation & Language research 28d ago Pretraining Data Can Be Poisoned through Computational Propaganda arXiv:2607.15267v1 Announce Type: cross Abstract: Poisoning pretraining data can introduce harmful behaviors to LMs that are difficult to detect and mitigate. Prior work on poisoning pretraining data has largely exploited established data sources such as Wikipedia, which do not… 9 Hugging Face Daily Papers research 28d ago BadWAM: When World-Action Models Dream Right but Act Wrong Abstract World-action models (WAMs) are emerging as a promising foundation for embodied control: rather than predicting actions alone, they learn representations that couple action generation with future world prediction. This coupling is often viewed as a source of robustness,… 11 r/LocalLLaMA community 28d ago Will we have a 27B model with Fable capabilities in 5 months? History says yes If history is any indication, open-source models in the 27B dense range should have caught up to what the US government banned two weeks ago because they thought they were too dangerous in less than half a year from now. Qwen 3.6 27B outperformed models that were considered… 4 Ars Technica — AI news-outlet 28d ago It's official: EU will force Google to share search data and open up AI on Android Google says these changes could endanger user privacy and security. 6 Hacker News — AI on Front Page community 28d ago Microsoft Comic Chat is now open source Article URL: https://opensource.microsoft.com/blog/2026/07/16/microsoft-comic-chat-is-now-open-source/ Comments URL: https://news.ycombinator.com/item?id=48936426 Points: 225 # Comments: 60 8 r/LocalLLaMA community 29d ago Mozilla’s State of Open Source AI Report Link to the report itself is at the bottom of the page.   submitted by   /u/pegasus912 [link]   [comments] 36 Page 8 of 10 · 500 articles ← Newer Older →