News / #security Tag Security 500 articles archived under #security · RSS Sign in to follow Latent.Space news-outlet 29d ago 🔬 The Lab of the Future Should Feel Like a Data Center — Andy Beam & Rafa Gómez-Bombarelli, Lila Sciences Lila is betting that science, not the internet, is the last untapped source of training data. We went to find out what that actually looks like in a room full of robots. 30 Simon Willison community 29d ago Quoting Linus Torvalds I realize that some people really dislike AI, but this is an area where I'm willing to absolutely put my foot down as the top-level maintainer. Linux is not one of those anti-AI projects, and if somebody has issues with that, they can do the open-source thing and fork it. Or… 28 Hugging Face Daily Papers research 29d ago Boogu-Image-0.1: Boosting Open-Source Unified Multimodal Understanding and Generation Abstract We introduce Boogu-Image-0.1, an open-source unified multimodal understanding and generation model family, comprising Base, Turbo, Edit, and Edit-Turbo variants. It delivers competitive performance in high-quality text-to-image generation, fast inference,… 10 arXiv — Machine Learning research 29d ago TSSM: Triaxial State Space Model for Global Station Weather Forecasting with Temporal-Variable-Historical Modeling arXiv:2607.13101v1 Announce Type: new Abstract: Global Station Weather Forecasting (GSWF) is pivotal for localized and extreme weather prediction over key regions. Despite efforts to exploit look-back windows, existing methods show limited accuracy gains and struggle with… 35 arXiv — Machine Learning research 29d ago PiVoT: A Variational Solution for Real-time Large-scale Multi-object Detection and Tracking under Heavy Clutter arXiv:2607.13891v1 Announce Type: new Abstract: Multi-object detection and tracking from noisy point clouds remain challenging in many data-scarce radar applications. Current Bayesian trackers based on Poisson measurement models offer a training-free solution but struggle to… 24 arXiv — Machine Learning research 29d ago Linear Independent Component Analysis via Optimal Transport arXiv:2607.14081v1 Announce Type: new Abstract: Linear Independent Component Analysis (ICA) recovers jointly independent source signals from their linear mixtures. To achieve this, classical ICA algorithms attempt to maximize non-Gaussianity, measured by negentropy, which is… 8 arXiv — Machine Learning research 29d ago SingGuard-NSFA: Extensible Guardrails for Agentic AI via Generative Reasoning and Real-Time Classification arXiv:2607.13081v1 Announce Type: cross Abstract: We present nsfaguard, a guardrail framework for securing agentic AI systems against operational threats, such as prompt injection, sensitive information extraction, malicious code requests, dangerous tool misuse, and resource… 5 Hugging Face Daily Papers research 29d ago MetaView: Monocular Novel View Synthesis with Scale-Aware Implicit Geometry Priors Abstract Current visual generation models are capable of producing high-quality content, yet they lack a coherent perception of the spatial structure. Existing generative novel view synthesis methods typically introduce explicit geometry priors, which enforce spatial consistency… 4 r/LocalLLaMA community 29d ago PSA: Nvidia's CMP 170HX Full Compute and Memory(80GB) may be unlockable via exploit If you don't know, the CMP 170HX is essentially a A100 that has had it's compute and memory crippled so it can only mine crypto. It was a product of the crypto craze, and was released shortly before the crypto crash. Well, I was scrolling around and found out, apparently, it can… 27 Vercel — AI dev-tools 29d ago Kimi K3 is now available on AI Gateway Kimi K3 from Moonshot AI is now available on AI Gateway. K3 is an open-source model with a 1M-token context window and native visual understanding, accepting text, image, and video inputs. Built for long-horizon software engineering, knowledge work, and deep reasoning, K3 is… 35 Simon Willison community 29d ago xai-org/grok-build, now open source xai-org/grok-build, now open source xAI's grok CLI tool faced severe community backlash yesterday when it became apparent that running the command in a directory could upload that entire directory to xAI's Google Cloud buckets. One user reported running it in their home… 26 Hacker News — AI on Front Page community 29d ago Governments, companies, nonprofits should invest in free, open source AI [pdf] Article URL: https://www.siegelendowment.org/wp-content/uploads/2026/07/fortune-david-siegel-open-source-ai.pdf Comments URL: https://news.ycombinator.com/item?id=48927095 Points: 242 # Comments: 84 20 Hacker News — AI on Front Page community 29d ago Grok Build is open source Article URL: https://github.com/xai-org/grok-build Comments URL: https://news.ycombinator.com/item?id=48926590 Points: 232 # Comments: 277 9 r/LocalLLaMA community 29d ago The Benchmarks of Thinking Machine's first open-source model Inkling   submitted by   /u/AloneCoffee4538 [link]   [comments] 33 llama.cpp releases dev-tools 29d ago b10031 tokenize : drop --stdin mutual-exclusion check ( #25672 ) match cli and completion, which don't enforce it macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework Linux: Ubuntu x64 (CPU) Ubuntu arm64 (CPU)… 13 TechCrunch — AI news-outlet 29d ago Hack suggests AI music generator Suno scraped YouTube for training data The hacker used an employee's credentials to access source code, which revealed how Suno scraped decades of audio. 29 r/LocalLLaMA community 29d ago Linus Torvalds tells people to stop attacking others for using AI The full quote: I realize that some people really dislike AI, but this is an area where I'm willing to absolutely put my foot down as the top-level maintainer. Linux is not one of those anti-AI projects, and if somebody has issues with that, they can do the open-source thing and… 23 r/MachineLearning community 29d ago All major robotics and VLA papers, ranked and benchmarked in a single place [P] Hi folks, There is now a dedicated Robotics page on Papers with Code that lists the major benchmarks, trending papers with linked code, and open-source artifacts. Find it here: https://paperswithcode.co/tasks/robotics… 24 r/LocalLLaMA community 29d ago The best model is the one you can actually run Don't get me wrong, all the big models are amazing, and every contribution to open source models is great. But I'm GPU poor and I can't use them locally. I'm currently running gemma-4-12b-it-qat-GGUF:UD-Q4_K_XL as my personal chat assistant, and I am so so happy with it! I still… 10 r/MachineLearning community 1mo ago How do you mathematically model an Unstoppable Force hitting an Immovable Object? [P] Or more broadly: how do you train a machine learning model to capture the nuances of entirely different, conflicting rule sets? I built a XGBoost classification pipeline to answer that. To stress-test the architecture across heterogeneous environments, I applied it to a highly… 22 OpenAI official-blog 1mo ago GPT-Red: Unlocking Self-Improvement for Robustness Explore GPT-Red, OpenAI’s automated red teaming system that uses self-play to improve AI safety, alignment, and prompt injection robustness. 14 arXiv — Machine Learning research 1mo ago Repairing Shape-Prior Shortcuts in Long-Range Single-Shot Fringe Projection Profilometry arXiv:2607.11928v1 Announce Type: new Abstract: Single-shot fringe projection profilometry (FPP) networks that regress depth directly can exploit a shape-prior shortcut, recovering depth from object boundaries rather than from fringe phase. On a photorealistic synthetic… 24 arXiv — Machine Learning research 1mo ago Scale-Aware Attention for Scarce Neural Data: An RG-Flow Transformer on Sleep-EDF EEG arXiv:2607.11950v1 Announce Type: new Abstract: Brain field potentials are scale-free: their power spectra follow a $1/f^{\beta}$ law whose aperiodic exponent $\beta$ tracks cortical state, and sleep depth in particular is a shift in $\beta$. We ask whether a transformer endowed… 11 arXiv — Machine Learning research 1mo ago Evaluating Reliability in Machine Learning Models for Early Chronic Kidney Disease Prediction: A Systematic Review of Data Leakage and Predictor Stability arXiv:2607.11963v1 Announce Type: new Abstract: The early detection of Chronic Kidney Disease using machine learning has attracted significant interest in healthcare-related computer science. Despite rapid advancements in this field, many reported studies remain inconsistent and… 33 arXiv — Machine Learning research 1mo ago LiteTopK: Exploiting the Curse of Dimensionality for a Fused Indexer-TopK Kernel in Long-Context Sparse Attention arXiv:2607.11976v1 Announce Type: new Abstract: Indexer-TopK, the operation to compute the scores and select the top-k candidates, is widely used by sparse attention kernels in large language models and vector retrieval in recommendation systems and vector databases. However,… 34 arXiv — Machine Learning research 1mo ago Adversarial Attacks on Online Handwriting using Salience-based Temporal Editing arXiv:2607.12500v1 Announce Type: new Abstract: Deep learning models for online handwriting recognition have been shown effective and are increasingly deployed in practical applications. However, their vulnerability to adversarial attacks is still a challenge. Existing… 26 arXiv — Machine Learning research 1mo ago Energy-Based Physics-Informed Form Finding for Clustered Tensegrity Structures arXiv:2607.12888v1 Announce Type: new Abstract: Tensegrity form-finding and physical property prediction are fundamental inverse problems in structural mechanics, which aim to determine equilibrium configurations and internal force distributions. These problems are challenging… 27 arXiv — Machine Learning research 1mo ago Contrastive Joint-Embedding Prediction for Representation Learning in Structural MRI arXiv:2607.11962v1 Announce Type: cross Abstract: Self-supervised learning offers a compelling approach for medical imaging, where labeled data are scarce and acquisition costs are high. We present COJEPA, a self-supervised framework for volumetric brain MRI that combines a… 28 arXiv — NLP / Computation & Language research 1mo ago Hybrid Continual Learning for Low-Resource Australian Aboriginal Language Identification arXiv:2607.11946v1 Announce Type: new Abstract: Language identification is an important step toward integrating endangered Australian Aboriginal languages (AALs) into speech technologies supporting language revitalisation and digital inclusion. However, extreme data scarcity… 9 arXiv — NLP / Computation & Language research 1mo ago Evaluating Health Misinformation in Low-Resource Languages: Integrating Small Language Models with a Culturally-Sensitive Responsible NLP Framework (Bangla as a Case Study) arXiv:2607.12336v1 Announce Type: new Abstract: Artificial Intelligence (AI) technologies, while serving as a foundational enabler for modern social media and digital health services, exert a bivalent effect by simultaneously acting as a combatant against and a spread vector for… 25 arXiv — NLP / Computation & Language research 1mo ago Translation as a Computationally Efficient Bridge: Feasibility of English BERT for Low-Resource Languages arXiv:2607.12612v1 Announce Type: new Abstract: BERT models have revolutionised Natural Language Processing (NLP) through their ability to process unstructured text across diverse domains. However, developing high-quality BERT models for non-English languages remains challenging… 20 r/MachineLearning community 1mo ago Things I got wrong building an incremental indexing pipeline [P] I've been working on incremental indexing pipelines lately, basically keeping a vector store in sync as the source data changes, and I keep finding the same bugs never show up until it's been running a while. Biggest one for me is deletes. I tested the "new doc comes in, gets… 10 r/LocalLLaMA community 1mo ago Me: one-shot programming is useless and should not be used as benchmark DeepSeek V4: hold my Atlas 500 SuperPod Credit goes to: kdzzzds on Bilibili Someone got access to a new version of DeepSeek V4 through A/B testing, and vibe coded / one-shot this No Man's Sky and Minecraft hybrid. Here is the chatlog: https://opncd.ai/share/fnOGJyIn And the source code:… 18 r/LocalLLaMA community 1mo ago Kimi K3 in the next few hours. Deepseek V4 GA later in the week. New Liquid models. New Mistral models sometime this month. And some rumours suggest GLM 5.5 is coming in August. Openweight AI is eating good. dam bois we eating good this week ngl, The velocity of the open_weight ecosystem right now is hitting a point where proprietary, closed-source APIs are losing their leverage on compute intelligence. When you have DeepSeek V4 dropping native MXFP4 mixtures of experts with massive… 10 TechCrunch — AI news-outlet 1mo ago Reflection inks $1B compute deal with Nebius Reflection AI has signed a $1 billion deal to access Nebius's compute. Reflection was founded in 2024 and is developing open source AI technology. 17 r/LocalLLaMA community 1mo ago Good podcasts I'm heading on vacation soon and want to download a few good podcasts about local LLMs, open-weight models, inference, tooling and the broader open-source AI ecosystem. Which podcasts or specific episodes do you genuinely recommend? I'm especially interested in technical… 18 r/MachineLearning community 1mo ago [N] AMA Reminder: Raffi Krikorian (CTO, Mozilla) Hello community, just a short reminder that Raffi Krikorian (CTO @ Mozilla) is live today for an AMA to discuss Mozilla's inaugural State of Open Source AI report. Topics include enterprise adoption, the real cost of "free"models, developer trust, Chinese open models and their… 28 arXiv — Machine Learning research 1mo ago ERP Data Provisioning Financial Control Testing arXiv:2607.09712v1 Announce Type: new Abstract: Financial control testing increasingly depends on representative enterprise resource planning (ERP) data in quality environments, yet direct production copies expose personal, supplier, banking, and commercially sensitive records.… 37 arXiv — Machine Learning research 1mo ago Quota Marketplace: Dynamic Pricing for Efficient Allocation of ML Training Resources arXiv:2607.09802v1 Announce Type: new Abstract: The escalating demand for Machine Learning (ML) training resources in recent years has resulted in a substantial gap between the high demand and the available supply. Efficient allocation of these scarce and expensive resources is… 23 arXiv — Machine Learning research 1mo ago Spectral Origins of the Self-Correction Blind Spot in Autoregressive Generation arXiv:2607.09803v1 Announce Type: new Abstract: Large autoregressive language models exhibit a self-correction blind spot: they reliably fix identical errors when attributed to an external source yet fail to fix the same errors in their own outputs. Prior work has documented… 12 arXiv — Machine Learning research 1mo ago RUBRIC: Realism--Utility Balanced Ranking for Imbalanced Classification arXiv:2607.09816v1 Announce Type: new Abstract: Class imbalance poses a fundamental challenge in risk-sensitive applications such as fraud detection and medical diagnosis, where minority-class samples are scarce yet critical for accurate classification. Existing oversampling… 22 arXiv — Machine Learning research 1mo ago ARMOR: Stabilizing On-Policy LLM RL with Off-Policy Anchor Samples arXiv:2607.10481v1 Announce Type: new Abstract: Reinforcement learning (RL) has significantly enhanced the reasoning capabilities of large language models (LLMs), yet the training process remains notoriously fragile. In this work, we investigate a critical source of this… 6 arXiv — Machine Learning research 1mo ago EvidentialRAG: Quantifying and Mitigating Information Conflict in Multi-Source Retrieval-Augmented Generation via Evidential Deep Learning arXiv:2607.10491v1 Announce Type: new Abstract: Retrieval-augmented generation grounds large language models in external evidence, but most pipelines still treat retrieved passages as deterministic and mutually consistent context. In open information environments, retrieved… 24 arXiv — Machine Learning research 1mo ago Learning to Fine-tune Foundation Models under Resource Limitations arXiv:2607.10694v1 Announce Type: new Abstract: We study the problem of optimal continual fine-tuning for a pre-trained Foundation Model deployed at a resource-limited device. At each time slot, a new batch of training data arrives, and the controller is faced with two options:… 22 arXiv — NLP / Computation & Language research 1mo ago Robust, Scalable Detection of Text Containment in Large Web-Crawled Corpora arXiv:2607.10020v1 Announce Type: new Abstract: We present FindMyText, an open-source Python package designed to efficiently assess whether a given text appears, in part or in full, within a text corpus. The tool builds on prior techniques for document fingerprinting, but… 25 arXiv — NLP / Computation & Language research 1mo ago Which Languages Transfer Best to Warlpiri? A Similarity-Based Study for Low-Resource ASR arXiv:2607.10256v1 Announce Type: new Abstract: This paper investigates how language similarity can improve cross-lingual transfer for automatic speech recognition (ASR) in extremely low-resource settings. Warlpiri, an Australian Aboriginal language, has very limited transcribed… 5 arXiv — NLP / Computation & Language research 1mo ago Polarization Detection: A Hybrid Approach with AfroXLMR-Social and DeBERTa for Low- and High-Resource Settings arXiv:2607.10312v1 Announce Type: new Abstract: The rapid proliferation of online polarization threatens social cohesion, necessitating robust automated detection systems that operate effectively across diverse linguistic contexts. This paper presents our system description for… 16 arXiv — NLP / Computation & Language research 1mo ago CAFE: A Compound-AI Factorial Evaluation Framework arXiv:2607.10380v1 Announce Type: new Abstract: We introduce CAFE (Compound-AI Factorial Evaluation), an open-source platform that brings design of experiments to the evaluation of compound AI systems (CAIS). Such systems expose many interchangeable choices - e.g. which… 6 arXiv — NLP / Computation & Language research 1mo ago Demographic Prompting at Scale: When More Attributes Hurt LLM--Human Agreement arXiv:2607.10590v1 Announce Type: new Abstract: We investigate how annotator demographic attributes, supplied as prompt cues, shape the alignment between large language model (LLM) predictions and human annotations across five tasks. Using five open-source LLMs, we… 26 arXiv — NLP / Computation & Language research 1mo ago Anamnesis: An Open-Source Platform for Large-Scale Backstory-Conditioned Survey Simulation arXiv:2607.10628v1 Announce Type: new Abstract: We present Anamnesis, an interactive system for demographically controllable survey simulation using large language models. Open-source, and designed for non-technical users/researchers, Anamnesis enables the prototyping and… 19 Page 9 of 10 · 500 articles ← Newer Older →