News / #rag Tag Rag 500 articles archived under #rag · RSS Sign in to follow arXiv — Machine Learning research 7d ago Geometric Mean Pooling for Equal-Weight Multiplicative Coarse-Graining arXiv:2609.21876v1 Announce Type: new Abstract: As an alternative to the additive and extremal biases of average and max pooling, we introduce Geometric Mean Pooling (GMP), a signed pooling operator that combines the product of feature signs with the geometric mean of feature… 25 arXiv — Machine Learning research 7d ago COMPLEX: A Closed-Form Certified Embedding of Multiparameter Persistence Modules arXiv:2609.22012v1 Announce Type: new Abstract: Every multiparameter persistence vectorization we know of carries a one-sided Lipschitz upper bound and nothing below it: without a lower gauge there is no sense in which the features are faithful, and no per-prediction guarantee… 23 arXiv — Machine Learning research 7d ago dSTAR: Straggler Tolerant and Byzantine Resilient Distributed SGD arXiv:2412.07151v1 Announce Type: cross Abstract: Distributed model training needs to be adapted to challenges such as the straggler effect and Byzantine attacks. When coordinating the training process with multiple computing nodes, ensuring timely and reliable gradient… 16 arXiv — NLP / Computation & Language research 7d ago HERMES: Contrast-Aware Knowledge Graph Reasoning from Clinical Notes for Patient Outcome Prediction arXiv:2609.20825v1 Announce Type: new Abstract: Clinical predictive models often rely on structured Electronic Health Record data, such as time-series and procedure codes. While recent approaches have begun leveraging unstructured clinical notes, they typically encode them as… 20 arXiv — Machine Learning research 7d ago Fragment-Aware Vision Transformers for Fresco-Fragment Style Classification arXiv:2609.21012v1 Announce Type: cross Abstract: Artistic style classification is usually studied on complete artworks, where models can exploit global composition, spatial organisation, and iconographic structure. In archaeological settings, however, artworks often survive… 15 arXiv — NLP / Computation & Language research 7d ago Do small language models know what they don't know? arXiv:2609.20824v1 Announce Type: new Abstract: We explore whether entropy-based confidence signals can be leveraged to improve the accuracy of Small Language Models (SLMs) with fewer than 3 billion parameters, running entirely on consumer hardware. We evaluate seven distinct… 29 arXiv — NLP / Computation & Language research 7d ago COAL-SQL: Coverage-Guided Augmentation and Failure-Driven Learning for Text-to-SQL Post-Training arXiv:2609.20842v1 Announce Type: new Abstract: Text-to-SQL translates natural-language questions into executable SQL queries, but open-source large language models still require task-specific post-training for complex, real-world SQL generation. Effective post-training requires… 28 arXiv — NLP / Computation & Language research 7d ago MIRAGE: Multi-Perspective Creative Language Model Reasoning with Reinforcement Learning Guidance arXiv:2609.21554v1 Announce Type: new Abstract: Recent advances in Large Language Models (LLMs) have revolutionized artificial intelligence and how human interact with AIs. Despite impressive advancements, LLMs struggle with complex mathematical, scientific, and logical tasks.… 14 arXiv — NLP / Computation & Language research 7d ago Analysing the Linearity of Linguistic Relations in Language Model Embedding Spaces arXiv:2609.21655v1 Announce Type: new Abstract: We propose a framework to analyse how strongly different linguistic relations are linearly encoded in language model embedding spaces. We formalise linear encoding via a constrained linear approximation over related and unrelated… 38 arXiv — NLP / Computation & Language research 7d ago CIBuzzBench: A Benchmark for Cross-Lingual Understanding of Chinese Internet Buzzwords arXiv:2609.21722v1 Announce Type: new Abstract: Chinese social media has generated a vast and continually evolving lexicon of internet buzzwords whose meanings are often non-literal and deeply rooted in local cultural and pragmatic contexts. Existing research has primarily… 24 arXiv — NLP / Computation & Language research 7d ago Per-Aetiology Contrastive Severity Embeddings with Phonological Pseudo-Labelling for Multilingual Dysarthric Speech arXiv:2609.21789v1 Announce Type: new Abstract: Most multilingual dysarthria-severity systems either train on a single aetiology-language pair or pool heterogeneous aetiologies into one label space. We test that pooling assumption with four matched HuBERT-base contrastive… 22 arXiv — NLP / Computation & Language research 7d ago Cross-Lingual Parkinson's Disease Severity Assessment Using Pre-trained Speech Embeddings: A Multi-Class Evaluation arXiv:2609.20875v1 Announce Type: cross Abstract: Parkinson's disease (PD) often manifests through speech impairments, facilitating accessible, non-invasive, and cost-effective severity assessment for early diagnosis and progression tracking. Despite advances in speech… 31 r/LocalLLaMA community 9d ago Tuning Qwen 3.8 27B and OMP as a coding agent on 2× 3090s Oh My Pi + vLLM on two 3090s. Average wait per turn went from 28s to 7s, mostly from changing omp settings: explicit effort level on every role (unset ones defaulted to xhigh) thinking_token_budget of 7500 maxTokens 8k → 32k (file writes were getting cut off) tool output over 10… 23 r/MachineLearning community 9d ago I posted my embedding migration project here, it got a lot of attention, so I added the features you guys said were missing [R] A little while ago, I posted about embedflow https://github.com/arnsri33/embedflow and it got a lot of attention. The basic idea was pretty simple keep your existing embedding index for candidate retrieval → rerank "k" candidates with the new embedding model → progressively… 6 GitHub Blog — AI & ML official-blog 9d ago Should you read the code, is RAG dead, and did Skills kill MCP? We dive into these questions and other AI hot takes on the latest episode of the GitHub Podcast. The post Should you read the code, is RAG dead, and did Skills kill MCP? appeared first on The GitHub Blog . 31 The Information — AI news-outlet 9d ago Communicate Technical Topics to a Non-Technical Audience with Google Gemini In a room full of IT professionals, words and phrases like “tokenization,” “retrieval-augmented generation,” and “context window” need no explanation. But when technology leaders present to their CEOs and other company leaders about topics like AI, they often find that they need… 10 arXiv — Machine Learning research 10d ago Generative Query Suggestion via Intent Coverage and Query-Level Credit Assignment arXiv:2609.19209v1 Announce Type: new Abstract: Generative query suggestion aims to enhance user engagement by anticipating user intents and recommending relevant follow-up queries. A central challenge is to generate slates whose individual queries are useful while the slate… 26 arXiv — Machine Learning research 10d ago QoS-Aware Federated Learning for Multimodal In-Cabin Interaction in Smart Vehicles arXiv:2609.20123v1 Announce Type: new Abstract: Modern smart vehicles leverage multimodal sensors, ranging from high-bandwidth vision systems to low-rate physiological monitors, to provide personalized in-cabin services. However, integrating high-fidelity multimodal fusion with… 34 arXiv — Machine Learning research 10d ago Subdomain-aware representation compression for pretrained image embeddings arXiv:2609.20213v1 Announce Type: new Abstract: Dimensionality reduction is a well-known technique for improving space efficiency, typically applied uniformly across an entire dataset. This paper investigates the possibilities of using dimensionality reduction techniques for… 16 arXiv — Machine Learning research 10d ago Seismic Site Response Prediction from Sparse Observations Using Finite-Element-Pretrained Latent Dynamics arXiv:2609.20451v1 Announce Type: new Abstract: Numerical site-response predictions often deviate from observations, yet correcting these discrepancies is difficult because records are limited in both sensor coverage and number of events. This study proposes the Transfer-Enabled… 35 arXiv — NLP / Computation & Language research 10d ago Less Is More: Graph-free Multimodal RAG via Multi-signal Late Fusion arXiv:2609.19417v1 Announce Type: new Abstract: Graph-based retrieval-augmented generation (RAG) is widely used for multimodal, cross-document question answering. However, building corpus-level graphs is expensive, slow to query, and difficult to maintain. We present TrioRAG, a… 18 arXiv — NLP / Computation & Language research 10d ago Semantic Layer Induction from Raw Telemetry via Hierarchical LLM and RAG Abstraction arXiv:2609.19615v1 Announce Type: new Abstract: Modern applications generate massive volumes of raw telemetry data, but translating those noisy, heterogeneous event streams into actionable business insights remains a fundamental challenge. Data engineers and analysts expend… 26 arXiv — NLP / Computation & Language research 10d ago Lens: Bringing the Right Semantic Perspective into Focus for Training-Free Multimodal Representation Learning arXiv:2609.20252v1 Announce Type: new Abstract: High-quality representations are essential for a wide range of downstream tasks. Dedicated embedding models are explicitly optimized for representation learning, yet their training data are often more limited in scale and diversity… 34 arXiv — NLP / Computation & Language research 10d ago Embedding Models Measure in Peculiar Ways arXiv:2609.20821v1 Announce Type: new Abstract: Embedding spaces define notions of semantic similarity and distance. We study whether those embeddings reflect physical measurements of mass, distance, time and volume, which admit a unique, objective notion of semantic equivalence… 12 arXiv — NLP / Computation & Language research 10d ago CovR: Coverage-Aware Hardware Verification via Reasoning-Guided Reinforcement Learning arXiv:2609.19189v1 Announce Type: cross Abstract: Design verification remains one of the most resource-intensive stages of hardware development, often consuming up to 70% of the total design effort. While recent work has explored using Large Language Models (LLMs) to automate… 33 arXiv — NLP / Computation & Language research 10d ago Scientific Image Quality Assessment via Multi-modal Retrieval-Augmented Generation arXiv:2609.19634v1 Announce Type: cross Abstract: This paper proposes a Retrieval-Augmented Generation (RAG) framework for scientific image quality assessment, designed to simultaneously address both the understanding track (SIQA-U) and the scoring track (SIQA-S) of the SIQA… 11 r/LocalLLaMA community 10d ago I just realized Courage used local LLMs to solve his problems before any of us ever did.   submitted by   /u/swagonflyyyy [link]   [comments] 17 Hugging Face Daily Papers research 11d ago Assessing nnU-Net Generalization across Brain Tumor Populations in BraTS-GoAT 2026 Abstract BraTS-GoAT evaluates tumor segmentation across heterogeneous populations. We trained a conventional 3D nnU-Net on 1,351 labeled cases using five-fold cross-validation and 1,000 epochs per fold. The final predictor averaged all folds and applied test-time mirroring. On… 11 r/LocalLLaMA community 11d ago Intel releases OpenVINO 2026.4 UPDATE https://github.com/ggml-org/llama.cpp/pull/29009 MERGED More Gen AI coverage and frameworks integrations to minimize code changes New models supported: On CPU: Gemma-3n On CPU, GPU: Kokoro-82M, Qwen3-VL-4B with eagle3, Qwen3-ASR, Muse Glimmer 30B, Qwen 3.8 27B, Gemma4… 22 Hugging Face Daily Papers research 11d ago Flattening Every Memory Peak in Long-Context Mixture-of-Experts Training Abstract Training a Mixture-of-Experts (MoE) model at long context or large batch size fails as soon as any one component's peak allocation exceeds device memory, so the target is every peak at once, not the average footprint. Four are left unbounded by the parallelism plans in… 5 arXiv — Machine Learning research 11d ago Beyond Static RAG: An Adaptive, Tri-Metric Routing Framework for Efficient Long-Context Inference on Commodity GPUs arXiv:2609.17564v1 Announce Type: new Abstract: Deploying retrieval-augmented generation (RAG) on commodity GPUs such as the NVIDIA T4 (16 GB VRAM) exposes a practical failure mode we call the Compression Paradox: neural prompt compression can add key-value (KV) cache contention… 7 arXiv — Machine Learning research 11d ago DSD: Learning Diverse and Reusable Motor Skills via Diffusion Skill Discovery arXiv:2609.17682v1 Announce Type: new Abstract: Humans efficiently learn new tasks by reusing a rich repertoire of motor skills across different goals and contexts. A similar strategy can also be used to enable simulated characters to efficiently perform new tasks by leveraging… 12 arXiv — Machine Learning research 11d ago EdgeReMIND: A Scalable, Top-Ranked Memorization Baseline for Temporal Multi-Relational Link Prediction arXiv:2609.17916v1 Announce Type: new Abstract: Temporal link prediction on the Temporal Graph Benchmark 2.0 (TGB 2.0) faces a scalability ceiling: on the benchmark's three largest datasets, every existing embedding method runs out of memory or exceeds the time budget. These… 13 arXiv — Machine Learning research 11d ago Beyond Embedding Transfer: Component Roles in Grokking Transfer and Stability arXiv:2609.18078v1 Announce Type: new Abstract: Warm-start transfer can make algorithmic tasks generalize rapidly, yet it is unclear which model components provide the gain and whether that gain remains stable under continued optimization. We study cross-operator transfer on… 37 arXiv — Machine Learning research 11d ago Behavioral Fingerprinting and Navigation Prediction in Web Browsing arXiv:2609.18273v1 Announce Type: new Abstract: Web browsing often appears ephemeral: users visit a few websites, complete a task, and move on. However, even short fragments of browsing activity can contain rich and structured behavioral signals. In this work, we conduct a… 27 arXiv — Machine Learning research 11d ago The evolution of sex for artificial intelligence: a population-genetic framework for multigenerational model populations arXiv:2609.18560v1 Announce Type: new Abstract: Some aspects of AI development resemble a population process in which models are specialised, retrained on the output of peers, or combined by averaging weights. These practices lead to generations of models, in the biological… 18 arXiv — Machine Learning research 11d ago Changepoint-Aware World Models: Detecting Dynamics Shifts and Recovering by Forgetting Stale Replay in Model-Based RL arXiv:2609.18950v1 Announce Type: new Abstract: A robot's learned model of its own dynamics is only valid until those dynamics change: actuators wear, payloads shift, and joints stiffen. A model-based agent that keeps training as if nothing happened adapts slowly, dragged back… 26 arXiv — NLP / Computation & Language research 11d ago Enhancing Extubation Failure Prediction with LLM-Derived Features from Respiratory Therapy Clinical Notes arXiv:2609.17532v1 Announce Type: new Abstract: Invasive mechanical ventilation is a lifesaving therapy, but timely, safe discontinuation is essential to preventing extubation failure (EF) and related risks to health. We present a novel approach to EF prediction that leverages… 30 arXiv — NLP / Computation & Language research 11d ago DualSQL: Text-to-SQL with Multi-Agent Reinforcement Learning arXiv:2609.18135v1 Announce Type: new Abstract: State-of-the-art Text-to-SQL systems are typically multi-agent pipelines centered around two fundamental tasks: schema linking and SQL generation. However, existing work trains separate models for each task, failing to leverage the… 24 arXiv — NLP / Computation & Language research 11d ago TeochewBench: A Human-Reviewed Benchmark for Teochew Hanzi Translation arXiv:2609.18156v1 Announce Type: new Abstract: Teochew has a substantial speaker community and exhibits distinctive lexical, syntactic, and pragmatic features, yet textual resources for evaluating large language models remain limited. We present TeochewBench, a human-reviewed… 28 arXiv — NLP / Computation & Language research 11d ago LocQE: Principled Domain Adaptation for Localisation Quality Estimation by Leveraging Post-Edits arXiv:2609.18720v1 Announce Type: new Abstract: Learned quality estimation (QE) models such as COMETKiwi are widespread and work well for general machine translation evaluation. However, they are known to struggle on unseen domains, limiting their performance in a real-world… 7 arXiv — NLP / Computation & Language research 11d ago Beyond frequency measures: Can contextual embeddings capture meaning change in scientific texts? arXiv:2609.18804v1 Announce Type: new Abstract: Identifying technological trends is a core scientometric task, yet traditional frequency-based approaches struggle to capture substantial meaning shifts of domain-specific terms. We hypothesise that contextual embeddings can… 37 arXiv — NLP / Computation & Language research 11d ago A Benchmark Suite and Ground-Truth Methodology for Formal Verification of IEC 61131-3 Ladder Diagram Programs arXiv:2609.18994v1 Announce Type: new Abstract: We present the first benchmark suite for formal verification of Programmable Logic Controller (PLC) programs that combines controlled ground truth with coverage of both textual (Structured Text, ST) and graphical (Ladder Diagram,… 30 arXiv — NLP / Computation & Language research 11d ago MIRAGE: How Conversation State Shapes Historical Evidence Use in Multimodal Personal Agents arXiv:2609.19059v1 Announce Type: new Abstract: Multimodal large language model (MLLM) agents are increasingly used as personal assistants for long-running tasks. Their utility depends on continuity: agents must retrieve and use earlier evidence across dialogue, files, and… 38 arXiv — NLP / Computation & Language research 11d ago Playing log(N)-Questions over Wikipedia Abstracts: Communication Efficiency Between Paired Frontier Models arXiv:2609.19113v1 Announce Type: new Abstract: We evaluate six frontier language models on the two-agent $\log(N)$-Questions game. A questioner sees $N$ Wikipedia lead paragraphs and must identify a secretly chosen target using exactly $\log_2 N$ yes/no questions. An answerer… 5 arXiv — NLP / Computation & Language research 11d ago ScienceIDE: Turning World's Scientific Codebase into Agent Learnable Environments arXiv:2609.19134v1 Announce Type: new Abstract: Scientific code repositories encode decades of human knowledge in executable models, methods, and tools. Yet fragmented toolchains, implicit domain conventions, and specialized correctness criteria make this knowledge difficult to… 10 Hugging Face Daily Papers research 11d ago ScienceIDE: Turning World's Scientific Codebase into Agent Learnable Environments Abstract Scientific code repositories encode decades of human knowledge in executable models, methods, and tools. Yet fragmented toolchains, implicit domain conventions, and specialized correctness criteria make this knowledge difficult to convert into reliable learning… 19 The Information — AI news-outlet 11d ago Apollo Backs AI Startups in Bid to Bankroll Hardware Boom Apollo Global Management, known for leveraged buyouts and hefty financing deals, has started investing in AI and related hardware startups with the ambition to lend money to them in the future. In one previously unreported deal, Apollo has already invested in the low tens of… 28 Vercel — AI dev-tools 11d ago Hobby projects now retain fewer deployments to free up storage Hobby projects now retain fewer deployments past the 30-day retention window. Hobby teams get 10GB of Deployment Storage . Every deployment you keep uses some of it, and going over the limit can block you from deploying until you free some up. Deployment Retention for Hobby… 15 Vercel — AI dev-tools 11d ago Secure Compute and Static IP builds start 64% faster Builds using Secure Compute or Static IPs now start 64% faster, with the average time from deployment creation to build start dropping from 6.7 seconds to 2.4 seconds. Previously, each build waited for a new build container to boot with its network configuration. These builds… 18 Page 3 of 10 · 500 articles ← Newer Older →