News / #paper Tag Research papers 500 articles archived under #paper · RSS Sign in to follow r/LocalLLaMA community 19h ago Adding logit penalty for "wait", "maybe" and "perhaps" to Qwen models improves their accuracy Meta came out with a banger paper https://arxiv.org/pdf/2606.00206 , but it did not look at various quantizations supported in llama.cpp. So I did a run on 50 random MATH-500 questions ( https://huggingface.co/datasets/HuggingFaceH4/MATH-500 ) and ran it on various quantizations… 10 Hacker News — AI on Front Page community 1d ago DeepSeek Elastic Compute (DSec) Article URL: https://arxiv.org/abs/2609.22978 Comments URL: https://news.ycombinator.com/item?id=49859112 Points: 200 # Comments: 61 33 r/LocalLLaMA community 2d ago Jev vs. Kev: open-source Jev alternative tested side by side We hosted Kev 4B (Jared Palmer's Apache-2.0 fine-tune of Qwen3.5-4B) and ran it side by side with Jev on the same endpoint to see how it compares. We built a fresh set of 362 items published after both models shipped (new arXiv papers, Stack Exchange questions, GitHub issues),… 9 arXiv — Machine Learning research 3d ago Stable and Faithful Explanations for Knowledge Tracing arXiv:2609.28502v1 Announce Type: new Abstract: Knowledge tracing (KT) models predict student performance opaquely, limiting pedagogical action. This study contributes a validation protocol testing predictive competitiveness (RQ1), explanation stability (RQ2) and… 11 arXiv — Machine Learning research 3d ago SMILESGNN: Interpretable Clinical Toxicity Prediction via SMILES-Graph Cross-Attention Fusion arXiv:2609.28553v1 Announce Type: new Abstract: Drug toxicity prediction is critical for reducing late-stage attrition in drug discovery, yet remains challenging due to severe class imbalance, scaffold-based generalization, and the clinical need for interpretable predictions.… 7 arXiv — Machine Learning research 3d ago CFD Correction of Open Tip Clearance Flow in a Compressor Cascade Using VAE Latent Space Adaptation arXiv:2609.28558v1 Announce Type: new Abstract: CFD predictions of open tip clearance flow in compressor cascades are subject to discrepancies relative to experiments, while experimental observations are sparse and high-resolution experimental ground truth is unavailable. This… 20 arXiv — Machine Learning research 3d ago CARE: Condition-Aware Representation Regularization for Diffusion Models arXiv:2609.28561v1 Announce Type: new Abstract: Recent advances in diffusion models highlight the importance of representation regularization for improving sample quality and training efficiency. However, commonly used regularization methods often overlook the built-in… 25 arXiv — Machine Learning research 3d ago SpaFactor: Lightweight Spatial Context-Aware Gene Program Modeling for Histology-to-Transcriptomics Inference arXiv:2609.28563v1 Announce Type: new Abstract: Spatial transcriptomics (ST) profiles gene expression within tissue architecture, but its cost and experimental complexity limit routine use. Predicting spatial expression from routinely available hematoxylin and eosin (HE) images… 34 arXiv — NLP / Computation & Language research 3d ago When Explanations Cannot Be Read: Measuring and Correcting SHAP and LIME Rendering for Right-to-Left Languages arXiv:2609.28565v1 Announce Type: cross Abstract: Post hoc explanation methods such as SHAP and LIME are widely used to interpret text classifiers, but their visualizations are mainly designed for left-to-right languages. When applied to right-to-left (RTL) languages such as… 32 arXiv — Machine Learning research 3d ago Leakage-Safe Machine Learning for Hydrogen Embrittlement Detection in 316L Stainless Steel: A Region-Held-Out Evaluation of Texture and Deep Features in SEM Micrographs arXiv:2609.28567v1 Announce Type: new Abstract: Scanning electron microscopy (SEM) is routinely used to characterize the microstructural changes caused by hydrogen embrittlement (HE) in structural steels. Machine learning can automate this characterization, but models are often… 6 arXiv — Machine Learning research 3d ago Time-Series Foundation Models That Understand Data Revisions arXiv:2609.28576v1 Announce Type: new Abstract: Historical observations are not always fixed: statistical agencies revise previously published values as new evidence arrives. Forecasting from a contemporary download can therefore expose a model to information unavailable at the… 13 arXiv — Machine Learning research 3d ago Uncovering Residential PV-EV Co-Adoption from Smart-Meter Data: Load Archetypes and Detection for Demand-Side Planning arXiv:2609.28578v1 Announce Type: new Abstract: The increasing adoption of electric vehicles (EVs) and rooftop photovoltaic (PV) systems is reshaping residential electricity demand and creating new challenges for demand-side management (DSM), tariff design, and low-voltage… 25 arXiv — Machine Learning research 3d ago Auditability Is Not One Property: Rule Overlap, Behavioural Agreement, and Composition in Reinforcement Learning arXiv:2609.28581v1 Announce Type: new Abstract: Reinforcement-learning (RL) policies are often distributed as opaque neural checkpoints, while training logs show that a run occurred without explaining what the policy learned. We study whether independently trained policies can… 24 arXiv — Machine Learning research 3d ago SGA: Uncertainty Quantification for Multi-Step Forecasting in Time Series Foundation Models arXiv:2609.28582v1 Announce Type: new Abstract: The recent emergence of Time Series Foundation Models (TSFMs) has significantly advanced multi-step forecasting performance, enabling accurate predictions over extended future horizons. However, existing TSFMs often suffer from… 5 arXiv — Machine Learning research 3d ago TAM-Chain: Multi-Scale Thyroid Cytology Classification via Absorbing Markov Chains and Shannon Entropy Uncertainty Quantification for False-Negative Suppression and Domain-Shift Adaptation arXiv:2609.28590v1 Announce Type: new Abstract: Background & Problem: Thyroid Fine-Needle Aspiration Biopsy (FNAB) cytology based on the Bethesda System plays a pivotal role in early thyroid cancer detection; however, deep learning approaches face substantial challenges… 34 arXiv — Machine Learning research 3d ago Learning to Discover Interesting Mathematics arXiv:2609.28603v1 Announce Type: new Abstract: Recently, Large Language Models (LLMs) have been increasingly able to solve advanced mathematical problems, including many that have been open for decades. This opens the door to expansion of mathematical knowledge at unprecedented… 4 arXiv — Machine Learning research 3d ago Physics-Informed Self-Supervised Learning for Joint Wire Calibration and Interaction Position Reconstruction in Multi-Wire Parallel Plate Avalanche Counters arXiv:2609.28604v1 Announce Type: new Abstract: Scientific instruments require accurate calibration to convert detector signals into reliable physical observables. Conventional calibration procedures typically rely on dedicated calibration measurements, analytical response… 15 arXiv — Machine Learning research 3d ago UO-FIE: Combining Exact-Label Supervision with Graded Utility for Factivity Inference arXiv:2609.28605v1 Announce Type: new Abstract: The Factivity Inference Evaluation 2026 (FIE2026) classifies Chinese context-hypothesis pairs into nine ordered factivity intervals. Its evaluation metric rewards both exact predictions and proximity to the correct interval, while… 19 arXiv — Machine Learning research 3d ago fable.intermittent: benchmarking probabilistic forecasting methods for intermittent time series arXiv:2609.28607v1 Announce Type: new Abstract: Intermittent time series are common in spare-parts demand and retail sales. Since the cost of forecast errors is typically asymmetric, decisions such as inventory control require the full predictive distribution rather than a point… 17 arXiv — Machine Learning research 3d ago RLVR landscapes for iterated multiplications can be benign: Insights from spin-glass theory arXiv:2609.28625v1 Announce Type: new Abstract: Despite the importance of reinforcement learning with verifiable rewards (RLVR), the extent to which it can learn new reasoning capabilities remains debated. Here we study the optimization landscape of RLVR on algorithmic tasks,… 36 arXiv — Machine Learning research 3d ago OPDiv: Optimal Selection of Top-K High-Scoring, Diverse Compounds arXiv:2609.28665v1 Announce Type: new Abstract: A virtual screening campaign may produce thousands of promising candidates, but only a small number can be purchased, synthesized, or tested. The practical question is how to select a set of compounds that both rank well and are… 21 arXiv — Machine Learning research 3d ago Beyond Static Graph World Models: Learning Stochastic Latent Dynamics over Evolving Topologies arXiv:2609.28670v1 Announce Type: new Abstract: Graph-based world models have recently emerged as a means of learning transitions over relational state representations. However, existing approaches are largely limited to fixed-topology graphs or deterministic, fully observable… 22 arXiv — Machine Learning research 3d ago Thinking Leakage: A Causal Audit of NoThink Post-Training in Hybrid Reasoning Models arXiv:2609.28682v1 Announce Type: new Abstract: Post-training hybrid reasoning models in NoThink mode has attracted growing interest as a way to improve performance while keeping inference fast. However, these gains may draw on thinking behavior already accessible through the… 4 arXiv — Machine Learning research 3d ago Federated Learning of AnDE Classifiers arXiv:2609.28695v1 Announce Type: new Abstract: This work presents a federated framework for training Averaged $n$-Dependence Estimators (AnDE) in distributed environments. The proposed method focuses on the discriminative setting, where model weights are learned locally and… 16 arXiv — Machine Learning research 3d ago LabFactory: Building and Evaluating Executable AI Labs arXiv:2609.28697v1 Announce Type: new Abstract: Scientific tasks specify a desired capability, but realizing it often requires building a computational system tailored to the task---acquiring data, designing representations, training models, implementing tools, and deciding how… 27 arXiv — Machine Learning research 3d ago Upholding Robustness in Federated Learning: Trends, Emerging Strategies, and Research Opportunities arXiv:2609.28722v1 Announce Type: new Abstract: While Federated Learning (FL) has been widely adopted for protecting user privacy in machine learning, it remains vulnerable to various robustness challenges, including performance-impairment risks, information-stealing threats,… 13 arXiv — Machine Learning research 3d ago Policy Complexity, Reaction Time, and Bounded Rationality in Reinforcement Learning arXiv:2609.28737v1 Announce Type: new Abstract: Biological agents do not learn under conditions of unlimited computation. For humans, learning and choice are shaped by constraints on perception, attention, and working memory, which limit how much state information guides… 22 arXiv — Machine Learning research 3d ago Evaluating Cross-region Generalization for Wavelet-Diffusion Precipitation Downscaling arXiv:2609.28749v1 Announce Type: new Abstract: Diffusion models have shown strong potential for kilometer-scale precipitation downscaling, but their performance in geographically unseen regions and event regimes remains insufficiently understood. Building on the wavelet… 18 arXiv — Machine Learning research 3d ago The Mechanics of Delta Learning: Target Design for Generalizable Scientific Machine Learning arXiv:2609.28782v1 Announce Type: new Abstract: In scientific machine learning, $\Delta$-learning trains models on residual errors relative to physical baselines, assuming that more accurate baselines with smaller residual scales inherently improve downstream performance. Here,… 11 arXiv — Machine Learning research 3d ago Vector Bellman Theory for Multichain Robust Average-Reward Markov Decision Processes arXiv:2609.28792v1 Announce Type: new Abstract: Robust average-reward Markov decision processes provide a fundamental framework for long-term performance optimization under uncertainty, and can have optimal long-run rewards that depend on the initial state. This state dependence… 9 arXiv — Machine Learning research 3d ago Monitoring Urban Traffic Dynamics at Fine Spatiotemporal Resolution Using Distributed Acoustic Sensing and Deep Learning arXiv:2609.28793v1 Announce Type: new Abstract: Mapping the distribution of traffic dynamics at high spatiotemporal resolution is a fundamental question in transportation research. Distributed acoustic sensing (DAS), an innovative seismic observation tool, emerges as a promising… 38 arXiv — Machine Learning research 3d ago Stream Recursion Model (SRM) arXiv:2609.28809v1 Announce Type: new Abstract: Mechanistic interpretability seeks to make verifiable statements about the internal behavior of large language models (LLMs). Many interpretability techniques struggle to scale with the increasing size and depth of architectures.… 7 arXiv — Machine Learning research 3d ago When Does Unsupervised Learning Succeed or Fail? A PoS Perspective on Reconstruction-Based Anomaly Detection arXiv:2609.28832v1 Announce Type: new Abstract: Reconstruction-based unsupervised learning can fail in two opposing ways: a model may reconstruct anomalies too accurately or discard valid nominal variation. Using the Pursuit of Subspaces hypothesis, we characterize these… 21 arXiv — NLP / Computation & Language research 3d ago LastOPD: Taming Collapse in Latent On-Policy Distillation arXiv:2609.28845v1 Announce Type: cross Abstract: On-policy distillation (OPD) corrects a student on the responses it writes, but its signal is the teacher's next-token distribution: it tells the student what the teacher says but misses how it thinks. Latent supervision promises… 28 arXiv — Machine Learning research 3d ago Image Fidelity is Not Field Fidelity: Joint Thermodynamic Reconstruction and Error Localization in Neural Tomography arXiv:2609.28868v1 Announce Type: new Abstract: Neural fields for scientific tomography are optimized from 2D images, but the actual quantity of interest is often a latent 3D physical field. Because the forward map is many-to-one, low 2D image error need not certify a correct 3D… 21 arXiv — Machine Learning research 3d ago Response-state Learning for Transferable Vibrational Spectroscopic Characterization with Electron Prior arXiv:2609.28935v1 Announce Type: new Abstract: Vibrational spectral prediction can become inaccurate when localized stereoelectronic environments perturb intermediate response states and high-risk response units dominate characteristic spectral fingerprints, making prediction… 21 arXiv — Machine Learning research 3d ago Spectral Graph Neural Networks with Hermite Polynomials: A Comprehensive Study arXiv:2609.28979v1 Announce Type: new Abstract: We study spectral graph neural networks built from Hermite polynomials and propose HermNet, a simple model that combines a nodewise predictor with normalized Hermite propagation. Its sparse recurrence requires neither… 5 arXiv — Machine Learning research 3d ago Automatic Rank Allocation for Low-Rank Adaptation in Large Language Models via lp Regularization arXiv:2609.28998v1 Announce Type: new Abstract: Low-rank adaptation (LoRA) has become a popular parameter-efficient fine-tuning method for large language models. A key challenge in LoRA is how to determine the rank of each adaptation matrix, as rank directly controls its… 29 arXiv — Machine Learning research 3d ago Learning from Mixed-Quality Deployment Experience for Robot Manipulation arXiv:2609.29000v1 Announce Type: new Abstract: Robot policies deployed in real environments naturally accumulate mixed-quality experience, including successful executions, partial progress, and failures. Although these rollouts provide valuable information for further learning,… 21 arXiv — Machine Learning research 3d ago Growth-Inspired Graph Generation and Inverse Design of Mechanical Lattices via Dot Matrices Database Augmentation and GCNN arXiv:2609.29024v1 Announce Type: new Abstract: Natural load-bearing and transport networks are not assembled in a single step; they emerge through a temporally ordered process of growth, branching, reinforcement, and loop formation. Inspired by this developmental logic, this… 29 arXiv — Machine Learning research 3d ago Generative Atmospheric Super-Resolution from Heterogeneous In Situ Observations through Composable Interfaces arXiv:2609.29027v1 Announce Type: new Abstract: Atmospheric observations are sparse, heterogeneous, and unevenly distributed, whereas many generative atmospheric models learn distributions over regularly gridded multivariate states. Once pretrained, diffusion models can supply… 27 arXiv — Machine Learning research 3d ago BranchShine-CR: Compact Multilingual IPA Transcription with Self-Conditioned CTC and Consistency Regularization arXiv:2609.29069v1 Announce Type: new Abstract: We introduce BranchShine-CR, a 25M-parameter model for multilingual transcription into the International Phonetic Alphabet (IPA). It combines log-mel features, a rotary-position E-Branchformer encoder, intermediate self-conditioned… 36 arXiv — Machine Learning research 3d ago Physics and Data Driven Transformer-Mamba Framework for Flow Field arXiv:2609.29087v1 Announce Type: new Abstract: While deep learning accelerates expensive partial differential equation solving in computational fluid dynamics (CFD), existing methods like PINNs and FNOs often struggle with generalization, noise robustness, and physical… 24 arXiv — Machine Learning research 3d ago Where Does Exactly-Once Live? Model, Harness, and Tool-Contract Effects on Duplicate Side Effects in LLM Agents arXiv:2609.29095v1 Announce Type: new Abstract: When a tool-using agent's write times out or returns a server error, the action may already have taken effect. Retrying blindly duplicates it -- a second charge, a second announcement, a second deployment -- while giving up skips… 13 arXiv — Machine Learning research 3d ago Downside-Controlled Online Forecast Combination under Delayed and Revised Outcomes arXiv:2609.29096v1 Announce Type: new Abstract: Post-hoc correction adjusts a forecaster that cannot be retrained, such as a foundation model, but a correction fitted where errors are stable can hurt where they shift. We aim for downside control: not much worse than the starting… 18 arXiv — Machine Learning research 3d ago Language Specificity vs. Domain Diversity: Benchmarking Transformers for Bangla Medical NER arXiv:2609.29101v1 Announce Type: new Abstract: Medical Named Entity Recognition (NER) for low-resource languages remains a challenging task due to high linguistic variability and a scarcity of domain-specific annotated corpora. This work presents a comprehensive empirical… 31 arXiv — Machine Learning research 3d ago A Concentration Bound for Two-Timescale Actor-Critic Algorithm arXiv:2609.29117v1 Announce Type: new Abstract: Significant research effort has been directed in recent years towards establishing both asymptotic and non-asymptotic convergence guarantees for two-timescale actor--critic algorithms, where the actor recursion is run on a slower… 13 arXiv — Machine Learning research 3d ago Not Every Token Is Worth Distilling: Selective Supervision for Direct-OPD arXiv:2609.29142v1 Announce Type: new Abstract: Direct On-Policy Distillation (Direct-OPD) transfers reinforcement-learning-induced policy improvements from a small model to a larger student by using the token-level log-ratio between post-RL and pre-RL checkpoints as dense… 13 arXiv — Machine Learning research 3d ago A Particle-Swarm-Assisted Gradient Meta-Learning Algorithm for Joint Transmit Precoding and STAR-RIS Coefficient Optimization arXiv:2609.29150v1 Announce Type: new Abstract: This paper investigates the joint optimization of the transmit precoder and the transmission/reflection coefficients of a simultaneously transmitting and reflecting reconfigurable intelligent surface (STAR-RIS) to maximize the… 38 arXiv — Machine Learning research 3d ago Edge AI on Constrained Devices for Binary Sleep-Wake Classification in Dynamic Environments arXiv:2609.29163v1 Announce Type: new Abstract: This paper presents an Edge AI-based system for detecting sleep and wake states in non-stationary mobile environments using resource-constrained embedded hardware. Conventional approaches relying on accelerometer-based activity… 13 Page 3 of 10 · 500 articles ← Newer Older →