arXiv — Machine Learning
500 articles archived · Visit source ↗ · RSS
-
arXiv — Machine Learning research 3d ago
Stable and Faithful Explanations for Knowledge Tracing
arXiv:2609.28502v1 Announce Type: new Abstract: Knowledge tracing (KT) models predict student performance opaquely, limiting pedagogical action. This study contributes a validation protocol testing predictive competitiveness (RQ1), explanation stability (RQ2) and…
11 -
arXiv — Machine Learning research 3d ago
SMILESGNN: Interpretable Clinical Toxicity Prediction via SMILES-Graph Cross-Attention Fusion
arXiv:2609.28553v1 Announce Type: new Abstract: Drug toxicity prediction is critical for reducing late-stage attrition in drug discovery, yet remains challenging due to severe class imbalance, scaffold-based generalization, and the clinical need for interpretable predictions.…
7 -
arXiv — Machine Learning research 3d ago
CFD Correction of Open Tip Clearance Flow in a Compressor Cascade Using VAE Latent Space Adaptation
arXiv:2609.28558v1 Announce Type: new Abstract: CFD predictions of open tip clearance flow in compressor cascades are subject to discrepancies relative to experiments, while experimental observations are sparse and high-resolution experimental ground truth is unavailable. This…
20 -
arXiv — Machine Learning research 3d ago
CARE: Condition-Aware Representation Regularization for Diffusion Models
arXiv:2609.28561v1 Announce Type: new Abstract: Recent advances in diffusion models highlight the importance of representation regularization for improving sample quality and training efficiency. However, commonly used regularization methods often overlook the built-in…
25 -
arXiv — Machine Learning research 3d ago
SpaFactor: Lightweight Spatial Context-Aware Gene Program Modeling for Histology-to-Transcriptomics Inference
arXiv:2609.28563v1 Announce Type: new Abstract: Spatial transcriptomics (ST) profiles gene expression within tissue architecture, but its cost and experimental complexity limit routine use. Predicting spatial expression from routinely available hematoxylin and eosin (HE) images…
34 -
arXiv — Machine Learning research 3d ago
When Explanations Cannot Be Read: Measuring and Correcting SHAP and LIME Rendering for Right-to-Left Languages
arXiv:2609.28565v1 Announce Type: new Abstract: Post hoc explanation methods such as SHAP and LIME are widely used to interpret text classifiers, but their visualizations are mainly designed for left-to-right languages. When applied to right-to-left (RTL) languages such as Urdu,…
32 -
-
arXiv — Machine Learning research 3d ago
Time-Series Foundation Models That Understand Data Revisions
arXiv:2609.28576v1 Announce Type: new Abstract: Historical observations are not always fixed: statistical agencies revise previously published values as new evidence arrives. Forecasting from a contemporary download can therefore expose a model to information unavailable at the…
13 -
arXiv — Machine Learning research 3d ago
Uncovering Residential PV-EV Co-Adoption from Smart-Meter Data: Load Archetypes and Detection for Demand-Side Planning
arXiv:2609.28578v1 Announce Type: new Abstract: The increasing adoption of electric vehicles (EVs) and rooftop photovoltaic (PV) systems is reshaping residential electricity demand and creating new challenges for demand-side management (DSM), tariff design, and low-voltage…
25 -
arXiv — Machine Learning research 3d ago
Auditability Is Not One Property: Rule Overlap, Behavioural Agreement, and Composition in Reinforcement Learning
arXiv:2609.28581v1 Announce Type: new Abstract: Reinforcement-learning (RL) policies are often distributed as opaque neural checkpoints, while training logs show that a run occurred without explaining what the policy learned. We study whether independently trained policies can…
24 -
arXiv — Machine Learning research 3d ago
SGA: Uncertainty Quantification for Multi-Step Forecasting in Time Series Foundation Models
arXiv:2609.28582v1 Announce Type: new Abstract: The recent emergence of Time Series Foundation Models (TSFMs) has significantly advanced multi-step forecasting performance, enabling accurate predictions over extended future horizons. However, existing TSFMs often suffer from…
5 -
-
arXiv — Machine Learning research 3d ago
Learning to Discover Interesting Mathematics
arXiv:2609.28603v1 Announce Type: new Abstract: Recently, Large Language Models (LLMs) have been increasingly able to solve advanced mathematical problems, including many that have been open for decades. This opens the door to expansion of mathematical knowledge at unprecedented…
4 -
-
arXiv — Machine Learning research 3d ago
UO-FIE: Combining Exact-Label Supervision with Graded Utility for Factivity Inference
arXiv:2609.28605v1 Announce Type: new Abstract: The Factivity Inference Evaluation 2026 (FIE2026) classifies Chinese context-hypothesis pairs into nine ordered factivity intervals. Its evaluation metric rewards both exact predictions and proximity to the correct interval, while…
19 -
arXiv — Machine Learning research 3d ago
fable.intermittent: benchmarking probabilistic forecasting methods for intermittent time series
arXiv:2609.28607v1 Announce Type: new Abstract: Intermittent time series are common in spare-parts demand and retail sales. Since the cost of forecast errors is typically asymmetric, decisions such as inventory control require the full predictive distribution rather than a point…
17 -
arXiv — Machine Learning research 3d ago
RLVR landscapes for iterated multiplications can be benign: Insights from spin-glass theory
arXiv:2609.28625v1 Announce Type: new Abstract: Despite the importance of reinforcement learning with verifiable rewards (RLVR), the extent to which it can learn new reasoning capabilities remains debated. Here we study the optimization landscape of RLVR on algorithmic tasks,…
36 -
arXiv — Machine Learning research 3d ago
OPDiv: Optimal Selection of Top-K High-Scoring, Diverse Compounds
arXiv:2609.28665v1 Announce Type: new Abstract: A virtual screening campaign may produce thousands of promising candidates, but only a small number can be purchased, synthesized, or tested. The practical question is how to select a set of compounds that both rank well and are…
21 -
arXiv — Machine Learning research 3d ago
Beyond Static Graph World Models: Learning Stochastic Latent Dynamics over Evolving Topologies
arXiv:2609.28670v1 Announce Type: new Abstract: Graph-based world models have recently emerged as a means of learning transitions over relational state representations. However, existing approaches are largely limited to fixed-topology graphs or deterministic, fully observable…
22 -
arXiv — Machine Learning research 3d ago
Thinking Leakage: A Causal Audit of NoThink Post-Training in Hybrid Reasoning Models
arXiv:2609.28682v1 Announce Type: new Abstract: Post-training hybrid reasoning models in NoThink mode has attracted growing interest as a way to improve performance while keeping inference fast. However, these gains may draw on thinking behavior already accessible through the…
4 -
arXiv — Machine Learning research 3d ago
Federated Learning of AnDE Classifiers
arXiv:2609.28695v1 Announce Type: new Abstract: This work presents a federated framework for training Averaged $n$-Dependence Estimators (AnDE) in distributed environments. The proposed method focuses on the discriminative setting, where model weights are learned locally and…
16 -
arXiv — Machine Learning research 3d ago
LabFactory: Building and Evaluating Executable AI Labs
arXiv:2609.28697v1 Announce Type: new Abstract: Scientific tasks specify a desired capability, but realizing it often requires building a computational system tailored to the task---acquiring data, designing representations, training models, implementing tools, and deciding how…
27 -
arXiv — Machine Learning research 3d ago
Upholding Robustness in Federated Learning: Trends, Emerging Strategies, and Research Opportunities
arXiv:2609.28722v1 Announce Type: new Abstract: While Federated Learning (FL) has been widely adopted for protecting user privacy in machine learning, it remains vulnerable to various robustness challenges, including performance-impairment risks, information-stealing threats,…
13 -
arXiv — Machine Learning research 3d ago
Policy Complexity, Reaction Time, and Bounded Rationality in Reinforcement Learning
arXiv:2609.28737v1 Announce Type: new Abstract: Biological agents do not learn under conditions of unlimited computation. For humans, learning and choice are shaped by constraints on perception, attention, and working memory, which limit how much state information guides…
22 -
arXiv — Machine Learning research 3d ago
Evaluating Cross-region Generalization for Wavelet-Diffusion Precipitation Downscaling
arXiv:2609.28749v1 Announce Type: new Abstract: Diffusion models have shown strong potential for kilometer-scale precipitation downscaling, but their performance in geographically unseen regions and event regimes remains insufficiently understood. Building on the wavelet…
18 -
arXiv — Machine Learning research 3d ago
The Mechanics of Delta Learning: Target Design for Generalizable Scientific Machine Learning
arXiv:2609.28782v1 Announce Type: new Abstract: In scientific machine learning, $\Delta$-learning trains models on residual errors relative to physical baselines, assuming that more accurate baselines with smaller residual scales inherently improve downstream performance. Here,…
11 -
arXiv — Machine Learning research 3d ago
Vector Bellman Theory for Multichain Robust Average-Reward Markov Decision Processes
arXiv:2609.28792v1 Announce Type: new Abstract: Robust average-reward Markov decision processes provide a fundamental framework for long-term performance optimization under uncertainty, and can have optimal long-run rewards that depend on the initial state. This state dependence…
9 -
-
arXiv — Machine Learning research 3d ago
Stream Recursion Model (SRM)
arXiv:2609.28809v1 Announce Type: new Abstract: Mechanistic interpretability seeks to make verifiable statements about the internal behavior of large language models (LLMs). Many interpretability techniques struggle to scale with the increasing size and depth of architectures.…
7 -
arXiv — Machine Learning research 3d ago
When Does Unsupervised Learning Succeed or Fail? A PoS Perspective on Reconstruction-Based Anomaly Detection
arXiv:2609.28832v1 Announce Type: new Abstract: Reconstruction-based unsupervised learning can fail in two opposing ways: a model may reconstruct anomalies too accurately or discard valid nominal variation. Using the Pursuit of Subspaces hypothesis, we characterize these…
21 -
arXiv — Machine Learning research 3d ago
LastOPD: Taming Collapse in Latent On-Policy Distillation
arXiv:2609.28845v1 Announce Type: new Abstract: On-policy distillation (OPD) corrects a student on the responses it writes, but its signal is the teacher's next-token distribution: it tells the student what the teacher says but misses how it thinks. Latent supervision promises…
28 -
-
arXiv — Machine Learning research 3d ago
Response-state Learning for Transferable Vibrational Spectroscopic Characterization with Electron Prior
arXiv:2609.28935v1 Announce Type: new Abstract: Vibrational spectral prediction can become inaccurate when localized stereoelectronic environments perturb intermediate response states and high-risk response units dominate characteristic spectral fingerprints, making prediction…
21 -
arXiv — Machine Learning research 3d ago
Spectral Graph Neural Networks with Hermite Polynomials: A Comprehensive Study
arXiv:2609.28979v1 Announce Type: new Abstract: We study spectral graph neural networks built from Hermite polynomials and propose HermNet, a simple model that combines a nodewise predictor with normalized Hermite propagation. Its sparse recurrence requires neither…
5 -
arXiv — Machine Learning research 3d ago
Automatic Rank Allocation for Low-Rank Adaptation in Large Language Models via lp Regularization
arXiv:2609.28998v1 Announce Type: new Abstract: Low-rank adaptation (LoRA) has become a popular parameter-efficient fine-tuning method for large language models. A key challenge in LoRA is how to determine the rank of each adaptation matrix, as rank directly controls its…
29 -
arXiv — Machine Learning research 3d ago
Learning from Mixed-Quality Deployment Experience for Robot Manipulation
arXiv:2609.29000v1 Announce Type: new Abstract: Robot policies deployed in real environments naturally accumulate mixed-quality experience, including successful executions, partial progress, and failures. Although these rollouts provide valuable information for further learning,…
21 -
-
arXiv — Machine Learning research 3d ago
Generative Atmospheric Super-Resolution from Heterogeneous In Situ Observations through Composable Interfaces
arXiv:2609.29027v1 Announce Type: new Abstract: Atmospheric observations are sparse, heterogeneous, and unevenly distributed, whereas many generative atmospheric models learn distributions over regularly gridded multivariate states. Once pretrained, diffusion models can supply…
27 -
arXiv — Machine Learning research 3d ago
BranchShine-CR: Compact Multilingual IPA Transcription with Self-Conditioned CTC and Consistency Regularization
arXiv:2609.29069v1 Announce Type: new Abstract: We introduce BranchShine-CR, a 25M-parameter model for multilingual transcription into the International Phonetic Alphabet (IPA). It combines log-mel features, a rotary-position E-Branchformer encoder, intermediate self-conditioned…
36 -
arXiv — Machine Learning research 3d ago
Physics and Data Driven Transformer-Mamba Framework for Flow Field
arXiv:2609.29087v1 Announce Type: new Abstract: While deep learning accelerates expensive partial differential equation solving in computational fluid dynamics (CFD), existing methods like PINNs and FNOs often struggle with generalization, noise robustness, and physical…
24 -
arXiv — Machine Learning research 3d ago
Where Does Exactly-Once Live? Model, Harness, and Tool-Contract Effects on Duplicate Side Effects in LLM Agents
arXiv:2609.29095v1 Announce Type: new Abstract: When a tool-using agent's write times out or returns a server error, the action may already have taken effect. Retrying blindly duplicates it -- a second charge, a second announcement, a second deployment -- while giving up skips…
13 -
arXiv — Machine Learning research 3d ago
Downside-Controlled Online Forecast Combination under Delayed and Revised Outcomes
arXiv:2609.29096v1 Announce Type: new Abstract: Post-hoc correction adjusts a forecaster that cannot be retrained, such as a foundation model, but a correction fitted where errors are stable can hurt where they shift. We aim for downside control: not much worse than the starting…
18 -
arXiv — Machine Learning research 3d ago
Language Specificity vs. Domain Diversity: Benchmarking Transformers for Bangla Medical NER
arXiv:2609.29101v1 Announce Type: new Abstract: Medical Named Entity Recognition (NER) for low-resource languages remains a challenging task due to high linguistic variability and a scarcity of domain-specific annotated corpora. This work presents a comprehensive empirical…
31 -
arXiv — Machine Learning research 3d ago
A Concentration Bound for Two-Timescale Actor-Critic Algorithm
arXiv:2609.29117v1 Announce Type: new Abstract: Significant research effort has been directed in recent years towards establishing both asymptotic and non-asymptotic convergence guarantees for two-timescale actor--critic algorithms, where the actor recursion is run on a slower…
13 -
arXiv — Machine Learning research 3d ago
Not Every Token Is Worth Distilling: Selective Supervision for Direct-OPD
arXiv:2609.29142v1 Announce Type: new Abstract: Direct On-Policy Distillation (Direct-OPD) transfers reinforcement-learning-induced policy improvements from a small model to a larger student by using the token-level log-ratio between post-RL and pre-RL checkpoints as dense…
13 -
-
arXiv — Machine Learning research 3d ago
Edge AI on Constrained Devices for Binary Sleep-Wake Classification in Dynamic Environments
arXiv:2609.29163v1 Announce Type: new Abstract: This paper presents an Edge AI-based system for detecting sleep and wake states in non-stationary mobile environments using resource-constrained embedded hardware. Conventional approaches relying on accelerometer-based activity…
13 -
-
arXiv — Machine Learning research 3d ago
BridgeMem: Causal Dyadic Transition Residuals for Temporal Knowledge Graph Forecasting
arXiv:2609.29268v1 Announce Type: new Abstract: Temporal knowledge graph forecasting aims to infer future relational facts from the temporal structure of observed events. Existing forecasters mainly summarize history through entity states, relation states, paths, or exact…
4 -
arXiv — Machine Learning research 3d ago
Learnable Time-Frequency Masks for Explaining Time-Series Classifiers
arXiv:2609.29270v1 Announce Type: new Abstract: Time-series explainability remains challenging because discriminative information is often encoded in latent frequency or time-frequency features rather than in the raw signal itself. Existing attribution methods typically operate…
12