arXiv — Machine Learning
500 articles archived · Visit source ↗ · RSS
-
arXiv — Machine Learning research 4d ago
The Drift Contract: Spectral Updates for Depth-Robust Local Learning
arXiv:2609.26811v1 Announce Type: new Abstract: Local learning trains each layer with its own auxiliary loss and no global backward pass, which makes layer updates structurally parallel. Two problems have kept it marginal: accuracy degrades as depth grows, and hyperparameters…
16 -
arXiv — Machine Learning research 4d ago
Signal2Symbol: Neuro-Symbolic Temporal Reasoning for Explainable Physiological Time-Series Anomaly Detection
arXiv:2609.26820v1 Announce Type: new Abstract: Physiological time series such as electrocardiograms (ECG) and electroencephalograms (EEG) exhibit complex temporal structure, substantial acquisition variability, and a strong need for transparent decision-making. Although deep…
4 -
arXiv — Machine Learning research 4d ago
HARN: Hierarchical Associative Resonance Network for Event-Driven Multi-Timeframe Forecasting
arXiv:2609.26822v1 Announce Type: new Abstract: Financial time series evolve across multiple temporal resolutions, challenging forecasting systems to incorporate newly available information without repeatedly recomputing unchanged representations. We introduce HARN, a…
15 -
arXiv — Machine Learning research 4d ago
What Makes a Terminal-Bench Task Hard? Separating Genuine Hardness from Fake-Hardness on an Adjudicated Agentic Corpus
arXiv:2609.26826v1 Announce Type: new Abstract: Frontier benchmarks need tasks that current models cannot solve. But a task that no model solves is not automatically a hard task. The same zero pass rate can come from a real capability gap, but it can also come from missing…
13 -
arXiv — Machine Learning research 4d ago
LWCal: Loss-Weighted Calibration for Tabular Classifiers with Noisy Calibration Labels
arXiv:2609.26839v1 Announce Type: new Abstract: Post-hoc probability calibration is usually evaluated under an optimistic assumption: the held-out calibration labels are clean. In many AI deployment settings, however, labels come from weak annotators, historical decisions,…
6 -
arXiv — Machine Learning research 4d ago
A Leakage-Aware Multimodal Evaluation Framework for Early Intraoperative Acute Kidney Injury Prediction
arXiv:2609.26848v1 Announce Type: new Abstract: Postoperative acute kidney injury (AKI) after major non-cardiac surgery carries substantial morbidity, yet early intraoperative risk stratification remains difficult. In this retrospective cohort study, we propose SynerT, a…
35 -
arXiv — Machine Learning research 4d ago
COPE: Continual Personalization of LLMs under Sparse User Feedback via User Embeddings and Self-Evaluation
arXiv:2609.26853v1 Announce Type: new Abstract: While Large Language Models (LLMs) have achieved remarkable results across various benchmarks, their alignment with normative values often results in homogenized responses that fail to address diverse user preferences. Existing…
5 -
arXiv — Machine Learning research 4d ago
QUARTET: Quad-branch cross-Attention and Random-walk Traces for Enhancing Transformers on Relational Graphs
arXiv:2609.26855v1 Announce Type: new Abstract: Relational Deep Learning (RDL) models multi-table databases as heterogeneous temporal graphs, and graph transformers currently achieve state-of-the-art performance on benchmarks like RelBench. However, the current leading model,…
20 -
arXiv — Machine Learning research 4d ago
Marginally Correct Tool Caches Can Reverse Group-Normalized Policy Updates
arXiv:2609.26866v1 Announce Type: new Abstract: Tool-result caching reduces repeated execution in agent training, but also couples rollout randomness. We study a two-action model in which independent and shared execution preserve every rollout's conditional reward distribution.…
24 -
arXiv — Machine Learning research 4d ago
PR-Smoother: Simulator-Preserving Non-Gaussian Smoothing for Data Assimilation
arXiv:2609.26890v1 Announce Type: new Abstract: Many physical data assimilation (DA) workflows require smoothing methods that represent non-Gaussian posteriors over physical state variables, scale to high-dimensional simulators, train from observation windows alone, and remain…
36 -
arXiv — Machine Learning research 4d ago
CORE-STACK+: Meta-Learning for Deep Stacked Generalization
arXiv:2609.26905v1 Announce Type: new Abstract: Stacking heterogeneous vision backbones (CNNs, ViTs, and hybrids) is the de facto recipe for accuracy, calibration, and robustness, yet two coupled pathologies limit its returns. Prediction-space multicollinearity ill-conditions…
9 -
arXiv — Machine Learning research 4d ago
On Preference Coverage Collapse from Hindsight Relabeling in Multi-Objective Reinforcement Learning
arXiv:2609.26918v1 Announce Type: new Abstract: Hindsight relabeling which retroactively replacing a transition's goal with the outcome the agent actually achieved is an effective tool for improving sample-efficiency in Reinforcement Learning (RL). A natural extension to…
23 -
-
-
arXiv — Machine Learning research 4d ago
CRISP: Scalable Importance-Stratified Coresets for Imbalanced Tabular Learning
arXiv:2609.26962v1 Announce Type: new Abstract: Large imbalanced tabular datasets make repeated gradient-boosted tree training expensive. Existing coreset methods often lose accuracy when most majority examples are removed. We present CRISP (Coreset Reduction via…
6 -
arXiv — Machine Learning research 4d ago
TinyUDE: Solver-Free Universal Differential Equations on Microcontrollers via Lie-Taylor Jet Matching
arXiv:2609.26972v1 Announce Type: new Abstract: Training Universal Differential Equations (UDEs) traditionally relies on backpropagating through numerical ODE solvers, creating memory footprints far exceeding the capabilities of edge microcontrollers. We present Lie-Taylor jet…
20 -
arXiv — Machine Learning research 4d ago
Resource-Efficient Distributed Recursive Gaussian Processes
arXiv:2609.26979v1 Announce Type: new Abstract: Gaussian processes (GPs) provide a flexible framework for learning unknown functions from noisy measurements while quantifying predictive uncertainty, making them well suited for estimation in multi-agent systems. However, when…
21 -
arXiv — Machine Learning research 4d ago
GeoRVQ: Decoder-aware geometry for residual-token prediction in physiological signals
arXiv:2609.27018v1 Announce Type: new Abstract: Residual vector quantization (RVQ) turns physiological waveforms into compact token sequences, but conventional masked modeling treats every incorrect token as equally costly. We propose GeoRVQ, a coarse-to-fine masked token model…
19 -
arXiv — Machine Learning research 4d ago
WTF?! Simulation-Free Reinforcement Learning with Wasserstein-Tilted Flow Maps
arXiv:2609.27033v1 Announce Type: new Abstract: Reward fine-tuning aims to update a pre-trained flow-based generative model to improve the downstream reward of its generated samples. Existing methods typically formulate this problem as sampling from a reward-tilted distribution,…
27 -
arXiv — Machine Learning research 4d ago
An open benchmark for machine learning-based polymer property prediction
arXiv:2609.27036v1 Announce Type: new Abstract: Polymer property prediction lacks open, standardized benchmarks that enable rigorous comparison of machine-learning methods, with existing resources covering only a narrow fraction of polymer architectures, such as homopolymers. We…
16 -
arXiv — Machine Learning research 4d ago
ChipMEM: Verification-Grounded Memory for EDA Agents
arXiv:2609.27067v1 Announce Type: new Abstract: Large language model (LLM)-based agents use Electronic Design Automation (EDA) tools to generate and revise register-transfer-level (RTL) designs under synthesis and verification feedback. Recent methods learn from this feedback by…
33 -
arXiv — Machine Learning research 4d ago
Local Evidence and Geometric Readout Repair in Trained GNNs
arXiv:2609.27092v1 Announce Type: new Abstract: Many node-classification GNNs apply a linear classifier to a nonnegative mixture of local messages. An error can reflect either poor mixture weights or a reachable logit set poorly positioned for the classifier. We separate these…
18 -
arXiv — Machine Learning research 4d ago
Learning Risk Scores Robust to Unobserved Confounders
arXiv:2609.27144v1 Announce Type: new Abstract: We consider the problem of learning risk scores to prioritize individuals for scarce resources or interventions, from historical observational data affected by unobserved confounding. Decisions about who receives scarce resources…
16 -
arXiv — Machine Learning research 4d ago
The Linear Representation Hypothesis Needs a Group Action
arXiv:2609.27158v1 Announce Type: new Abstract: To make claims about representations that generalize beyond a particular trained model, we need to specify when two representations should count as equivalent. The Linear Representation Hypothesis is often discussed without making…
13 -
arXiv — Machine Learning research 4d ago
Scaling of Capability and Efficiency at Inference Time in Large Reasoning Models
arXiv:2609.27166v1 Announce Type: new Abstract: Capability and efficiency are two key dimensions of reasoning in large language models (LLMs). Capability refers to the ability to solve a given problem correctly, whereas efficiency refers to the ability to do so with limited…
29 -
arXiv — Machine Learning research 4d ago
Data-driven discrete-time deep recurrent neural network-based modeling for dissipative systems
arXiv:2609.27186v1 Announce Type: new Abstract: Physical AI has gained increasing attention for its role in developing AI systems that better understand, predict, and control real-world dynamics. Achieving this requires AI models that not only achieve high prediction accuracy…
22 -
arXiv — Machine Learning research 4d ago
ZO-COSMO: Index-Free One-Hop Mixing for Decentralized Zeroth-Order Optimization
arXiv:2609.27199v1 Announce Type: new Abstract: Sparse communication in decentralized zeroth-order learning requires compatible peer-state coordinates. We characterize this one-hop condition and develop \textsf{ZO-COSMO}, coupling two-query estimation with average-preserving…
11 -
arXiv — Machine Learning research 4d ago
A Systematic Benchmark of Explainable Methods for Temporal Attribution in Sequential Recommendation Systems
arXiv:2609.27201v1 Announce Type: new Abstract: Sequential RecSys are central to modern personalization, exploiting user's historical interaction sequences to drive next-step decisions. Deep learning models, particularly CNN and Transformer-based architectures, have proven…
32 -
arXiv — Machine Learning research 4d ago
Scalable Subgraph Sampling via Resistance Curvature
arXiv:2609.27209v1 Announce Type: new Abstract: Subgraph sampling reduces the training cost of large-scale graph neural networks, but sampling criteria may overlook the geometric roles of edges. We propose a resistance-curvature-guided sampling framework built on ERC-LG, a…
29 -
arXiv — Machine Learning research 4d ago
Tail-Aware Geometry Learning for Conformal Ellipsoids
arXiv:2609.27221v1 Announce Type: new Abstract: This paper studies multivariate conformal prediction (CP), a distribution-free uncertainty quantification framework with finite-sample coverage guarantees. The efficiency of multivariate prediction sets hinges critically on the…
37 -
arXiv — Machine Learning research 4d ago
A Scaling Study for fMRI Foundation Models
arXiv:2609.27232v1 Announce Type: new Abstract: Scaling laws have guided large-model development in computer vision and natural language processing, but the relationships among data, model size, and compute remain unclear for functional magnetic resonance imaging (fMRI)…
26 -
-
arXiv — Machine Learning research 4d ago
Full-Covariance Smoothing of Bayesian Neural Networks for Online Adaptation
arXiv:2609.27244v1 Announce Type: new Abstract: A neural network's layers can be treated as time steps of a state-space model, turning Bayesian training into a smoothing problem: a forward pass propagates Gaussian moments through the network, and a backward Rauch--Tung--Striebel…
35 -
arXiv — Machine Learning research 4d ago
Repurposing Pre-trained LLMs as High Fidelity Continuous Text Autoencoders
arXiv:2609.27248v1 Announce Type: new Abstract: Next-token prediction has enabled highly fluent autoregressive language models, but it represents global structure only indirectly through sequential factorization. In contrast, high-fidelity autoencoders have become a standard…
22 -
arXiv — Machine Learning research 4d ago
What Converges in the Platonic Representation Hypothesis? Structure over Geometry
arXiv:2609.27252v1 Announce Type: new Abstract: The Platonic Representation Hypothesis suggests that increasingly capable models converge toward shared representations. Recent work narrows this claim to shared local neighborhood relationships, finding that capacity-dependent…
15 -
arXiv — Machine Learning research 4d ago
Graph Learning with Spectral Connectivity Priors for Scarce Data
arXiv:2609.27278v1 Announce Type: new Abstract: Learning a sparse graph from scarce data is practically important but challenging. Motivated by the desirable combination of local sparsity and strong global connectivity exhibited by expander-like graphs, we propose spectral…
5 -
arXiv — Machine Learning research 4d ago
SR-Fraud: An Outcome-Supervised Reflective LLM Agent Framework for Non-Stationary Payment Fraud Detection
arXiv:2609.27287v1 Announce Type: new Abstract: Real-time payment fraud detection is a non-stationary streaming prediction problem: adversaries adapt before supervised labels mature, and localized burst attacks can cause losses before retraining. Production systems typically…
31 -
arXiv — Machine Learning research 4d ago
NGN: Learning Neural Network Size as a Differentiable Count
arXiv:2609.27291v1 Announce Type: new Abstract: Neural network size is usually chosen before training, separating architecture selection from weight optimization. We introduce the Neurogenesis Network (NGN), a differentiable parameterization for learning how many ordered…
30 -
arXiv — Machine Learning research 4d ago
KITE: KV-Invariant Transformer Expansion for Efficient Agentic LLM Scaling
arXiv:2609.27294v1 Announce Type: new Abstract: Scaling a language model is not only a question of final quality: the architectural choice determines how much computation is spent during training, prompt processing, and autoregressive decoding to achieve certain model quality.…
14 -
arXiv — Machine Learning research 4d ago
Live Assistant: Learning Whether, When, and Whom to Assist in Real-World Live Social Streams
arXiv:2609.27303v1 Announce Type: new Abstract: Livestreams are long-lasting interactive environments where audiovisual content, viewer activity, host behavior, and platform signals evolve together, creating assistance needs that emerge from the stream itself. We introduce…
29 -
arXiv — Machine Learning research 4d ago
Discrete Diffusion Models via Evolving Variational Autoregressive Networks
arXiv:2609.27306v1 Announce Type: new Abstract: Conventional score-based diffusion models learn scores without representing normalized densities, whereas tractable normalized models support both sampling and direct likelihood evaluation. A recent tensor-network approach provides…
5 -
arXiv — Machine Learning research 4d ago
Quantization-Robust Unlearning through the Lens of Retain-Forget Loss Landscapes Interaction
arXiv:2609.27355v1 Announce Type: new Abstract: Unlearning ensures LLM compliance by removing the influence of private or copyrighted training data. However, since LLM models typically undergo post-training compression, like quantization, in practical deployment, it has been…
11 -
arXiv — Machine Learning research 4d ago
Anomaly-Free Self-Optimization via AUC Bounds
arXiv:2609.27362v1 Announce Type: new Abstract: Anomalies are rare, and anomalous data are often unavailable during development, making it difficult to determine which anomaly detection models and configurations will generalize to unseen anomalies. Recent approaches address this…
15 -
arXiv — Machine Learning research 4d ago
Forecast Workflow Bench: Evaluating Language-Model Decisions with Budgeted Forecast Tools
arXiv:2609.27385v1 Announce Type: new Abstract: Time-series foundation models (TSFMs) provide forecasts for operational decisions, but accuracy alone does not determine their value. Evaluating agents that use these models requires measuring decision quality and forecast cost.…
27 -
arXiv — Machine Learning research 4d ago
Active Learning for Biodiversity Monitoring: From Label Efficiency to Reliable Ecological Inference
arXiv:2609.27409v1 Announce Type: new Abstract: Limited expert annotation capacity is a pervasive constraint in biodiversity monitoring. Passive acoustic recorders and camera traps generate data faster than experts can analyse them. Machine learning (ML) models can process these…
24 -
arXiv — Machine Learning research 4d ago
When Labels Are Scarce: An Oscillatory State Space Model for Vibration Diagnosis
arXiv:2609.27411v1 Announce Type: new Abstract: Machine fault diagnosis from vibration requires learning from scarce labelled fault recordings while meeting the computational constraints of edge devices for local inference. We introduce DualRes, a compact oscillatory state-space…
30 -
arXiv — Machine Learning research 4d ago
Counterfactual Constraint-Conditioned On-Policy Distillation for Multi-Constraint Instruction Following
arXiv:2609.27421v1 Announce Type: new Abstract: Multi-constraint instruction following requires a model to respond to a query under many simultaneously active constraints. Even strong instruction-tuned models still routinely violate some of them. Existing approaches either…
14 -
arXiv — Machine Learning research 4d ago
Stable Neural Decoding Across Sessions via Task-Conditioned Latent Alignment for Brain-Machine Interfaces
arXiv:2609.27441v1 Announce Type: new Abstract: Achieving stable long-term neural decoding in invasive brain-machine interfaces (BMIs) remains challenging due to variations in recorded neural populations across sessions. Current latent alignment approaches may overlook…
8 -
arXiv — Machine Learning research 4d ago
Quantum Reinforcement Learning for Cost and Delay Tradeoffs in Quantum Cloud Orchestration
arXiv:2609.27446v1 Announce Type: new Abstract: Quantum cloud computing, delivered through the quantum-as-a-service (QaaS) model, provides access to quantum computing resources. However, applying uniform time-based pricing across fundamentally heterogeneous quantum resources…
30 -