arXiv — Machine Learning
500 articles archived · Visit source ↗ · RSS
-
arXiv — Machine Learning research 3d ago
Causal State-Space Model for Causal Inference: Estimating Longitudinal Individual Treatment Effects
arXiv:2608.08288v1 Announce Type: new Abstract: Estimating counterfactual outcomes over time from longitudinal observational data is central to clinical decision support. Existing methods rely on domain confusion -- adversarial training that renders representations invariant to…
35 -
arXiv — Machine Learning research 3d ago
A Controlled Study of Feature-Based Knowledge Distillation Across Student Designs
arXiv:2608.08294v1 Announce Type: new Abstract: Knowledge distillation trains a smaller student to match the outputs of a larger teacher. Feature-based methods also align intermediate representations, but this extra constraint may affect students differently. We study this…
23 -
arXiv — Machine Learning research 3d ago
Machine-Learning-Based Diagnostic Framework for Passive Ultrasonic Detection of Railway Wheel Defects
arXiv:2608.08301v1 Announce Type: new Abstract: Reliable identification of railway wheel defects is important for safety and maintenance. This study develops a machine-learning-based diagnostic framework for multi-class defect identification using passive air-coupled ultrasonic…
10 -
-
arXiv — Machine Learning research 3d ago
Spatial Heterogeneity-Aware Multi-Hazard Susceptibility and Risk Mapping at Regional Scale
arXiv:2608.08321v1 Announce Type: new Abstract: Floods and landslides often co-occur, but their relationships with environmental controls vary spatially. This study develops a spatial heterogeneity-aware framework for flood-landslide susceptibility and relative-risk mapping in…
12 -
arXiv — Machine Learning research 3d ago
PRISM: A Predictive Protocol for Permutation Optimization via Landscape Diagnostics
arXiv:2608.08344v1 Announce Type: new Abstract: Permutation optimization arises whenever the components of a system are fixed but their ordering affects performance. We introduce PRISM, a predictive protocol for permutation optimization that measures a fitness landscape before…
8 -
arXiv — Machine Learning research 3d ago
Correlation flow governs learning at criticality
arXiv:2608.08350v1 Announce Type: new Abstract: The initialisation of deep neural networks determines whether information and gradients can propagate across depth, yet a unified theory connecting these properties to learning dynamics remains elusive. Combining mean-field theory…
17 -
arXiv — Machine Learning research 3d ago
Unimodality-Promoting Regularized Learning for Ordinal Regression
arXiv:2608.08359v1 Announce Type: new Abstract: Ordinal regression, also called ordinal classification, is classification of ordinal data, in which the underlying target variable is categorical and considered to have a natural ordinal relation. Previous works have indicated…
37 -
arXiv — Machine Learning research 3d ago
Exact Rank and Convex Calibration Dimension Lower Bounds for the Multi-Label F1 Loss
arXiv:2608.08399v1 Announce Type: new Abstract: The instance-wise $F_1$ measure is a central performance measure for multi-label classification. For a problem with $s$ labels, it defines a $2^s\times 2^s$ loss matrix. Previous work exhibited $s^2+1$-coordinate affine and shifted…
19 -
arXiv — Machine Learning research 3d ago
Rethinking Learning-Based Influence Maximization: Simple Neural Surrogates and Native Discrete Search
arXiv:2608.08406v1 Announce Type: new Abstract: Existing learning-based influence maximization frameworks rely heavily on complex neural architectures and continuous optimization over seed representations. We challenge this paradigm with SIMBA, a diffusion-model-agnostic…
8 -
arXiv — Machine Learning research 3d ago
Constrained Learning with Universally Learnable Concept Classes
arXiv:2608.08414v1 Announce Type: new Abstract: We study constrained statistical learning over infinite-dimensional hypothesis classes in the fully nonconvex setting, and establish universal PACC learnability of the solutions of dual algorithms: Probably Approximately Correct on…
16 -
arXiv — Machine Learning research 3d ago
Optimal Learning Under Tsybakov Noise
arXiv:2608.08416v1 Announce Type: new Abstract: Probably Approximately Correct (PAC) learning [Val84] is a fundamental learning model that has been extensively investigated. In this model, $\mathcal{H} \subseteq \{0,1\}^{\mathcal{X}}$ is a concept class, and $h^*\in\mathcal{H}$…
28 -
arXiv — Machine Learning research 3d ago
FSTC-Encoder: Feature--Spatial--Temporal Correlation Learning for Generalizable RF Sensing
arXiv:2608.08439v1 Announce Type: new Abstract: Heterogeneous RF sensing differs substantially in feature structure, spatial layout, and temporal scale, making existing models difficult to reuse across devices, environments, and RF modalities. We propose FSTC-Encoder, which…
19 -
arXiv — Machine Learning research 3d ago
MGMCL: Multi-Granularity Manifold Contrastive Learning With Neural ODEs for Cross-Subject EEG Emotion Recognition
arXiv:2608.08440v1 Announce Type: new Abstract: Cross-subject electroencephalogram (EEG)-based emotion recognition remains challenging due to substantial inter-individual variability and discrete formulation that overlooks affective continuity. Existing methods operate in…
22 -
arXiv — Machine Learning research 3d ago
No Unique Minimizer, No Problem: On the Consistency of Robust Neural Classifiers
arXiv:2608.08489v1 Announce Type: new Abstract: Neural network classifiers trained by cross-entropy minimization are highly sensitive to label noise and adversarial contamination. While robust alternatives offer bounded influence and resistance to corruption, their statistical…
29 -
arXiv — Machine Learning research 3d ago
Out-of-Distribution Federated Distillation with Domain-Aware Proxy
arXiv:2608.08525v1 Announce Type: new Abstract: Federated Learning is a distributed machine learning paradigm that trains a global model by aggregating local clients without sharing private data of each client. Federated Distillation (FD) builds upon this paradigm by leveraging…
29 -
arXiv — Machine Learning research 3d ago
Task-to-Model Optimization for Enterprise LLM Coding Assistants: A Data-Driven Framework for Cost-Optimal Routing
arXiv:2608.08528v1 Announce Type: new Abstract: Enterprise AI coding assistants incur substantial inference spend, and naive token-cost minimization often fails to reduce end-to-end cost once retries, escalations, and developer wait time are included. We present Task-to-Model…
33 -
arXiv — Machine Learning research 3d ago
Can Graph Learning Learn Circuits?
arXiv:2608.08536v1 Announce Type: new Abstract: Circuit localization is a mechanistic interpretability task whose goal is to identify a sparse subgraph of a transformer's computation graph sufficient to reproduce a particular behavior. Most established methods localize circuits…
26 -
arXiv — Machine Learning research 3d ago
When Skills Meet Safety: Benchmarking and Characterizing the Adaptive Jailbreak Robustness of Skill-Merged LLMs
arXiv:2608.08542v1 Announce Type: new Abstract: Model merging has become the default way to give an aligned language model new skills without retraining: a practitioner folds task vectors from math, code, or domain specialists into a safety-aligned base using task arithmetic,…
12 -
arXiv — Machine Learning research 3d ago
Neural Message Passing on Structural Interaction Graphs for Fully-Inductive Graph Neural Networks
arXiv:2608.08567v1 Announce Type: new Abstract: A central obstacle in building graph foundation models is the input heterogeneity in terms of feature space dimensionality, semantics, and structure. Such heterogeneity limits the capability of graph neural networks to generalize…
11 -
arXiv — Machine Learning research 3d ago
Robust Reputation-Driven Crowdsourced Federated Learning
arXiv:2608.08574v1 Announce Type: new Abstract: Crowdsourced Federated Learning (CrowdFL) extends traditional federated learning by enabling open and heterogeneous participation through a crowdsourcing paradigm. In this setting, reputation-driven incentive mechanisms are…
32 -
-
arXiv — Machine Learning research 3d ago
Multi-Agent Reinforcement Learning via Agent-Specific Preference
arXiv:2608.08604v1 Announce Type: new Abstract: Multi-agent reinforcement learning (MARL) is a powerful framework for solving complex collaborative tasks, but it relies heavily on well-defined global reward functions. Designing such rewards is challenging, especially in systems…
9 -
arXiv — Machine Learning research 3d ago
Domain-Aware Pruning: Sparsity and Domain Generalization via Regularized Probabilistic Masking
arXiv:2608.08624v1 Announce Type: new Abstract: Domain generalization (DG) and neural network pruning are conventionally treated as distinct objectives, targeting out-of-distribution (OOD) robustness and model efficiency, respectively. In this work, we bridge this gap by…
20 -
arXiv — Machine Learning research 3d ago
Trajectory Design and Budgeted Querying for Digital Twin Calibration
arXiv:2608.08631v1 Announce Type: new Abstract: Digital-twin calibration requires interaction data that is expensive to collect. We study two acquisition decisions: which trajectories to generate, and when to spend a limited budget on privileged parameter measurements. Our…
7 -
arXiv — Machine Learning research 3d ago
Exact Rank-Space KL Projection for Shared-Marginal Low-Rank Factors: Application to Doubly Stochastic Clustering
arXiv:2608.08642v1 Announce Type: new Abstract: We study exact Kullback--Leibler (KL) projection for low-rank factorizations whose two nonnegative factors have prescribed row marginals and a shared, learned column marginal. For arbitrary positive row marginals of equal total…
24 -
arXiv — Machine Learning research 3d ago
Path-dependent Discrete Amortized Inference
arXiv:2608.08644v1 Announce Type: new Abstract: We consider the problem of sampling compositional and discrete objects from a given unnormalized posterior distribution. Notably, recent studies have shown that this problem can be efficiently solved by learning a deterministic…
16 -
arXiv — Machine Learning research 3d ago
Multi-Relational Knowledge Graph Enhanced Embedding for Trajectory-User Linking
arXiv:2608.08646v1 Announce Type: new Abstract: Trajectory-User Linking (TUL) aims to identify the owner of an anonymous trajectory from a set of candidate users, providing a basis for user mobility analysis and personalized location-aware services. Existing methods often learn…
6 -
arXiv — Machine Learning research 3d ago
LegoLM: Structured Weight Sharing for Large Language Models
arXiv:2608.08652v1 Announce Type: new Abstract: We present \LegoLM{}, a structured weight-sharing compression framework for large language models grounded in a systematic study of why global weight sharing fails and how to fix it. We identify two distinct failure modes.…
12 -
arXiv — Machine Learning research 3d ago
Catastrophic Forgetting in Continual Reinforcement Learning
arXiv:2608.08673v1 Announce Type: new Abstract: This work explores the relationship between task similarity and catastrophic forgetting in reinforcement learning. Catastrophic forgetting, the phenomenon in machine learning of losing the ability to effectively perform on previous…
21 -
arXiv — Machine Learning research 3d ago
Backward Compatibility in Tree-Based Explanations and Enhanced CART Algorithm
arXiv:2608.08674v1 Announce Type: new Abstract: In the operation of machine learning models, model update is a fundamental process that requires careful consideration of its impact on downstream decision-making. Particularly when operating explainable models, changes in…
27 -
arXiv — Machine Learning research 3d ago
Efficient Test-Time Scaling for LLM-based Time Series Forecasting
arXiv:2608.08675v1 Announce Type: new Abstract: Long-term time series forecasting benefits from preserving global structure such as trends and seasonality. Recent LLM-based forecasters often improve accuracy through test-time scaling (e.g., iterative refinement), but these…
27 -
arXiv — Machine Learning research 3d ago
RippleKV: Cross-Layer KV Cache Allocation via Perturbation Propagation
arXiv:2608.08684v1 Announce Type: new Abstract: Long-context LLM inference is bottlenecked by KV cache memory, yet distributing a limited cache budget across layers remains challenging. Existing methods rely on proxies such as layer depth, attention statistics, or representation…
34 -
arXiv — Machine Learning research 3d ago
Loss-Resilient Wireless Video Token Communication over Block Fading Channels
arXiv:2608.08698v1 Announce Type: new Abstract: Video token communication represents video content as discrete tokens that differ in their importance to reconstruction and exhibit temporal dependencies. When these tokens are packetized for wireless transmission, block fading can…
16 -
arXiv — Machine Learning research 3d ago
Gaming Without an Attacker: Benchmark Fingerprinting in LLM-Driven Search Under Selection Pressure
arXiv:2608.08722v1 Announce Type: new Abstract: Benchmarks for systems that are optimized against the evaluation signal measure something different from what they claim. We document this concretely in two GPU-kernel-optimization suites with held-out generalization gates:…
24 -
arXiv — Machine Learning research 3d ago
PAST: Privileged Adaptation from Complete Student Trajectories for On-Policy Self-Distillation
arXiv:2608.08726v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) uses a privileged teacher to supervise a reasoning model on prefixes sampled from its own rollouts. Yet each rollout also reveals how the student's response unfolds and whether it succeeds,…
12 -
arXiv — Machine Learning research 3d ago
Measuring and Reducing WebGPU Dispatch Overhead for LLM Inference
arXiv:2608.08730v1 Announce Type: new Abstract: Large Language Models are deployed to multiple types of environments, from internet browsers to edge devices, and WebGPU serves as a modern cross-platform standard. The engines for browser-based LLM inference have proliferated, yet…
19 -
-
-
arXiv — Machine Learning research 3d ago
Quantum-Classical Physics-Informed Kolmogorov-Arnold Networks for Solving Fuzzy Differential Equations
arXiv:2608.08782v1 Announce Type: new Abstract: In this study, we propose a quantum-classical physics-informed Kolmogorov-Arnold network (QCPIKAN) dedicated to the solution of fuzzy differential equations. The network takes the spatiotemporal coordinates and membership level as…
28 -
arXiv — Machine Learning research 3d ago
Distilling Vision-Language Models for Robust Traffic Sign Perception in Autonomous Vehicles
arXiv:2608.08815v1 Announce Type: new Abstract: Traffic sign recognition (TSR) models based on deep neural networks achieve strong clean-data performance but remain vulnerable to physically realizable adversarial attacks, including shadow perturbations, natural-light…
15 -
-
arXiv — Machine Learning research 3d ago
The Cost of Adaptivity: Matching Lower Bounds Across Learning Problems
arXiv:2608.08826v1 Announce Type: new Abstract: Adaptive procedures must work without nuisance information an oracle may use, such as a gradient scale or smoothness index, and robust procedures may have to answer queries whose coordinate and inspection time are chosen only after…
7 -
arXiv — Machine Learning research 3d ago
Beyond Routing: Decoupling Expert Dispatch and Aggregation in Sparse Mixture-of-Experts
arXiv:2608.08853v1 Announce Type: new Abstract: Sparse Mixture-of-Experts (MoE) routers commonly use the same scores both to select experts and to weight their already-computed outputs. We study whether these two roles, dispatch and aggregation, should be coupled. On pretrained…
11 -
arXiv — Machine Learning research 3d ago
Agentic Anomaly Detection with ORCA-Style Dynamic Inductive Bias Adaptation in Multimodal Wearable Time Series Data
arXiv:2608.08859v1 Announce Type: new Abstract: Wireless Body Area Networks (WBANs) generate multivariate physiological time series that are highly nonstationary and must often be processed under strict computational and memory constraints. A critical yet underexplored challenge…
19 -
arXiv — Machine Learning research 3d ago
Approximation Rates for Metaplectic Neural Networks
arXiv:2608.08872v1 Announce Type: new Abstract: In this paper we develop quantitative approximation results for shallow neural networks constructed using a dictionary based on metaplectic operators. First, we extend the concept of Barron spaces by considering a symplectically…
29 -
arXiv — Machine Learning research 3d ago
DistillCache: KL-Guided Adaptive KV-Cache Eviction for Memory-Efficient LLM Inference
arXiv:2608.08878v1 Announce Type: new Abstract: Transformer-based large language models (LLMs) achieve strong performance across many tasks, but their Key-Value (KV) cache grows linearly with sequence length, creating a severe memory bottleneck for long-context inference.…
32 -
arXiv — Machine Learning research 3d ago
Federated Attention Autoencoders with a Stochastic Aggregation Scheme for Anomaly Detection
arXiv:2608.08906v1 Announce Type: new Abstract: Outlier detection in decentralized data environments is a challenging task for many machine learning implementations, particularly in settings where data cannot be shared. Recently, there have been advances in federated outlier…
25 -
arXiv — Machine Learning research 4d ago
Latent Fact-Checking: Detecting Misinformation through Activation Engineering
arXiv:2608.06417v1 Announce Type: new Abstract: The proliferation of misinformation online has driven demand for scalable detection systems. While most existing approaches rely on surface-level linguistic features or external knowledge retrieval, we examine truthfulness as a…
36 -
arXiv — Machine Learning research 4d ago
Risk-Aware Decision Policies for Agents Under Noisy Perception
arXiv:2608.06420v1 Announce Type: new Abstract: Perception in biological systems is inherently noisy, requiring organisms to make decisions under uncertainty where misclassification can be costly or fatal. We present an Artificial Life predator-prey model of foraging under noisy…
23