arXiv — Machine Learning
500 articles archived · Visit source ↗ · RSS
-
arXiv — Machine Learning research 2d ago
Optimize Cheap, Deploy Strong: Cost-Aware Cross-Tier Transfer for Evolutionary Optimization
arXiv:2608.10694v1 Announce Type: new Abstract: Evolutionary optimization of LLM prompts and agentic programs (e.g., GEPA) is dominated by fitness evaluation: scoring each candidate runs an answering LLM over a validation set, so the evaluator's price tier dictates total search…
25 -
arXiv — Machine Learning research 2d ago
ProTAGAD: A Foundation Model for TAG Anomaly Detection with Decoupled Topological and Textual Prototypes
arXiv:2608.10699v1 Announce Type: new Abstract: Text-Attributed Graphs (TAGs), endowed with abundant textual content along with topological structures, have emerged as a versatile backbone for real-world anomaly detection spanning large language model security, social network…
24 -
arXiv — Machine Learning research 2d ago
Your LLM, Your Style: Behavioral Mode Axes for LLM Behavioral Control
arXiv:2608.10703v1 Announce Type: new Abstract: Large language models (LLMs) increasingly act in interactive settings where their behavioral styles affect user experience, safety, and downstream decision making. Existing LLM personality studies largely rely on self-report…
18 -
arXiv — Machine Learning research 2d ago
SQuaT: Self-Supervised Knowledge Distillation via Student-Aware Quantized Teacher Features
arXiv:2608.10709v1 Announce Type: new Abstract: Quantization-Aware Training (QAT) enables the deployment of quantized models with minimal accuracy degradation. However, in practical scenarios, training labels are often unavailable due to privacy, copyright, or cost constraints.…
5 -
arXiv — Machine Learning research 2d ago
Long-Time Trajectory Approximation via SA-NODEs: Model Predictive and Floquet Strategies
arXiv:2608.10738v1 Announce Type: new Abstract: We study the approximation of dynamical systems by semi-autonomous neural ordinary differential equations (SA-NODEs) over long time horizons. For a single network trained on the whole horizon, the available error bound deteriorates…
27 -
arXiv — Machine Learning research 2d ago
Path Integral Value Matching for Linear Quadratic Stochastic Optimal Control
arXiv:2608.10777v1 Announce Type: new Abstract: Linear Quadratic Stochastic Optimal Control (LQ-SOC) establishes a fundamental framework for steering noisy dynamical systems and has recently gained renewed interest in the machine learning community. However, current…
26 -
arXiv — Machine Learning research 2d ago
MoE Proxy Models for Low-Cost Failure Reproduction and Diagnosis in LLM RL Post-Training
arXiv:2608.10823v1 Announce Type: new Abstract: Reinforcement learning (RL) post-training of large language models (LLMs) is computationally intensive and involves complex system pipelines with substantial debugging overhead. In practice, factors such as framework adaptation,…
16 -
arXiv — Machine Learning research 2d ago
TACTICL: Task-Aware Compression of Tabular ICL Models
arXiv:2608.10837v1 Announce Type: new Abstract: The strong performance of foundation models for tabular tasks comes at substantial inference costs. Distilling models into task-specific architectures reduces model size and computational demands but also sacrifices in-context…
5 -
arXiv — Machine Learning research 2d ago
Diffract: Spectral View of LLM Domain Adaptation
arXiv:2608.10850v1 Announce Type: new Abstract: We study continual pre-training (CPT) as a mechanism for adapting general-purpose large language models to specialized domains: mathematics, instruction, code, and natural text. Using singular value decomposition of weight…
6 -
arXiv — Machine Learning research 2d ago
FiGuRO: Intrinsic Dimension Estimation for Multi-Modal Data
arXiv:2608.10857v1 Announce Type: new Abstract: Determining the complexity, or Intrinsic Dimension (ID), of data is fundamental to efficient and interpretable representation learning. This is particularly challenging in multi-modal settings when trying to learn disentangled…
10 -
arXiv — Machine Learning research 2d ago
Can Bayesian Optimization Efficiently Find a Strong Single Expert in Neural Thickets?
arXiv:2608.10867v1 Announce Type: new Abstract: Gradient-free post-training has emerged as a compelling alternative to gradient-based optimization for large language models (LLMs), but existing approaches remain costly. We ask whether structured search can identify a strong…
28 -
arXiv — Machine Learning research 2d ago
Optimistic Rates for Multiclass PAC Learning
arXiv:2608.10869v1 Announce Type: new Abstract: Worst-case multiclass bounds do not become smaller when the best classifier is already nearly correct: what is missing is an optimistic rate, a guarantee whose fluctuation scales with the oracle risk itself. For a class of…
5 -
arXiv — Machine Learning research 2d ago
Benchmarking Time Series Generation Methods for Privacy-Preserving Forecasting
arXiv:2608.10891v1 Announce Type: new Abstract: Time series forecasting in privacy-sensitive domains often requires training models on released data rather than original observations. Synthetic time series generation has been developed primarily for data augmentation, where…
6 -
arXiv — Machine Learning research 2d ago
Partially Observable Learning for Multi-Platform Dispatch Optimization
arXiv:2608.10897v1 Announce Type: new Abstract: Instant delivery platforms have become a critical component of urban logistics, increasingly relying on crowdsourced couriers to fulfill highly dynamic orders. In real-world systems, couriers are not exclusive to a single platform…
33 -
arXiv — Machine Learning research 2d ago
ReOrder-OPD:Reliability-Aware Prompt Ordering for On-Policy Distillation
arXiv:2608.10905v1 Announce Type: new Abstract: On-policy distillation (OPD) applies token-level teacher supervision to student-generated trajectories, but this supervision is not always reliable. Existing methods use local confidence or teacher-student agreement to weight,…
7 -
arXiv — Machine Learning research 2d ago
Physics-informed Diffusion Generative Model for Time-Series Data Synthesis in Dynamic Systems
arXiv:2608.10941v1 Announce Type: new Abstract: Industrial time-series signals, such as turbine temperature and rotational speed in aero-engines, are essential for monitoring the health and operational status of complex dynamical systems. However, collecting such data is often…
23 -
arXiv — Machine Learning research 2d ago
GARLIC: Graph Attention-based Relational Learning of Multivariate Time Series in Intensive Care
arXiv:2608.10969v1 Announce Type: new Abstract: Healthcare data, such as Intensive Care Unit (ICU) records, comprise heterogeneous multivariate time series sampled at irregular intervals with pervasive missingness. However, clinical applications demand predictive models that are…
27 -
-
arXiv — Machine Learning research 2d ago
Derivative Computation in PINNs: Automatic Differentiation, Finite Differences and Beyond
arXiv:2608.11020v1 Announce Type: new Abstract: We systematically investigate finite-difference (FD) derivative computation in Physics-Informed Neural Networks (PINNs) as an alternative to automatic differentiation (AD). On three benchmark PDEs we show that, with a properly…
31 -
arXiv — Machine Learning research 2d ago
Mapping and Measuring the Behavioral Evolution of Large Language Models
arXiv:2608.11027v1 Announce Type: new Abstract: Benchmark leaderboards summarize how well a language model performs, but not how its behavior relates to that of other models or changes across generations. We characterize the output behavior of 32 models from six families using…
27 -
arXiv — Machine Learning research 2d ago
ReRound: Reconstructive Rounding to Resolve Midpoint Ambiguity in Calibration-Free LLM Quantization
arXiv:2608.11045v1 Announce Type: new Abstract: ReRound (Reconstructive Rounding) is a post-training quantization method that addresses the midpoint ambiguity inherent in standard round-to-nearest (RTN) schemes when quantizing weights near the centers of quantization intervals.…
30 -
arXiv — Machine Learning research 2d ago
Efficient Hypergradient Descent for Inverse Reinforcement Learning
arXiv:2608.11052v1 Announce Type: new Abstract: Inverse reinforcement learning (IRL) aims to recover a reward function under which the resulting policy reproduces the behavior observed in expert demonstrations. A natural approach is to formulate IRL as a bilevel optimization…
38 -
arXiv — Machine Learning research 2d ago
Uncertainty-Aware Deep Learning for Genomics Applications: Insights from an Empirical Study
arXiv:2608.11054v1 Announce Type: new Abstract: Deep learning models have emerged as the standard computational tool for a wide range of applications in genomics. Yet, uncertainty quantification (UQ) -- and more specifically, the reliability of different uncertainty estimates in…
25 -
arXiv — Machine Learning research 2d ago
Batch Size or Negatives? A Selection Rule for Memory-Constrained Recommender Training
arXiv:2608.11061v1 Announce Type: new Abstract: Large-scale neural recommender systems are typically trained with a softmax cross-entropy objective over the full item vocabulary. For a typical large number of possible items $K$, the final classification layer dominates memory,…
38 -
arXiv — Machine Learning research 2d ago
Cross-View Feature Matching: Survey, Benchmarking, and Foundation-Model Perspectives
arXiv:2608.11093v1 Announce Type: new Abstract: Cross-view feature matching aims to establish reliable correspondences across images with large viewpoint variations. Over the past decade, the field has evolved from task-specific models toward increasingly unified and…
35 -
arXiv — Machine Learning research 2d ago
Two-stage Odd Residual Flows for Mean-Preserving Probabilistic Time Series Forecasting
arXiv:2608.11114v1 Announce Type: new Abstract: Probabilistic forecasting plays an essential role in risk-sensitive decision-making, particularly in long-horizon settings. However, existing approaches often face a fundamental trade-off between distributional flexibility and…
31 -
arXiv — Machine Learning research 2d ago
A Recommendation System Approach for Interference-Robust Sensor Subset Selection
arXiv:2608.11143v1 Announce Type: new Abstract: This paper develops a method for sensor-subset selection for tracking. Prior work showed that low-cost acoustic Received Signal Strength Indicator (RSSI) measurements can be used to recommend subsets of sensor nodes whose expensive…
38 -
arXiv — Machine Learning research 2d ago
DACRI: Decision-Aware Causal Intervention Ranking for Critical Supply Chains
arXiv:2608.11154v1 Announce Type: new Abstract: Detecting or attributing a supply-chain disruption is not the same as selecting the intervention that maximizes recoverable net value. We present CriticalSCM-Bench v1, a controlled synthetic benchmark with causal ground truth,…
35 -
arXiv — Machine Learning research 2d ago
Hierarchical Empirical-Bayes Naive Bayes: Minimax Smoothing and Calibration with AODE Extension
arXiv:2608.11162v1 Announce Type: new Abstract: The Naive Bayes (NB) classifier remains a standard choice for categorical data, yet its widely used smoothing rules, such as Laplace, Lidstone, Krichevsky-Trofimov, and the $m$-estimate, all prescribe a fixed smoothing strength…
14 -
arXiv — Machine Learning research 2d ago
Beyond a Bag of Features: Set-Level Instability in Sparse Autoencoders
arXiv:2608.11197v1 Announce Type: new Abstract: Shani et al. (2026) show that LLM representations broadly recover human category boundaries, while failing to reflect fine-grained typicality structure. Their analysis uses cosine similarity over dense model representations. We…
22 -
arXiv — Machine Learning research 2d ago
Quantifying the noise sensitivity of the Wasserstein metric for images
arXiv:2510.01015v3 Announce Type: cross Abstract: Wasserstein metrics are increasingly adopted as similarity scores for images. We consider the sensitivity of Wasserstein metrics with respect to pixel-wise additive noise when the images are treated as discrete measures on the…
38 -
arXiv — Machine Learning research 2d ago
Optimized Sequential Testing for Binary Ensemble Classifiers
arXiv:2606.15237v1 Announce Type: cross Abstract: Ensemble classifiers are predictive models that combine the results of simpler base models, often by majority vote. A classic example is random forests, which combine the predictions of decision trees. Ensembles that use more…
8 -
arXiv — Machine Learning research 2d ago
Divergent Response Modes in Frontier Language Models Under Steering Pressure
arXiv:2608.06578v1 Announce Type: cross Abstract: Frontier language models are trained using distinct data, objectives, and safety pipelines. Whether these differences produce measurably different behaviors under explicit steering pressure remains underexplored. This study…
35 -
arXiv — Machine Learning research 2d ago
HyperShape: Hyperelasticity Across Diverse Shapes
arXiv:2608.09938v1 Announce Type: cross Abstract: Hyperelastic deformations are highly sensitive to domain geometry and boundary conditions, making generalization across both a critical capability for neural operators applied to these problems. However, existing benchmarks for…
34 -
-
arXiv — Machine Learning research 2d ago
EweAcT: Ewe behaviour aligned to accelerometer data for activity monitoring in extensive grazing systems
arXiv:2608.09943v1 Announce Type: cross Abstract: Monitoring livestock behaviour under extensive conditions would provide valuable insights to assess animal adaption to environmental perturbations in agroecological systems (e.g., heat waves, parasitism, predator attacks). Animal…
35 -
arXiv — Machine Learning research 2d ago
An adaptive and evolvable deep reinforcement learning framework for weather prediction
arXiv:2608.09948v1 Announce Type: cross Abstract: No single AI weather model excels at all variables, pressure levels, and lead times. Rather than building yet another architecture, we reframe the forecasting problem as one of coordination. Here we present Feitian Adaptive…
28 -
arXiv — Machine Learning research 2d ago
AIFS-TC: A simple correction competitive with the operational frontier for tropical cyclone intensity forecasting
arXiv:2608.09959v1 Announce Type: cross Abstract: AI weather models are in the process of revolutionising weather forecasting. While these models have been shown to achieve superior performance to physics-based NWP in forecasting tropical cyclone (TC) tracks, they dramatically…
24 -
arXiv — Machine Learning research 2d ago
Projected climate memory and inherited warm-tail risk in accelerated European summer warming
arXiv:2608.09966v1 Announce Type: cross Abstract: European summer warming reflects interactions among background change, persistent ocean--land--circulation states, and same-season variability. We develop an empirical reduced-dynamics framework that decomposes regional summer…
16 -
arXiv — Machine Learning research 2d ago
SPOTting the Future: Lookahead Explanations for Deep Reinforcement Learning
arXiv:2608.09967v1 Announce Type: cross Abstract: Deep reinforcement learning (DRL) agents achieve strong performance in complex environments, yet their decision-making processes remain difficult to interpret. We introduce SPOT (Sampling Policy Observation Tree), a novel…
35 -
arXiv — Machine Learning research 2d ago
Do AI weather models miss extremes?
arXiv:2608.09972v1 Announce Type: cross Abstract: First-generation AI weather models are often reported to underperform at extremes, mostly in reanalysis-based evaluations of deterministic regression systems. We verify eleven physical and AI forecast systems against European…
30 -
arXiv — Machine Learning research 2d ago
MIDAS: Mutual Information Disentanglement with Uncertainty-Aware Fusion for Incomplete Multimodal Sentiment Analysis
arXiv:2608.09986v1 Announce Type: cross Abstract: Most existing multimodal sentiment analysis approaches assume access to complete multimodal inputs. However, real-world applications frequently encounter incomplete or corrupted modalities, posing a critical challenge. Although…
29 -
arXiv — Machine Learning research 2d ago
Knowledge-Guided 3D CT Generation: A Conditioning-Centric Taxonomy
arXiv:2608.09992v1 Announce Type: cross Abstract: Controllable generation guided by external knowledge is a key requirement in modern generative deep learning applications, enabling the synthesis of samples with explicit constraints on semantic content, structural properties,…
6 -
arXiv — Machine Learning research 2d ago
Energy and Performance Benchmarking of Deep Learning Models for Breast Cancer Detection
arXiv:2608.09996v1 Announce Type: cross Abstract: Recent advances in machine learning have greatly improved breast cancer detection, enabling more accurate and timely diagnosis. Deep learning (DL) models show strong potential for medical image analysis; however, as their…
25 -
-
arXiv — Machine Learning research 2d ago
Do LLM Recommenders Know When They're Hallucinating? Auditing Confidence Calibration in Catalog Faithfulness
arXiv:2608.10008v1 Announce Type: cross Abstract: LLM recommenders for top-$K$ item suggestion regularly emit titles outside the target catalog. Prior audits measure this as a binary out-of-domain rate; none ask whether the model knew it was hallucinating. We jointly audit…
36 -
arXiv — Machine Learning research 2d ago
HIPNO: Symmetry-Aware Physics-Informed Neural Operators for Noninvasive Hemodynamic Inference
arXiv:2608.10011v1 Announce Type: cross Abstract: Continuous hemodynamic monitoring guides treatment decisions in surgery and intensive care. However, gold-standard signals are only measured in severe cases due to risks associated with invasive measurement. In this work, we…
30 -
arXiv — Machine Learning research 2d ago
Deep Learning-Based Statistical Downscaling of Sea Surface Temperature Using a Residual Corrective Neural Network
arXiv:2608.10022v1 Announce Type: cross Abstract: The large-scale oceanic and atmospheric forecasts provided by global climate models typically lack sufficient resolution to accurately capture the response of the coastal ocean to atmospheric forcing and coastal circulation that…
37 -
arXiv — Machine Learning research 3d ago
Application of Artificial Intelligence for Fraudulent Banking Operations Recognition
arXiv:2608.07471v1 Announce Type: new Abstract: This study considers the task of applying artificial intelligence to recognize bank fraud. In recent years, due to the COVID19 pandemic, bank fraud has become even more common due to the massive transition of many operations to…
4 -
arXiv — Machine Learning research 3d ago
Data-Driven Fire-Zone Segmentation for Improved Short-Term Wildfire Prediction
arXiv:2608.07472v1 Announce Type: new Abstract: Wildfire prediction models typically discretize study areas into uniform grids, ignoring the heterogeneous spatial distribution of ignitions. We challenge this paradigm by showing that how data is discretized matters more than…
35