arXiv — Machine Learning
500 articles archived · Visit source ↗ · RSS
-
arXiv — Machine Learning research 4d ago
Every Cache Entry Earns Its Place: Global Allocation of Resolution and Coverage for KV Cache Compression
arXiv:2608.07001v1 Announce Type: new Abstract: As large language models (LLMs) process increasingly long contexts, KV cache storage and repeated access have become a major bottleneck. Existing KV cache compression methods rely on predefined, fixed compression rules and are…
30 -
arXiv — Machine Learning research 4d ago
Hyperbolic Graph Embedders for Link Prediction and Topology Reconstruction
arXiv:2608.07029v1 Announce Type: new Abstract: Hyperbolic embeddings provide compact geometric representations of complex networks in hyperbolic spaces, but systematic comparisons of methods developed in machine learning, network science, and algorithmics remain rare. We…
18 -
arXiv — Machine Learning research 4d ago
Accounting Graph Transformer for Short-History Multi-KPI Forecasting in Small Businesses
arXiv:2608.07037v1 Announce Type: new Abstract: Small businesses often have only 12-24 months of accounting history, yet planning and risk workflows require coordinated forecasts across financial statements. We study joint 12-month forecasting of 13 income-statement,…
24 -
arXiv — Machine Learning research 4d ago
Beyond Isolation: Unlocking Reinforcement Learning Component Synergy for Sample-Efficient Continuous Control
arXiv:2608.07086v1 Announce Type: new Abstract: Reinforcement learning systems are significantly more complex than other machine learning paradigms due to inherent properties, causing RL system design to jointly account for many tightly coupled factors. Despite advances in…
34 -
arXiv — Machine Learning research 4d ago
Synthetic LiDAR Data Generation and Deterministic Downsampling for Point Cloud Classification on the Edge
arXiv:2608.07106v1 Announce Type: new Abstract: Deploying three-dimensional deep learning frameworks to low-power embedded processors is bottlenecked by the unstructured nature of spatial data and the resource-intensive distance sorting algorithms often used before neural…
25 -
arXiv — Machine Learning research 4d ago
Modular TTT: Rethinking Test-Time Training as Composable Modules
arXiv:2608.07110v1 Announce Type: new Abstract: Test-time training (TTT) views sequence modeling as an online learning problem in which fast weights are updated by an internal learning rule. Despite the growing number of TTT variants, existing approaches typically hard-code each…
34 -
arXiv — Machine Learning research 4d ago
Online Conformal Prediction Beyond Feedback
arXiv:2608.07139v1 Announce Type: new Abstract: Uncertainty quantification is essential when deploying machine learning models in safety-critical applications. Online conformal prediction (OCP) provides theoretically principled uncertainty quantification for arbitrary black-box…
37 -
arXiv — Machine Learning research 4d ago
Interpretable reinforcement learning with decision-tree pruning
arXiv:2608.07151v1 Announce Type: new Abstract: Reinforcement learning policies are difficult to inspect, but interpreting them is a prerequisite for trustworthiness. Converting a trained policy into explicit decision-tree rules improves transparency and the resulting artifacts…
10 -
arXiv — Machine Learning research 4d ago
Machine Learning-Based Inter-Crystal Scatter Recovery for Ultra-High Resolution PET Imaging
arXiv:2608.07155v1 Announce Type: new Abstract: Inter-crystal scatter (ICS) events pose a significant challenge in ultrahigh- resolution positron emission tomography (UHR-PET), especially as detector crystals become smaller and their readouts increasingly segmented. Current…
7 -
arXiv — Machine Learning research 4d ago
Capacity Confounds and Coverage Guarantees in Adaptive Sub-model Federated Learning
arXiv:2608.07157v1 Announce Type: new Abstract: Sub-model federated learning lets resource-constrained clients train width-reduced versions of a global model, but existing methods allocate capacity by device resources alone. A natural next step, allocating capacity by each…
19 -
arXiv — Machine Learning research 4d ago
Edge Sparsification via Temporal Forman-Ricci Curvature for Dynamic Graph Learning
arXiv:2608.07158v1 Announce Type: new Abstract: Temporal graph learning has become essential for analyzing real-world systems whose interactions continuously evolve over time, including financial transaction networks, communication systems, and online social platforms. However,…
9 -
arXiv — Machine Learning research 4d ago
Fluid-DiT: Graph-Free Diffusion Transformers for Fluid Flow Simulations Learning
arXiv:2608.07161v1 Announce Type: new Abstract: Simulating complex fluid flows requires capturing full equilibrium distributions rather than just mean trajectories, yet high-fidelity solvers remain computationally prohibitive. Recent advances, such as Diffusion Graph Networks…
27 -
arXiv — Machine Learning research 4d ago
Momba: Network Modernization Improves Multi-Objective Reinforcement Learning
arXiv:2608.07180v1 Announce Type: new Abstract: Recent advances in deep reinforcement learning (RL) have shown that improving neural network architectures can yield substantial gains in sample efficiency and asymptotic performance without altering the underlying algorithms. In…
37 -
arXiv — Machine Learning research 4d ago
Conformal Fusion Under Missing Modalities
arXiv:2608.07183v1 Announce Type: new Abstract: Multimodal fusion architectures typically assume all modalities are available at inference, yet sensor failures, acquisition variability, and cost constraints routinely produce incomplete observations. Existing work treats modality…
32 -
arXiv — Machine Learning research 4d ago
MAUPITI: On-Device Prototype-Based Learning on a Smart Infrared Sensor
arXiv:2608.07192v1 Announce Type: new Abstract: Low-resolution infrared (IR) array sensors represent an interesting solution for privacy-preserving human sensing in embedded systems. In this letter, we describe a smart multi-pixel IR sensor integrating a 16$\times$16 thermal…
13 -
arXiv — Machine Learning research 4d ago
An AI4AI Framework for Visual Token Pruning
arXiv:2608.07193v1 Announce Type: new Abstract: Visual-token pruning can substantially reduce the inference cost of multimodal large language models (MLLMs), yet existing methods largely rely on fixed, handcrafted heuristics and costly expert trial and error. As pruning…
11 -
arXiv — Machine Learning research 4d ago
Stochastic Autoregressive Learning
arXiv:2608.07224v1 Announce Type: new Abstract: Motivated by LLMs, which generate outputs by iteratively sampling from next-token distributions, we introduce a PAC-learning model for binary stochastic autoregressive learning. This generalizes the deterministic autoregressive…
13 -
arXiv — Machine Learning research 4d ago
Learning Suffers More Than the Policy Class Under Partial Observability: A Closed-Form Analysis
arXiv:2608.07228v1 Announce Type: new Abstract: When a reinforcement learning agent cannot observe the full state, we usually blame its policies: it cannot see enough to represent a good one. We show that in a solvable case the bigger problem lies elsewhere. Even when a good…
35 -
arXiv — Machine Learning research 4d ago
TOFD: Target-Oriented Feature Decoupling against Poisoning Attacks in Split Federated Learning
arXiv:2608.07274v1 Announce Type: new Abstract: Split Federated Learning (SFL) facilitates privacy-preserving collaborative training with reduced client-side overhead. However, its split architecture introduces unique attack surfaces, rendering it vulnerable to diverse poisoning…
22 -
arXiv — Machine Learning research 4d ago
A foundation-model approach to pediatric headache classification from rs-fMRI
arXiv:2608.07287v1 Announce Type: new Abstract: Headache is the most common neurological disorder in children and substantially affects quality of life. We investigated whether resting-state functional MRI (rs-fMRI) can support pediatric headache classification using machine…
4 -
arXiv — Machine Learning research 4d ago
FUSE: Feature-Wise Unified Specialization with Cross-Column Exchange for Mixed-Type Tabular Flow Matching
arXiv:2608.07294v1 Announce Type: new Abstract: Generating mixed-type tabular data requires jointly modeling diverse feature distributions and their complex cross-column dependencies. Variational flow matching handles distinct endpoints via factorized distributions, yet leaves…
26 -
arXiv — Machine Learning research 4d ago
From Optimal Actions to World Models: Identifiability of Transition Kernels in Discounted MDPs
arXiv:2608.07301v1 Announce Type: new Abstract: We study what can be recovered about the transition probabilities of a Markov decision process from optimal actions alone. This is closely related to the inverse problem considered by Letcher et al., who ask when the dynamics can…
29 -
arXiv — Machine Learning research 4d ago
Is SwiGLU's Open Positive Tail Necessary? Evidence from Closed-Tail Gating with MemGLU
arXiv:2608.07323v1 Announce Type: new Abstract: We test whether decoder-only language-model FFNs require SwiGLU's open positive tail. We introduce MemGLU as a closed-tail comparator derived from a memristive branch geometry. Across paired 9M and 30M pretraining runs with three…
27 -
arXiv — Machine Learning research 4d ago
When GNNs Fail: Quantifying and Overcoming Temporal Correlation Volatility in Time Series
arXiv:2608.07333v1 Announce Type: new Abstract: Modeling multivariate time series by representing them as graphs, where individual series act as nodes and pairwise temporal corre- lations serve as edges, has gained significant traction. Recent advances in Graph Neural Networks…
6 -
arXiv — Machine Learning research 4d ago
Aftab: A Comprehensive Benchmark of CNN Encoders and Advanced Value Functions in Parallelized Q-Networks
arXiv:2608.07335v1 Announce Type: new Abstract: Recent advancements in deep reinforcement learning have increasingly favored simplified, highly parallelized paradigms. Notably, the Parallelized Q-Network (PQN) algorithm achieves stable off-policy learning without relying on…
9 -
arXiv — Machine Learning research 4d ago
Residual Algebra for Representation-Preserving Learning
arXiv:2608.07349v1 Announce Type: new Abstract: Learning from heterogeneous representations is usually reduced to feature concatenation, which erases which representation produced an error. We instead algebraize the residual: a representation is a typed object that owns both a…
14 -
arXiv — Machine Learning research 4d ago
Trajectory-Relative Hindsight Distillation for Agentic Reinforcement Learning
arXiv:2608.07371v1 Announce Type: new Abstract: Recent agentic reinforcement learning methods use hindsight to complement sparse outcome rewards. However, a completed rollout can yield many such signals, leaving their appropriate allocation across turns unclear. We introduce…
7 -
arXiv — Machine Learning research 4d ago
Omni-modal decomposition autoencoders learn full-stack wearable disentangled representations
arXiv:2608.07385v1 Announce Type: new Abstract: Learning disentangled representations is a key requirement for developing versatile, general-purpose, and sustainable models in multi-modal wearable computing. However, existing approaches do not operate as full-stack wearable…
14 -
arXiv — Machine Learning research 4d ago
FedDOSE: Federated Learning Framework Decomposing Site Effects for Modeling Brain Dynamic Functional Connectivity
arXiv:2608.07393v1 Announce Type: new Abstract: Functional Magnetic Resonance Imaging ( fMRI ) data are often pooled into collaborative multi-site consortia, as deep learning models for analyses require large datasets to generalize well. While Federated Learning (FL) offers a…
10 -
arXiv — Machine Learning research 4d ago
Beyond Post-Hoc Temperature Scaling: Bilevel Optimization for LLM Calibration
arXiv:2608.07419v1 Announce Type: new Abstract: Preference alignment often makes large language models (LLMs) overconfident and poorly calibrated. Traditional post-hoc temperature scaling is inherently domain-dependent: a temperature fitted on one domain does not generalize…
11 -
arXiv — Machine Learning research 4d ago
Beyond Myopic World Models: Long-Horizon End-to-End Training for Direct Future Prediction
arXiv:2608.07420v1 Announce Type: new Abstract: World models are expected to support imagination over extended temporal horizons, yet most are still trained through local few-step prediction objectives and deployed by recursively rolling out their own predictions. This creates a…
5 -
arXiv — Machine Learning research 4d ago
Diffusion LLMs as Targets and Adversaries: Mechanistic Safety Exploits
arXiv:2608.07430v1 Announce Type: new Abstract: Diffusion Large Language Models (DLLMs) replace autoregressive next-token prediction with iterative parallel denoising, yet their internal safety mechanisms remain poorly understood. In this work, we investigate DLLMs both as…
36 -
arXiv — Machine Learning research 4d ago
A proximal subgradient method for nonconvex stochastic optimization under the Kurdyka-{\L}ojasiewicz condition
arXiv:2608.05460v1 Announce Type: cross Abstract: This work introduces a proximal stochastic subgradient method for minimizing the sum of an expected cost, whose integrand is potentially nonsmooth and nonconvex, and a lower semicontinuous, prox-bounded function. We target a…
20 -
arXiv — Machine Learning research 4d ago
Interpretable Unsupervised Community Detection with LLM-Symbolized Structured Processes
arXiv:2608.06402v1 Announce Type: cross Abstract: Community detection is a fundamental task in graph analytics that aims to identify cohesive groups of entities with similar behaviors or interests. Classic objective-driven methods struggle with complex graph structures, while…
28 -
arXiv — Machine Learning research 4d ago
UAV3DCrop: Benchmarking 3D Reconstruction in Repeated Multi-Angle UAV Crop Surveys
arXiv:2608.06404v1 Announce Type: cross Abstract: Accurate 3D crop monitoring underpins data-driven precision agriculture by enabling field-scale analysis of plant structure, growth dynamics, and management response. Modern 3D reconstruction methods perform strongly on generic…
36 -
arXiv — Machine Learning research 4d ago
Deep Evidential Regression for Sparse Forest Height Estimation from Multimodal Satellite Imagery
arXiv:2608.06406v1 Announce Type: cross Abstract: Accurate estimation of forest height from satellite imagery is essential for applications such as carbon accounting, biodiversity monitoring, and ecosystem management. While recent deep learning approaches provide accurate…
38 -
arXiv — Machine Learning research 4d ago
Certified Feedforward Tracking for Unknown Nonlinear Systems via Invertible Neural Networks
arXiv:2608.06419v1 Announce Type: cross Abstract: In this paper, we address the certification of datadriven feedforward control for periodic tracking of unknown nonlinear systems under partial state measurements. To this end, we adopt an invertible neural network (INN) as a…
37 -
arXiv — Machine Learning research 4d ago
Multi Codec Discrete Diffusion Model for Text Guided Speech Inpainting and Editing
arXiv:2608.06424v1 Announce Type: cross Abstract: Speech recordings often contain missing, corrupted, or incorrect regions that must be reconstructed or modified without re-synthesizing the entire utterance. Speech inpainting restores missing segments, whereas speech editing…
29 -
arXiv — Machine Learning research 4d ago
NTDH: Complex Reasoning for Comprehensive Affective Analysis
arXiv:2608.06425v1 Announce Type: cross Abstract: Comprehensive affective analysis is challenging for two reasons: it spans heterogeneous prediction tasks with continuous, ordinal, and multi-label outputs, and affective meaning is context-dependent, requiring conflicting cues to…
9 -
arXiv — Machine Learning research 4d ago
Recovering Lesion Parameters from Aphasic Picture Naming Error Profiles in Large Language Models
arXiv:2608.06429v1 Announce Type: cross Abstract: Interpretability methods for large language models (LLMs) describe internal state but do not directly test whether that state is causally sufficient to produce the observed behavior. In earlier work, we lesioned LLMs to produce…
28 -
arXiv — Machine Learning research 4d ago
Fast and Accurate: An Adaptive VLA Inference Framework through Environment-aware Model Selection
arXiv:2608.06434v1 Announce Type: cross Abstract: Embodied intelligence demands both long-horizon reasoning and real-time closed-loop responsiveness. Recent dual-system Vision-Language-Action (VLA) architectures combine fast reactive control with slow deliberative reasoning to…
19 -
arXiv — Machine Learning research 4d ago
Game-Theoretic Inverse Reinforcement Learning for Modeling Competitive Human Driving: A Cut-in Prediction Study
arXiv:2608.06445v1 Announce Type: cross Abstract: Capturing the strategic decision-making inherent in competitive human driving is critical for autonomous vehicle safety and traffic simulation. This study demonstrates that game-theoretic Inverse Reinforcement Learning (IRL)…
25 -
arXiv — Machine Learning research 4d ago
FedTransKD-IDS: Robust Federated Transfer Learning with Knowledge Distillation for Intrusion Detection in IoT
arXiv:2608.06447v1 Announce Type: cross Abstract: In modern distributed network environments, particularly in Internet of Things infrastructures and 5G networks, stringent privacy preservation and scalability requirements have created significant challenges for intrusion…
10 -
-
arXiv — Machine Learning research 4d ago
LyEvO: Lyapunov-Guided Evolutionary Optimization for Safe and Robust Sim-to-Real Policy Learning
arXiv:2608.06481v1 Announce Type: cross Abstract: Training controllers that are safe and robust in simulation, and systematically assessing their readiness for real-world deployment, remain key challenges in sim-to-real transfer. To address this, we propose LyEvO, a…
31 -
-
arXiv — Machine Learning research 4d ago
TaskSense: Focusing on What Matters in World Models
arXiv:2608.06544v1 Announce Type: cross Abstract: World models for visual control typically learn compact latent states by reconstructing observations, implicitly encouraging representations to preserve information across the entire visual input. However, task-relevant content…
9 -
arXiv — Machine Learning research 4d ago
Cascade: Exploiting SLO-Aware latency budget for fair and high goodput LLM inference serving
arXiv:2608.06557v1 Announce Type: cross Abstract: The reasoning and agentic capabilities of large language models have expanded the range of applications they support, from short interactive exchanges to long, compute-heavy requests. LLM serving platforms today define…
34 -
arXiv — Machine Learning research 7d ago
MS-MLB: An Open Machine Learning Benchmark for Blood-Based MS Classification
arXiv:2608.05196v1 Announce Type: new Abstract: Multiple sclerosis (MS) is diagnosed through clinical assessment, magnetic resonance imaging, laboratory evidence when appropriate, and exclusion of better explanations. Blood RNA expression data may contain disease associated…
14 -
arXiv — Machine Learning research 7d ago
When Do Corrective Features Help? An Agent for Corrective Feature Discovery on Black-Box Forecasters
arXiv:2608.05207v1 Announce Type: new Abstract: Frozen pretrained forecasters often fail in structured, recurring ways that are costly to repair through fine-tuning. We study corrective feature discovery: mining interpretable features of a frozen forecaster's residual to drive a…
33