News / #paper Tag Research papers 500 articles archived under #paper · RSS Sign in to follow arXiv — Machine Learning research 1d ago Learning with Bilevel-Minimax Optimization for Efficient and Reliable Transfer Attacks arXiv:2608.11815v1 Announce Type: new Abstract: Transfer-based adversarial attacks craft adversarial examples using surrogate models to mislead black-box victim models. Beyond perturbation generation, transferability is fundamentally governed by the coupling of initialization,… 18 arXiv — NLP / Computation & Language research 1d ago Towards Understanding On-Policy Distillation through the Lens of Test-Time Scaling arXiv:2608.11829v1 Announce Type: cross Abstract: On-policy distillation (OPD) has emerged as a promising post-training technique for enhancing LLM reasoning. It is commonly believed to enable the student model to distill knowledge from a stronger teacher model, thereby… 29 arXiv — Machine Learning research 1d ago Kernel Methods for Learning Operators with Multiple Inputs and Outputs arXiv:2608.11831v1 Announce Type: new Abstract: Learning mappings between infinite-dimensional objects is a central challenge in scientific machine learning. We introduce a general kernel-based encoder-decoder framework for operator learning that separates observation,… 12 arXiv — Machine Learning research 1d ago Air Quality Station Simulation via LSTM and Attention-Based Modelling arXiv:2608.11839v1 Announce Type: new Abstract: Poor air quality in urban areas is driven by a complex chain of processes and presents a significant public health concern. To better understand and control the mechanisms that determine air quality, cities deploy networks of… 24 arXiv — Machine Learning research 1d ago Small-Scale Experiments: Are We There Yet? arXiv:2608.11859v1 Announce Type: new Abstract: Scaling laws promised cost-effective experiments; six years later, they have yet to fully deliver. Instead, researchers have found them unreliable at small scales (starting at 4M parameters) and concluded that sizable models cannot… 28 arXiv — Machine Learning research 1d ago Forward and Inverse Virtual Metrology for Phototransistor Gain: A Hierarchical, Uncertainty-Aware Approach for Small Production Datasets arXiv:2608.11868v1 Announce Type: new Abstract: The customization, optimization and stabilization of the process flow of a silicon bipolar phototransistor commits months of cleanroom time before a finished device can be measured, so a model that predicts device gain from process… 37 arXiv — Machine Learning research 1d ago DCM Bandits: Multiplayer Information Asymmetric Cascading Bandits for Multiple Clicks arXiv:2608.11873v1 Announce Type: new Abstract: In this work, we extend the Dependent Click Model (DCM) Bandits to a multiplayer information-asymmetric setting, where multiple agents interact with a shared ranked list and may observe multiple clicks per session, introducing new… 30 arXiv — Machine Learning research 1d ago Disentangling the Expressivity of RoPE arXiv:2608.11909v1 Announce Type: new Abstract: Two accounts recur in explanations of the success of rotary position embeddings (RoPE). Expressivity studies associate periodic position information with modular predicates, whereas mechanistic and long-context studies emphasize… 16 arXiv — Machine Learning research 1d ago A Factor Graph Approach to Scalable Multi-Output Gaussian Process Regression arXiv:2608.11917v1 Announce Type: new Abstract: Multi-output Gaussian process regression scales cubically in the number of observations times outputs, and dense kernel-matrix methods need bespoke handling whenever different outputs are observed at different inputs. We express… 36 arXiv — Machine Learning research 1d ago Distillation of Foundation Models for Time-dependent PDEs arXiv:2608.11937v1 Announce Type: new Abstract: Foundation models for time-dependent partial differential equations (PDEs) are trained on large and diverse collections of physical systems and can generalize effectively to new downstream tasks. After fine-tuning on only a few… 25 arXiv — Machine Learning research 1d ago TailBooster: A Dual-Layer Generative Framework for Extreme Value Augmentation with Operational Validity Enforcement arXiv:2608.11951v1 Announce Type: new Abstract: Extreme events in air transport, such as severe arrival delays and abnormal air times, cause cascading network disruptions with substantial operational, economic, and safety costs. Such events are rare in historical records,… 38 arXiv — Machine Learning research 1d ago LoongReflect: Boosting Long-Horizon Reflection in Search Agents via Global Perspective Distillation arXiv:2608.11967v1 Announce Type: new Abstract: Large language model agents increasingly rely on long-horizon reasoning to solve complex tasks involving planning, tool use, and memory. A critical capability in such settings is reflection: assessing trajectory progress,… 31 arXiv — Machine Learning research 1d ago TESLA: Taylor Expansion of Sinusoidal Learnable Activations arXiv:2608.11970v1 Announce Type: new Abstract: The parity problem--deciding whether the number of ones in a binary vector is odd or even--remains challenging for standard neural networks due to linear inseparability and the need for global interactions. We propose TESLA, an… 15 arXiv — Machine Learning research 1d ago Remote Sensing and Machine Learning-Based Analysis of Land Use and Vegetation Change in Dhaka District, Bangladesh arXiv:2608.12001v1 Announce Type: new Abstract: Rapid urbanization in Dhaka District, Bangladesh has triggered substantial alterations in land use and environmental conditions, necessitating systematic monitoring for informed urban planning and ecological sustainability. This… 27 arXiv — Machine Learning research 1d ago Dual-Model Sentiment Analysis of Consumer Reviews in the Retail Coffee Sector Using Machine Learning and Deep Learning Approaches arXiv:2608.12007v1 Announce Type: new Abstract: Consumer reviews play an important role in shaping brand perception and business strategies, particularly in service-driven industries such as retail coffee. This study presents a comparative sentiment analysis framework for… 15 arXiv — Machine Learning research 1d ago Reducing Symmetry Increase in Equivariant Neural Networks arXiv:2608.12010v1 Announce Type: new Abstract: Equivariant Neural Networks (ENNs) have empowered numerous applications in scientific fields. Despite their remarkable capacity for representing geometric structures, ENNs suffer from degraded expressivity when processing symmetric… 11 arXiv — Machine Learning research 1d ago SoftWater: Class-Aware Rate Allocation for Softmax Quantization arXiv:2608.12026v1 Announce Type: new Abstract: Post-training quantization pipelines routinely leave the softmax output layer in high precision. Yet in small LLMs with modern vocabularies, the head holds 15--30\% of all parameters, so a nominal ``2-bit'' model with an fp16 head… 27 arXiv — Machine Learning research 1d ago Uncertainty-Aware Probabilistic Constrained Clustering from Entangled Pairwise Supervision arXiv:2608.12027v1 Announce Type: new Abstract: Pairwise constrained clustering typically relies on hard must-link/cannot-link labels, whereas realistic pairwise supervision may be real-valued and entangle intrinsic ambiguity, expert judgment, and stochastic corruption. Existing… 13 arXiv — Machine Learning research 1d ago Clustered Randomized Smoothing for Stochastic Prediction Functions arXiv:2608.12037v1 Announce Type: new Abstract: Modern stochastic predictors can model rich, multi-modal outcome distributions. However, this expressive power comes with challenges in ensuring robust predictions $-$ a critical requirement in safety-critical domains. Randomized… 8 arXiv — Machine Learning research 1d ago Towards Truly Unsupervised Evaluation of Feature Selection arXiv:2608.12057v1 Announce Type: new Abstract: Feature selection is one of the most important and fundamental tasks in data mining, tackled by a family of methods with an established set of evaluation techniques to measure the quality of a specific method. Most of the methods… 9 arXiv — Machine Learning research 1d ago Faithful, Sufficient and Understandable: Rethinking Graph Counterfactual Explanations via Discrete Diffusion Inversion arXiv:2608.12083v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) achieve strong predictive performance on graph-structured data across domains such as chemistry, biology, and network analysis, yet they provide no intrinsic explanation of their predictions. This… 23 arXiv — Machine Learning research 1d ago NAE: Normalizing AutoEncoder arXiv:2608.12084v1 Announce Type: new Abstract: We consider the setting of Normalizing flows with approximate inverses, an established paradigm spanning both full-dimensional ($d=D$) and bottleneck ($d<D$) settings, and group these models under the term flow autoencoders. We… 9 arXiv — Machine Learning research 1d ago Task- and dataset-specific information in protein language models arXiv:2608.12090v1 Announce Type: new Abstract: Protein language models (PLMs) have transferred the latest advances from natural language processing to computational biology. These models, trained on large corpora of protein sequence data, are widely used to translate amino acid… 17 arXiv — Machine Learning research 1d ago Confidence Calibration of Deep Learning Systems arXiv:2608.12100v1 Announce Type: new Abstract: In high-stakes applications, reliable confidence estimates are as important as the predictions themselves. Confidence calibration ensures that predicted probabilities reflect the likelihood of correctness, making it essential for… 36 arXiv — Machine Learning research 1d ago Beyond Parameter Space: NTK-Guided Personalized Aggregation for Robust Federated Learning arXiv:2608.12108v1 Announce Type: new Abstract: Federated learning (FL) enables collaborative model training across distributed clients while keeping data local. A central challenge is determining which client updates are beneficial for aggregation with respect to each client's… 6 arXiv — Machine Learning research 1d ago Attractor Image-Based Deep Learning of Arterial Pulse Waves for Age Classification arXiv:2608.12117v1 Announce Type: new Abstract: Arterial pulse waveform morphology evolves with age, reflecting structural and functional changes in the cardiovascular system. Thus, vascular age is a valuable surrogate marker of cardiovascular health, and premature vascular… 22 arXiv — Machine Learning research 1d ago Adversarial Resilience of Poisson-Process Submodular Maximization over Matroids: From Robust Offline Optimization to Full-Bandit Learning arXiv:2608.12134v1 Announce Type: new Abstract: We study nonnegative submodular maximization subject to a general matroid when the offline algorithm is given an arbitrary controlled value oracle. Our main result is an adversarial resilience theorem for the Spiteful Greedy Swap… 6 arXiv — Machine Learning research 1d ago HYDRA: Hyperbolic Dynamic Representation Architecture for Kolmogorov-Arnold Networks arXiv:2608.12194v1 Announce Type: new Abstract: Kolmogorov-Arnold Networks (KANs) enhance nonlinear function approximation by replacing scalar weights with learnable univariate functions. However, assigning an independent function to every connection results in substantial… 4 arXiv — Machine Learning research 1d ago ScreenShot: A Foundation Model for Few-Shot Combination Drug Screening arXiv:2608.12219v1 Announce Type: new Abstract: Treating patients with combinations of drugs reduces the risk of resistance to any individual drug. Finding effective combinations is difficult because the large search space makes combinatorial screens prohibitively expensive,… 6 arXiv — Machine Learning research 1d ago An Efficient Near-Optimal Algorithm for Adversarial $m$-Set Bandits arXiv:2608.12231v1 Announce Type: new Abstract: We study adversarial combinatorial bandits with $m$-set actions, where at each round the learner selects $m$ out of $d$ items and observes only the aggregate loss of the selected items. The resulting action set contains… 31 arXiv — Machine Learning research 1d ago Calibration Bets on the Past: Post-Training Quantization for Financial Time-Series Forecasting arXiv:2608.12259v1 Announce Type: new Abstract: Financial forecasting models are typically developed in full precision, yet production deployment often requires low-precision inference to reduce memory and computational cost. Post-training quantization (PTQ) enables such… 38 arXiv — Machine Learning research 1d ago Earth observation embeddings are effective sub-grid descriptors for probabilistic weather downscaling arXiv:2608.12271v1 Announce Type: new Abstract: Global weather reanalyses and forecasts resolve the evolving atmospheric state on coarse grids, but site-specific applications require predictions at arbitrary locations where near-surface conditions also depend on unresolved… 21 arXiv — Machine Learning research 1d ago A Framework for Designing Reward Functions: From Objectives to Features to Human-Aligned Reward Functions arXiv:2608.12302v1 Announce Type: new Abstract: We present a formal process to enable non-experts to instantiate and iterate on human-aligned reward functions, i.e. reward functions that adhere to a given preference ordering over trajectories. Given a task described in natural… 20 arXiv — Machine Learning research 1d ago Redistribution-based Cost Inference Improves Sparse Safe Offline RL arXiv:2608.12306v1 Announce Type: new Abstract: Safe offline RL typically assumes access to dense per-step cost annotations, but in practice supervisors provide only trajectory-level stop-feedback: a binary signal at the first unsafe transition, with no per-step attribution. We… 29 arXiv — NLP / Computation & Language research 1d ago AI4AI at Test-Time: Strong-to-Weak Capability Transfer via Harnesses arXiv:2608.12307v1 Announce Type: cross Abstract: Recent work on distillation transfers the capabilities of large models to smaller ones often by updating the latter's parameters, through teacher forcing, on-policy distillation, and related training-time methods. In this paper,… 33 arXiv — NLP / Computation & Language research 1d ago Reproducing and Stress-Testing Two Approaches to LLM Reasoning Reliability: Test-Time Probability Aggregation and Logic-Representation Editing arXiv:2608.08514v1 Announce Type: cross Abstract: We independently reproduce two recent methods for making large language model (LLM) reasoning more reliable, and stress-test them across domains and models (RPC across four new task domains with Qwen3-8B, LCF across four 7-8B… 35 arXiv — NLP / Computation & Language research 1d ago Detecting a Route Flip Is Easier Than Knowing Whether to Fix It: Causal Route-Mediated Damage in Quantized Mixture-of-Experts arXiv:2608.11212v1 Announce Type: cross Abstract: Top-k Mixture-of-Experts (MoE) routing is discontinuous, so a deployment-motivated numerical disturbance -- simulated 4-bit KV-cache quantization read by a protected BF16 gate -- pushes tokens across decision boundaries and flips… 18 arXiv — NLP / Computation & Language research 1d ago Backtrader-Bench: Benchmarking LLM Agents on Algorithmic Trading with Self-Generated MCQs arXiv:2608.11232v1 Announce Type: new Abstract: Evaluating LLM coding agents in algorithmic trading is difficult because static benchmarks risk data contamination and numerical backtest outputs require ground truth from actual code execution. We present Backtrader-Bench, a… 18 arXiv — NLP / Computation & Language research 1d ago Retrofitting Recurrent Depth into a Pretrained Language Model: Installation, Extrapolation, Transfer, and Retention at Two Parameter Budgets arXiv:2608.11233v1 Announce Type: new Abstract: A dense, pretrained language model can be retrofitted with recurrent depth and learn an iterative latent transition that persists after outcome-only annealing. Qwen2.5-0.5B-Instruct is split into a Prelude, a weight-tied Recurrent… 19 arXiv — NLP / Computation & Language research 1d ago TRACE Bench: Task-driven Roleplay Agentic Checklist Evaluation arXiv:2608.11236v1 Announce Type: new Abstract: Roleplay evaluation should do more than assign a single score: it should reveal which role requirements were tested, which failed, and which dialogue evidence supports the judgment. We propose TRACE Bench, a task-driven agentic… 28 arXiv — NLP / Computation & Language research 1d ago Lost in Compaction: Evaluating Side-Constraint Loss under Context Compaction arXiv:2608.11242v1 Announce Type: new Abstract: When the context window is under pressure, LLM systems compact prior context to continue ongoing tasks. We identify a class of user-issued instructions, Session Constraints (SCs), such as "do not delete any emails until I confirm,"… 27 arXiv — NLP / Computation & Language research 1d ago Diffuse to Compress: Leveraging Diffusion LMs for Lossless Compression arXiv:2608.11249v1 Announce Type: new Abstract: We study the problem of lossless text compression, motivated by the rapid growth in the collection and storage of digital textual data - including plain text, source code, and structured formats such as XML - and by recent advances… 37 arXiv — NLP / Computation & Language research 1d ago Gloss-Free Representation Learning for Cross-Dataset Sign Spotting arXiv:2608.11332v1 Announce Type: new Abstract: Sign-language research for resource-constrained languages is often limited by the cost of dense linguistic labels such as glosses, temporal boundaries, and sign order. Broadcast news offers a practical alternative by pairing… 38 arXiv — NLP / Computation & Language research 1d ago Better, Faster, Stronger: Programmatic Skill Learning Best Reduces Agent Cost arXiv:2608.11338v1 Announce Type: new Abstract: Recently, the practice of augmenting LLM agent capability with skills has gained prevalence. We explore the cost effective adaptation of agents to novel domains by means of learning skills. Existing works focus on performance gain… 19 arXiv — NLP / Computation & Language research 1d ago Self-Evolving Embodied Agents via Skill-Harness Evolution arXiv:2608.11350v1 Announce Type: new Abstract: Embodied agents are increasingly built as systems around foundation models, where performance depends not only on model weights but also on the skills, context, action interfaces, and execution harness surrounding the model. While… 12 arXiv — NLP / Computation & Language research 1d ago ODE-Based Transformer Decoders for Iterative Sign Language Translation arXiv:2608.11352v1 Announce Type: new Abstract: Sign language translation has achieved strong results with Transformer architectures, yet recent improvements largely rely on scaling model capacity at the cost of increased computation. We propose a parameter-efficient alternative… 25 arXiv — NLP / Computation & Language research 1d ago Measure, Don't Optimize: Forecasting Recovery in LLM Unlearning arXiv:2608.11408v1 Announce Type: new Abstract: Prior white-box studies show that large language models can retain latent traces of target knowledge after unlearning, even when the knowledge is no longer expressed in their outputs. However, existing audits remain limited to… 25 arXiv — NLP / Computation & Language research 1d ago Is Convergence Inevitable? Tracing Output Homogeneity Back to Base Models arXiv:2608.11426v1 Announce Type: new Abstract: The lack of diversity in LM content is widely attributed to the alignment process, but how and where exactly in the pipeline this collapse begins is unknown. We argue that output homogeneity is likely learned during the pretraining… 12 arXiv — NLP / Computation & Language research 1d ago Stigma and Support in Online Sexual Violence Narratives on Reddit arXiv:2608.11433v1 Announce Type: new Abstract: Online communities increasingly provide spaces where survivors of sexual violence can share their experiences and seek support. Although prior research has examined stigma and social support separately, less is known about how… 22 arXiv — NLP / Computation & Language research 1d ago DonorRank: Donor Language Selection for Low-Resource Cross-Lingual Speech Recognition arXiv:2608.11441v1 Announce Type: new Abstract: Low-resource automatic speech recognition (ASR) commonly relies on cross-lingual transfer, where models are adapted from higher-resource donor languages. However, selecting donors remains challenging for spontaneous speech from… 18 Page 6 of 10 · 500 articles ← Newer Older →