arXiv — Machine Learning
500 articles archived · Visit source ↗ · RSS
-
arXiv — Machine Learning research 4d ago
Joint Distribution Alignment for Universal Domain Adaptation
arXiv:2608.24429v1 Announce Type: new Abstract: Unsupervised domain adaptation (UDA) has been widely concerned in the fields of machine learning, pattern recognition, and computer vision. Traditional UDA learning usually assumes that the label spaces of the source and target…
17 -
arXiv — Machine Learning research 4d ago
Evaluating Deep Multivariate Imputation Models on Wearable Device Data
arXiv:2608.24436v1 Announce Type: new Abstract: Wearable device data enables continuous health monitoring, but suffers from structured missingness: features sharing a physical sensor drop out together. Deep imputation methods such as BRITS and SAITS have seen limited evaluation…
33 -
arXiv — Machine Learning research 4d ago
WarpSAC: Towards the Pinnacle of Scalable Off-policy RL by Rethinking Exploration and Exploitation
arXiv:2608.24479v1 Announce Type: new Abstract: Massively parallel simulation changes the data regime in which off-policy reinforcement learning (RL) is trained, challenging stabilizers designed for data-limited replay. Through controlled experiments across eight benchmark…
35 -
arXiv — Machine Learning research 4d ago
Beyond Static Interpretability: Anticipating Post-SFT Mechanisms from Pre-SFT Parameters for Better Tuning
arXiv:2608.24482v1 Announce Type: new Abstract: Mechanistic Localization bridges mechanistic interpretability and post-training optimization by isolating critical parameters via interpretative approaches and then guiding parameter-efficient Supervised Fine-Tuning (SFT) in a…
38 -
arXiv — Machine Learning research 4d ago
Where Entropy Is Measured Matters: Policy Geometry in Bounded Continuous-Control PPO
arXiv:2608.24488v1 Announce Type: new Abstract: Many continuous-control policies are optimized as unbounded Gaussians and then mapped into bounded actions. We show that where entropy is measured changes the policy geometry learned by proximal policy optimization (PPO). In an…
29 -
arXiv — Machine Learning research 4d ago
When Do Supervised UQ Ensembles Improve LLM Hallucination Detection? A Robustness Study
arXiv:2608.24492v1 Announce Type: new Abstract: Uncertainty quantification (UQ) methods are widely used for hallucination detection in large language models (LLMs) in closed-book settings where ground-truth evidence is unavailable at inference time. Prior work has proposed…
17 -
-
arXiv — Machine Learning research 4d ago
From Numerical Simulators of PDEs to Neural Emulators and Back
arXiv:2608.24547v1 Announce Type: new Abstract: Simulation is central to modern engineering and science, but the cost of numerical solvers for partial differential equations (PDEs) remains a bottleneck whenever fast or many-query evaluations are required. Neural emulators…
28 -
arXiv — Machine Learning research 4d ago
Persistent Cross Entropy
arXiv:2608.24549v1 Announce Type: new Abstract: Persistent entropy is the Shannon entropy of a persistence-based probability measure defined on a persistence diagram. However, its cross-entropy version is not naturally defined because two persistence diagrams generally have…
4 -
arXiv — Machine Learning research 4d ago
FraudBench: Protocol-Sensitive Benchmarking of Adversarial Robustness for Financial Risk Assessment
arXiv:2608.24551v1 Announce Type: new Abstract: Machine learning models are widely used in financial fraud and credit-risk detection, yet their adversarial robustness remains difficult to evaluate because financial tabular data involve domain-specific constraints, severe class…
25 -
arXiv — Machine Learning research 4d ago
SeisMamba: Low-Latency Single-Station Seismic Magnitude Estimation for Spatially Distributed Earthquake Early Warning
arXiv:2608.24561v1 Announce Type: new Abstract: Rapid earthquake magnitude estimation is central to earthquake early warning, yet many operational systems depend on dense regional seismic networks and region-specific calibration. This creates a spatial coverage barrier for…
11 -
arXiv — Machine Learning research 4d ago
Across the Loss Landscape with Progressive Growth
arXiv:2608.24568v1 Announce Type: new Abstract: Deep neural networks generalize well despite their highly nonconvex, overparameterized loss landscapes, a phenomenon often associated with the geometry of the minima found by stochastic optimization. We study how incremental…
8 -
arXiv — Machine Learning research 4d ago
IAPO: Influence-Aware Policy Optimization for Credit Assignment in Multi-Turn Service Agents
arXiv:2608.24588v1 Announce Type: new Abstract: Large Language Model (LLM) agents increasingly solve long-horizon tasks through multi-turn interactions with users and external tools. In these settings, relevant task information often unfolds over time rather than being fully…
5 -
arXiv — Machine Learning research 4d ago
Delayed Optimizer-State Transport Shapes Short-Horizon Training Decisions
arXiv:2608.24593v1 Announce Type: new Abstract: Adaptive optimizers retain gradient history in moment variables, allowing a local change in loss weighting to alter later updates. We examine whether this delayed transport is large enough to change prospective short-horizon…
21 -
-
arXiv — Machine Learning research 4d ago
Conditional GraphGANFed: Optimizing Graph-Structured Molecule Generation in Federated Generative Adversarial Networks
arXiv:2608.24610v1 Announce Type: new Abstract: Generative adversarial networks (GANs) have garnered considerable attention in molecular discovery for their ability to generate novel and high-quality molecules. To efficiently train a GAN model while preserving data privacy,…
18 -
arXiv — Machine Learning research 4d ago
Bandit Submodular Maximization under Matroid Constraints: Learning Compressed Exchange Policy
arXiv:2608.24627v1 Announce Type: new Abstract: We study adversarial bandit maximization of monotone submodular functions under a matroid constraint. For a rank-$k$ matroid on $n$ elements, we give a randomized oracle-polynomial algorithm that makes one feasible value query per…
7 -
arXiv — Machine Learning research 4d ago
Data Leakage Inflates Generalizability of Power Outage Prediction Models
arXiv:2608.24665v1 Announce Type: new Abstract: Power outage prediction models are increasingly used in assessments of climate-driven infrastructure risk, yet current evaluation practices obscure whether these models generalize to the novel conditions such applications require.…
6 -
-
arXiv — Machine Learning research 4d ago
On-policy Distillation with Verifiable Reward
arXiv:2608.24696v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) and on-policy distillation (OPD) have become two widely adopted paradigms for post-training large language models. However, RLVR suffers from sparse task-level feedback, while…
26 -
arXiv — Machine Learning research 4d ago
Single State Update Predictive Coding training for Time Series Forecasting and Anomaly Detection
arXiv:2608.24697v1 Announce Type: new Abstract: Predictive Coding (PC) is a neural learning paradigm that enables parallelizable neural network layer updates. However, the main bottleneck of PC Networks (PCN) is the sequential backwards error propagation. To tackle this, we…
37 -
arXiv — Machine Learning research 4d ago
Parameter-Level Attribution of Symmetry in Trained Networks Though Parameter-Wise Functional Sensitivity
arXiv:2608.24700v1 Announce Type: new Abstract: When a network has learned a function with a known symmetry, can that symmetry be moved through the parametrisation---is there a motion in parameter space realising the group action in function space? We formulate this as a lifting…
9 -
arXiv — Machine Learning research 4d ago
Constrained Hyperparameter Optimization for Streaming Data
arXiv:2608.24712v1 Announce Type: new Abstract: Optimization of hyperparameters is a critical factor to obtain optimal model performance. While existing research has predominantly concentrated on batch-learning scenarios, addressing the complexities inherent in data streams…
37 -
arXiv — Machine Learning research 4d ago
Enhancing Bayesian Optimization and Active Learning Through Kernel Diversity
arXiv:2608.24721v1 Announce Type: new Abstract: Hyperparameter selection remains a key challenge in Bayesian optimization (BO) and Bayesian active learning (AL), as model misspecification can lead to suboptimal performance, while more accurate fully Bayesian treatments typically…
35 -
arXiv — Machine Learning research 4d ago
Parameter-Efficient Self-Supervised Adaptation for EEG-FM under Fixed Computational Budgets
arXiv:2608.24727v1 Announce Type: new Abstract: EEG foundation models pretrained via self-supervised learning promise transferable representations, but their generalization remains limited, especially across diverse clinical datasets. Full fine-tuning is impractical for…
10 -
arXiv — Machine Learning research 4d ago
Optimal Alternating Regret for Online Learning and Games
arXiv:2608.24731v1 Announce Type: new Abstract: We settle the minimax-optimal alternating regret, a regret notion motivated by alternating learning dynamics in games, for both online linear optimization (OLO) and online convex optimization (OCO). For OLO over the probability…
5 -
arXiv — Machine Learning research 4d ago
$(\text{DNN})^2$: Doubly Non-Negative Relaxations for Deep Neural Networks
arXiv:2608.24743v1 Announce Type: new Abstract: Existing linear program (LP) and semidefinite program (SDP) relaxations for rectified linear unit (ReLU) neural network (NN) verification yield overly-conservative safety guarantees due to significant relaxation gaps. While the…
6 -
arXiv — Machine Learning research 4d ago
Beyond Uniform Local Isometry and Topology: FactoMap for Disentangled Representations
arXiv:2608.24762v1 Announce Type: new Abstract: Many disentanglement methods represent generative factors using Euclidean product coordinates, although the underlying factor spaces may wrap, collapse, or have position-dependent geometry. We introduce factor-space structure,…
8 -
arXiv — Machine Learning research 4d ago
LION: A Clifford Neural Paradigm for Multimodal-Attributed Graph Learning
arXiv:2608.24795v1 Announce Type: new Abstract: Recently, the rapid advancement of multimodal domains has driven a data-centric paradigm shift in graph ML, transitioning from text-attributed to multimodal-attributed graphs. This advancement significantly enhances data…
14 -
arXiv — Machine Learning research 4d ago
MDTE: Minority-Aware Diffusion over Temporal Edge Events for Imbalanced Node Classification
arXiv:2608.24812v1 Announce Type: new Abstract: Class-imbalanced node classification on temporal graphs is challenging because majority-dominated temporal propagation progressively assimilates minority representations, while conventional node and neighborhood information…
28 -
arXiv — Machine Learning research 4d ago
Effective Learning Rate Governs Loss Dynamics in Language Model Pretraining
arXiv:2608.24814v1 Announce Type: new Abstract: We uncover ELR collapse in language model pretraining: learning rate (LR) and parameter norm govern loss dynamics primarily through their ratio, the effective learning rate (ELR). When ELR is matched across runs, their loss…
12 -
arXiv — Machine Learning research 4d ago
A Geometric Theory of Robust Fairness Audits
arXiv:2608.24818v1 Announce Type: new Abstract: Neighborhood-based fairness audits evaluate individual fairness by comparing predictions among similar individuals in feature space. Despite their widespread use, little is known about the robustness of the auditing procedure…
17 -
arXiv — Machine Learning research 4d ago
BioKERN: Biological Kernel Regularization for Histology-to-Transcriptomics Neighborhood Retrieval
arXiv:2608.24823v1 Announce Type: new Abstract: Spatially resolved biology requires representations that preserve biological neighborhood structure rather than only exact cross-modal correspondences. Existing histology--transcriptomics objectives can emphasize instance-level…
10 -
arXiv — Machine Learning research 4d ago
Bellman Calibration for Marginalized Importance Weighting in Offline Reinforcement Learning
arXiv:2608.24858v1 Announce Type: new Abstract: Marginalized importance weighting evaluates a target policy by reweighting offline state-action samples with its discounted occupancy ratio, characterized by an adjoint Bellman equation. Existing minimax, primal-dual, and fitted…
15 -
arXiv — Machine Learning research 4d ago
Improving Cross-Problem Vehicle Routing with Locally Augmented Preferences and Representation Disentanglement
arXiv:2608.24859v1 Announce Type: new Abstract: Multi-task vehicle routing problem (VRP) solvers seek to handle multiple VRP variants within a single unified model, avoiding the need to train a separate model for every variant. In spite of recent progress, current approaches…
12 -
arXiv — Machine Learning research 4d ago
Symbolic Classification-Enabled LHC Limits Online BSM Global Fits
arXiv:2605.22330v1 Announce Type: cross Abstract: Global fits of Beyond the Standard Model (BSM) physics often involve a two-way interplay between theory and experiment. Theoretical models provide guidance for experimental searches, while experimental results, in turn, constrain…
13 -
arXiv — Machine Learning research 4d ago
Finite-Sample Metric Non-Collapse for Geometrically Supervised Latent World Models in Control
arXiv:2608.07265v2 Announce Type: cross Abstract: We establish a finite-sample learning-to-control theory for geometrically supervised latent models of nonlinear deterministic systems. Geometric supervision is used only during training: simulator state, proprioception, or state…
15 -
arXiv — Machine Learning research 4d ago
DiD It in 87 Minutes: A Label-Free Softmax-to-Linear Adaptation of Vision Transformers for Object Detection
arXiv:2608.22368v1 Announce Type: cross Abstract: While linear attention is a compelling mechanism for high-resolution object detection due to its reduced cost for global token mixing, converting the Softmax-attention ViT backbone of a trained detector into a linear-attention…
28 -
arXiv — Machine Learning research 4d ago
InfoDPP-PAC: Principled Patch Selection for Whole Slide Image Analysis
arXiv:2608.23574v1 Announce Type: cross Abstract: Each WSI slide contains thousands of candidate tissue patches, while supervision is usually available only at slide level. Existing bag-construction strategies like Uniform extraction and handcrafted heuristics do not control…
7 -
arXiv — Machine Learning research 4d ago
Transformer Accelerator (TFA): A Macro-Op INT8 Hardware Chip for Transformer Inference and Machine Translation
arXiv:2608.23582v1 Announce Type: cross Abstract: We present the Transformer Accelerator (TFA), a synthesizable, parameterizable INT8 memory-to-memory engine for transformer inference. One time-multiplexed datapath handles prompt processing and autoregressive generation. TFA…
38 -
arXiv — Machine Learning research 4d ago
StateTune: Transforming LLM-Assisted EDA Flow Tuning into a Stateful, Closed-Loop Process
arXiv:2608.23601v1 Announce Type: cross Abstract: EDA flow parameter tuning is critical for quality-of-results~(QoR), yet the parameter space is large, tightly coupled, and full evaluations are prohibitively expensive. Prior LLM-assisted tuners mainly use the LLM as an external…
10 -
arXiv — Machine Learning research 4d ago
When May an Agent Stop? Evidence-Carrying Termination for Tool-Using LLMs
arXiv:2608.23623v1 Announce Type: cross Abstract: Tool-using agents must decide when to stop. Existing systems already gate terminal success, certify execution traces, or enforce runtime polici es, but do not test this particular receipt-, scope-, and closed-replay design at the…
8 -
-
arXiv — Machine Learning research 4d ago
Replicable Conformal Prediction
arXiv:2608.23638v1 Announce Type: cross Abstract: Two analysts who calibrate the same predictive model on independent samples will deploy different prediction sets every time, because the calibration threshold inherits the randomness of the data. Wherever deployments must be…
20 -
arXiv — Machine Learning research 4d ago
Contextual Embedding Evidence for Main--Light Verb Distinctions in Urdu
arXiv:2608.23645v1 Announce Type: cross Abstract: Urdu light verbs contribute schematic event-structural meaning while remaining lexically related to corresponding main verbs. This study tests representational predictions derived from Butt's analysis using contextual embeddings…
32 -
arXiv — Machine Learning research 4d ago
MolEmb: Multimodal Large Language Models Can Be Strong Molecular Embedding Models
arXiv:2608.23646v1 Announce Type: cross Abstract: Molecular embedding models can serve as foundational infrastructure for computational chemistry and drug discovery, where reusable vector representations support property prediction, virtual screening, and retrieval. Most…
6 -
arXiv — Machine Learning research 4d ago
Scaling Reinforcement Learning for Diffusion Models via Velocity Matching
arXiv:2608.23664v1 Announce Type: cross Abstract: Reward fine-tuning is becoming an important tool for adapting diffusion models to human preferences and task-specific objectives, but existing methods largely inherit policy-gradient machinery from large language models. Unlike…
17 -
arXiv — Machine Learning research 4d ago
Automata from Agent Traces: Failure and Next-Step Prediction
arXiv:2608.23670v1 Announce Type: cross Abstract: LLM-based agents execute multi-step tasks, but their behavioral structure remains opaque: long unstructured traces resist the safety auditing and runtime monitoring that deployment requires. Existing approaches operate per-trace…
5 -
arXiv — Machine Learning research 4d ago
A Hybrid Two-Stage Machine Learning Pipeline for Fault Detection and Classification in Power Transmission Systems
arXiv:2608.23726v1 Announce Type: cross Abstract: Rapid and accurate fault detection in high-voltage transmission networks is essential for grid reliability and equipment protection. Transmission fault datasets are frequently imbalanced, and certain fault types produce…
8 -
arXiv — Machine Learning research 4d ago
S-matrix informed neural networks for amplitude analysis
arXiv:2608.23750v1 Announce Type: cross Abstract: Reconstructing scattering amplitudes from finite, noisy, and mutually inconsistent measurements is an ill-posed inverse problem common to many reactions relevant to particle physics. We introduce S-matrix informed neural networks…
15