arXiv — Machine Learning
500 articles archived · Visit source ↗ · RSS
-
arXiv — Machine Learning research 9d ago
Towards On-Board Implementation of ML-Based Helicopter Weight Estimator
arXiv:2608.19210v1 Announce Type: new Abstract: This paper focuses on the implementation of a novel supervised Machine Learning model for estimating helicopter weight during takeoff, utilizing extensive datasets from Airbus's global in-service fleet. The study details a learning…
34 -
arXiv — Machine Learning research 9d ago
Triangular Fuzzy Rescaling Distance
arXiv:2608.19234v1 Announce Type: new Abstract: Decision-making in complex systems often involves dealing with imprecise or uncertain information, frequently represented using fuzzy sets, particularly Triangular Fuzzy Numbers (TFNs). A crucial aspect of many fuzzy methods is the…
28 -
arXiv — Machine Learning research 9d ago
Holtercare-Bench: A Multimodal Benchmark for Evaluating Long-Term Dynamic ECG Analysis
arXiv:2608.19297v1 Announce Type: new Abstract: While multimodal large language models (MLLMs) excel in medical applications, most of them favor static images or short-term signals. In the critical field of dynamic electrocardiograms (ECG), models struggle with complex temporal…
19 -
arXiv — Machine Learning research 9d ago
Quantum Kernel Estimation for the Discovery of Early Lung Cancer Detection
arXiv:2608.19304v1 Announce Type: new Abstract: Lung cancer screening with low-dose chest computed tomography reduces mortality, but its impact is limited by uptake, adherence, and management challenges. Blood-based cell-free DNA (cfDNA) biomarkers offer a complementary…
26 -
arXiv — Machine Learning research 9d ago
Improved Confidence Estimates for Black-Box Large Language Models
arXiv:2608.19323v1 Announce Type: new Abstract: Uncertainty quantification (UQ) is essential for the safe deployment of large language models (LLMs). Existing methods, from verbalized confidence to ones requiring multiple generations, are often zero-shot and produce scores…
16 -
arXiv — Machine Learning research 9d ago
Mechanistic Tomography: Designed Measurement for Control-Oriented Interpretability
arXiv:2608.19338v1 Announce Type: new Abstract: Mechanistic interpretability seeks quantities that models do not expose directly: represented states, component effects, interactions, and responses to interventions. Patching, gradients, Hessian-vector products, and subset…
17 -
arXiv — Machine Learning research 9d ago
Uncovering the Limits of Proof Sharing for Neural Networks
arXiv:2608.19351v1 Announce Type: new Abstract: Robustness verification of neural networks is increasingly important, due to their use in many critical domains. In certain scenarios, proof sharing has been shown to accelerate incomplete verification techniques by reusing…
36 -
arXiv — Machine Learning research 9d ago
Longitudinal Bayesian Learning of Continuous Disease Position across the Alzheimer's Disease Continuum
arXiv:2608.19436v1 Announce Type: new Abstract: Alzheimer's disease (AD) progresses as a continuous biological process, whereas most existing neuroimaging-based artificial intelligence methods remain limited to discrete diagnosis or clinical score prediction from cross-sectional…
9 -
arXiv — Machine Learning research 9d ago
Quantifying Event Impacts on Time Series via Multiscale Contrastive Learning
arXiv:2608.19447v1 Announce Type: new Abstract: Shocks that spread through the web, such as cybersecurity breach disclosures, can abruptly disrupt financial time series and cause substantial abnormal losses. While these events are disclosed as discrete records through news…
31 -
arXiv — Machine Learning research 9d ago
LLM as Detector: An In-context Learning Approach for Tabular Anomaly Detection
arXiv:2608.19463v1 Announce Type: new Abstract: Anomaly detection in tabular data is challenging because abnormal samples often arise as violations of cross-feature dependencies rather than simple marginal deviations. Existing detectors rely on geometric or reconstruction…
35 -
-
arXiv — Machine Learning research 9d ago
DeltaMomentum: A Key-Value based Anisotropic Momentum Update via Delta Rule
arXiv:2608.19491v1 Announce Type: new Abstract: Most modern optimizers form their momentum as an exponential moving average (EMA) of past gradients, forgetting every direction at one fixed rate. However, the inputs a deep network sees during training can be highly anisotropic,…
37 -
arXiv — Machine Learning research 9d ago
Beyond Multimodal Alignment: Certifying Physical Language through Response Substitution and Ordered Execution
arXiv:2608.19492v1 Announce Type: new Abstract: World models increasingly treat compact multimodal representations as interfaces between perception and physical interaction, yet existing probes do not establish whether different sensors carry the same executable meaning or…
20 -
arXiv — Machine Learning research 9d ago
Empirical Characterization of Learning Geometry in Hybrid Quantum Forecasting Models
arXiv:2608.19497v1 Announce Type: new Abstract: We characterize the learning dynamics of a compact hybrid quantum forecasting model through comparison with a structurally aligned classical baseline. Using stationary harmonic-mixture and nonstationary chirp benchmarks with…
4 -
arXiv — Machine Learning research 9d ago
In Two Minds about Lifelong Learning: Exploring Hemispheric Redundancy and Specialisation in Neural Models
arXiv:2608.19514v1 Announce Type: new Abstract: Persistent intelligent systems require the ability to learn continually, but current machine learning approaches face significant challenges in this area compared to biological learning systems. Machine learning algorithms…
27 -
arXiv — Machine Learning research 9d ago
Continuous Adversarial MeanFlow Transfer
arXiv:2608.19540v1 Announce Type: new Abstract: Training fast generators on new domains with limited data remains challenging for two reasons. First, adapting a pretrained diffusion or flow model to a new domain leaves its costly multi-step sampling unaddressed, and existing…
26 -
arXiv — Machine Learning research 9d ago
DraftFM: A FoundationModel for Day-Zero Drafting in Magic: The Gathering
arXiv:2608.19568v1 Announce Type: new Abstract: Drafting a new Magic: The Gathering expansion begins before any pick from it has been observed: the complete card list is public, but the draft logs that supervised pick models train on do not yet exist. We study this day-zero…
24 -
arXiv — Machine Learning research 9d ago
A Two-Stage Time-Aware Transformer for Short-Horizon AECOPD Risk Prediction
arXiv:2608.19578v1 Announce Type: new Abstract: Acute exacerbation of chronic obstructive pulmonary disease (AECOPD) can worsen rapidly, making timely prediction a clinical priority. Most existing machine learning approaches rely on episodically collected clinical variables,…
21 -
-
arXiv — Machine Learning research 9d ago
Unregularized Convergence of Single-Loop, Entropy-Regularized Natural Actor-Critic
arXiv:2608.19587v1 Announce Type: new Abstract: While entropy regularization is widely used to stabilize and accelerate Natural Policy Gradient methods, its ability to yield faster convergence rates for the unregularized objective remains underexplored. Existing analyses often…
35 -
-
arXiv — Machine Learning research 9d ago
Time-Uniform Self-Normalized Concentration for Discounted Least Squares: Limits and Corrections
arXiv:2608.19643v1 Announce Type: new Abstract: Self-normalized concentration inequalities are standard tools in bandit and reinforcement-learning analyses. A widely used weighted extension claims an analogous time-uniform guarantee for discounted least-squares estimators in…
20 -
arXiv — Machine Learning research 9d ago
DeltaML-Bench: Evaluating Machine Learning Agents on Real-World Research Repositories
arXiv:2608.19653v1 Announce Type: new Abstract: Autonomous agents for machine learning experimentation must navigate heterogeneous repositories, repair training pipelines, and evaluate candidate improvements under realistic compute constraints. Existing benchmarks only partially…
6 -
arXiv — Machine Learning research 9d ago
Rationally Enriched Chebyshev Trunk Bases for DeepONet Surrogates of High P\'eclet Entrance Transport
arXiv:2608.19658v1 Announce Type: new Abstract: This study demonstrates a rationally enriched Chebyshev (REC) trunk for deep operator network (DeepONet) surrogate models of singularly perturbed and high-P\'eclet transport problems whose solution profiles are characterized by…
32 -
arXiv — Machine Learning research 9d ago
FleetSieve: Decision-Critical Profiling for SLO-Aware LLM Fleet Configuration
arXiv:2608.19659v1 Announce Type: new Abstract: Choosing tensor-parallel (TP) degrees and replica counts for an LLM serving fleet is difficult because performance is not monotonic in TP and the feasible choice can change with load. Exhaustive profiling resolves this uncertainty,…
38 -
-
arXiv — Machine Learning research 9d ago
A Locally Tokenized Generative Model for Robust Time-Series Watermarking
arXiv:2608.19727v1 Announce Type: new Abstract: Watermarking is a central tool for provenance in generative models, yet its application to multivariate time series remains hindered by reliability failures under post-editing attacks. We show that existing detectors, which rely on…
21 -
arXiv — Machine Learning research 9d ago
RecPFN: Prior-Fitted Networks for In-Context-Based Recommendations
arXiv:2608.19735v1 Announce Type: new Abstract: We introduce RecPFN, a prior-fitted network that brings in-context learning to sequential recommendation. RecPFN is pretrained entirely on synthetic clickstream environments sampled from a broad structural causal prior, enabling it…
11 -
arXiv — Machine Learning research 9d ago
Truncate Bad, Upweight Good: BoN-Style Distillation via Rank-Based Classification
arXiv:2608.19748v1 Announce Type: new Abstract: Inference-time selection methods, such as Best-of-N, improve generation by sampling a pool of candidates and selecting the top-ranked completion according to a reward model. Distillation seeks to amortize this procedure into a…
28 -
arXiv — Machine Learning research 9d ago
Credit Without Ground Truth: Auditing Step-Level Credit Assignment in LLM Agents Against Executed Replay
arXiv:2608.19760v1 Announce Type: new Abstract: Audited against causal ground truth from executed replay in a single-agent tool environment (ALFWorld), none of the step-level credit signals used to train LLM agents -- LLM-judge scores, outcome-conditioned logprob ratios, or the…
8 -
arXiv — Machine Learning research 9d ago
Finite-Horizon Input-Output Dynamics of Minibatch Perturbations in AdamW
arXiv:2608.19762v1 Announce Type: new Abstract: A minibatch can influence training beyond the update at which it is observed because AdamW stores past gradient information in its optimizer states. We study this delayed effect through paired trajectories that differ only in one…
33 -
arXiv — Machine Learning research 9d ago
Unsupervised Anomaly Detection Using Flow Matching on Tabular Data
arXiv:2608.19801v1 Announce Type: new Abstract: Financial anomaly detection often relies on large unlabeled transaction logs, where anomalous samples may already be present during training. Such training-set contamination violates the clean-normal data assumption underlying many…
4 -
arXiv — Machine Learning research 9d ago
MileGPO: Milestone Inference with Local Evidence for Graph-Based Policy Optimization of Long-Horizon LLM Agents
arXiv:2608.19803v1 Announce Type: new Abstract: Credit assignment is challenging in long-horizon agentic reinforcement learning, where supervision often comes only from final rewards. Existing methods refine trajectory-level signals into step-level credits through step grouping…
30 -
arXiv — Machine Learning research 9d ago
Answer-Level Trust Selection for Physical Vision-Language Reasoning
arXiv:2608.19807v1 Announce Type: new Abstract: Vision-language models (VLMs) can estimate physical quantities such as duration, speed, and acceleration from visual observations, but existing benchmarks primarily assess overall model performance against annotated ground truth.…
4 -
arXiv — Machine Learning research 9d ago
FAR-DPO: Feasibility-Aware and Robust Direct Preference Optimization for Cyclic Peptide Design
arXiv:2608.19808v1 Announce Type: new Abstract: Cyclic peptides are emerging as promising molecular scaffolds in drug discovery due to their high binding affinity and structural stability. However, extending generative models from linear to cyclic peptide design remains…
8 -
arXiv — Machine Learning research 9d ago
Adaptive Probabilistic Shielding by Learning MDPs for Safe Reinforcement Learning
arXiv:2608.19836v1 Announce Type: new Abstract: Probabilistic shielding is a technique for safe reinforcement learning (RL). Typically, a static observer -- called the shield -- constrains the learning agent's actions to those for which acting safely remains feasible.…
30 -
arXiv — Machine Learning research 9d ago
Inadvertent Context Leakage in Language Models
arXiv:2608.19857v1 Announce Type: new Abstract: For AI agents to be useful beyond simple chat, they must hold sensitive user context such as calendars, credentials, health records, and financial data. We study whether the mere presence of such secrets in a model's context window…
13 -
arXiv — Machine Learning research 9d ago
Online Test-Time Adaptation for Generalizable Dynamic Graph Anomaly Detection
arXiv:2608.19858v1 Announce Type: new Abstract: Generalizable dynamic graph anomaly detection (DGAD) enables pretrained detectors to identify anomalies in unseen target domains without costly retraining. However, existing methods often fail for two reasons. First, they mainly…
26 -
-
arXiv — Machine Learning research 9d ago
Evidence Before Expansion: Reuse, Spawn, or Defer in Lifelong Expert Pools
arXiv:2608.19888v1 Announce Type: new Abstract: Streaming systems that maintain a pool of expert models must repeatedly decide whether to reuse an existing expert for arriving data, spawn a new one, or defer. We present a decision layer that makes all three outcomes…
32 -
arXiv — Machine Learning research 9d ago
Reliable Neural Collapse Approximation for Open-World Test-Time Adaptation
arXiv:2608.19890v1 Announce Type: new Abstract: Test-Time Adaptation (TTA) methods aim to bridge the domain gap between the source and target domains. However, traditional TTA methods become ineffective when the label distribution shift occurs, a challenge commonly referred to…
6 -
arXiv — Machine Learning research 9d ago
PETA:Parameter-Efficient Test-Time Adaptation for Virtual Screening
arXiv:2608.19906v1 Announce Type: new Abstract: Accurately ranking active ligands for a target protein pocket from massive chemical libraries remains a central challenge in virtual screening. DrugCLIP and its recent extensions substantially accelerate this process by encoding…
17 -
arXiv — Machine Learning research 9d ago
Multi-Source Wasserstein Distributionally Robust Graph Learning
arXiv:2608.19914v1 Announce Type: new Abstract: Network topology inference from graph signals is central to graph signal processing with applications in neuroscience, sensor, and social networks. In practice, target-domain samples are scarce while heterogeneous source-domain…
32 -
-
arXiv — Machine Learning research 9d ago
G-MARK: Grounded Multi-Agent Reasoning for Cooperative Driving via Knowledge Graphs
arXiv:2608.19964v1 Announce Type: new Abstract: Autonomous driving systems must operate under partial observability, where safety-critical objects may be occluded or visible only to neighboring connected vehicles. Vehicle-to-vehicle cooperation can reduce this uncertainty, but…
24 -
arXiv — Machine Learning research 9d ago
Green BOA: Determining the environmental break-even point for ML-based data compression
arXiv:2608.19994v1 Announce Type: new Abstract: We summarise the outcome of two summer internship projects based at the University of Manchester, focused on the break-even point in terms of environmental sustainability for ML-based data compression algorithms. Using the example…
10 -
arXiv — Machine Learning research 9d ago
Scale-Aware Pretraining of Time Series Foundation Models via Multi-Patch Token Alignment and Hybrid Masking
arXiv:2608.20005v1 Announce Type: new Abstract: Pretraining time series foundation models across heterogeneous datasets necessitates effective handling of varying sampling frequencies. Current methods either employ dataset-specific patch sizes and separate FFNs, leading to…
31 -
arXiv — Machine Learning research 9d ago
Systematic Evaluation of TabPFN-TS for Zero-Shot Probabilistic Heat Load Forecasting in District Heating Networks
arXiv:2608.20024v1 Announce Type: new Abstract: District heating energy hubs require reliable heat load forecasts for efficient operational scheduling. Conventional forecasting workflows train system-specific models on historical data, which can become burdensome when networks…
11 -
arXiv — Machine Learning research 9d ago
CLaST: Context-aware Contrastive VAE for Probabilistic Time Series Forecasting
arXiv:2608.20025v1 Announce Type: new Abstract: Probabilistic forecasting models are widely used for time series forecasting in domains such as energy systems, finance, medicine, and transportation. In recent years, deep generative models have shown strong results on…
32 -
arXiv — Machine Learning research 9d ago
An Inclusive and Lightweight Approach to Federated Continual Learning for Cultural Heritage
arXiv:2608.20038v1 Announce Type: new Abstract: Artificial intelligence can support cultural heritage and digital humanities through large-scale retrieval and analysis of digitized collections. However, cultural heritage data are often distributed across institutions,…
13