arXiv — Machine Learning
500 articles archived · Visit source ↗ · RSS
-
arXiv — Machine Learning research 2d ago
TEMPLAR Wales: A georeferenced environmental and toponymic dataset of Welsh settlements
arXiv:2608.26970v1 Announce Type: new Abstract: Place names provide persistent records of how landscapes have been described and organised, but their quantitative reuse requires explicit separation between mapped places, lexical annotations and environmental measurements.…
36 -
arXiv — Machine Learning research 2d ago
Terrain signatures in Welsh settlement names
arXiv:2608.26978v1 Announce Type: new Abstract: Landscapes are named, but whether names retain measurable environmental information beyond broad geographic structure is rarely tested. We analysed 3,757 Welsh settlements using a frozen, source-audited 24-element lexical…
33 -
arXiv — Machine Learning research 2d ago
Decentralized Multitask Learning over Learned Task Graphs
arXiv:2608.26989v1 Announce Type: new Abstract: This paper investigates decentralized multitask learning over networks when the underlying task relationships are unknown. While existing graph-regularized multitask frameworks typically assume a known structure, practical settings…
36 -
arXiv — Machine Learning research 2d ago
Benchmarking_Fast_Domain_Adaptation_for_Unsupervised_Speech_Units
arXiv:2608.26992v1 Announce Type: new Abstract: Representation learning has attracted great atten- tion and managed to reach good performances as a pretraining method for downstream tasks or as a first step towards unsu- pervised speech modeling. Yet, little is known about how…
4 -
arXiv — Machine Learning research 2d ago
Disentangling Optimization Scale from Preference Scale in DPO
arXiv:2608.27032v1 Announce Type: new Abstract: Direct Preference Optimization (DPO) is a widely used objective for aligning language models from preference data, with the coefficient $\beta$ commonly interpreted as controlling the KL constraint to a reference policy. We show…
25 -
arXiv — Machine Learning research 2d ago
Performance Foundations of Parallel & Distributed Reasoning Language Models
arXiv:2608.27046v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) and other RL-style post-training paradigms have been used for aligning large language models (LLMs) with reasoning standards. The resulting recent Reasoning Language Models…
12 -
arXiv — Machine Learning research 2d ago
Soft Active Electromyography Interface for Machine Learning-Enabled Silent Speech Recognition
arXiv:2608.27048v1 Announce Type: new Abstract: Silent speech recognition (SSR) provides an alternative communication pathway in the absence of audible speech. However, conventional approaches are limited by the need for constant facial attachment, privacy concerns, and unstable…
33 -
arXiv — Machine Learning research 2d ago
Unifying Detection and Adaptation in Task-Free Continual Learning
arXiv:2608.27070v1 Announce Type: new Abstract: To mitigate catastrophic forgetting in downstream continual learning (CL) for large language models (LLMs), existing methods typically constrain parameter updates or introduce task-specific adaptation modules. However, these…
34 -
arXiv — Machine Learning research 2d ago
Emotional Preferences as Goal-Priority Regulation
arXiv:2608.27072v1 Announce Type: new Abstract: A core question in decision-making for agents is whether the relative priorities of competing lower-level objectives can be determined by emotional preferences autonomously generated by higher-level goals, rather than being…
26 -
arXiv — Machine Learning research 2d ago
Tabular Deep Learning for Algorithmic Trading: Cross-Regime Bayesian Optimisation for Equity Signal Generation
arXiv:2608.27076v1 Announce Type: new Abstract: Algorithmic trading now represents a market exceeding $20 billion, where even marginal gains in signal robustness can translate into economically significant returns. Existing evaluations of equity prediction models do not…
24 -
-
arXiv — Machine Learning research 2d ago
TRACE-CRC: Trajectory-Adaptive Conformal Risk Control for Multi-Step Channel State Information Prediction
arXiv:2608.27124v1 Announce Type: new Abstract: Reliable prediction of time-varying channel state information (CSI) is essential for efficient wireless communication. Each CSI frame is a matrix-valued representation of the wireless channel response, and a sequence of CSI frames…
5 -
arXiv — Machine Learning research 2d ago
Ultra Low-Power, Lightweight, Probabilistic RSS-Based Path Reconstruction: A System for Landscape-Scale Bee Tracking
arXiv:2608.27152v1 Announce Type: new Abstract: Applications in fields such as movement ecology, Internet of Things or robotics share the need for systems that localize devices that are too small and power constrained to implement GNSS (Global Navigation Satellite Systems).…
37 -
arXiv — Machine Learning research 2d ago
Inductive Correlation Clustering with Graph Neural Networks
arXiv:2608.27153v1 Announce Type: new Abstract: Correlation Clustering (CC) is a natural formulation of clustering in combinatorial optimization, which uses a graph representation of the input and does not require a pre-specified number of clusters. Given $n$ objects and a…
6 -
arXiv — Machine Learning research 2d ago
Diffusion Policies for Short-Horizon Planning in Robot Crowd Navigation
arXiv:2608.27158v1 Announce Type: new Abstract: Robot crowd navigation requires safe and efficient decision-making under dense, dynamic, and multimodal human--robot interactions. Existing reinforcement-learning methods typically output a single reactive action at each timestep,…
10 -
arXiv — Machine Learning research 2d ago
TraceBench: Controlled Evaluation of LLM Agents for Time-Series Root-Cause Attribution
arXiv:2608.27182v1 Announce Type: new Abstract: LLM agents are increasingly applied to anomaly detection and root-cause analysis in time-series observations collected from real-world systems; however, their performance on these tasks has not been systematically evaluated under…
10 -
arXiv — Machine Learning research 2d ago
When Interference Graphs Evolve: Doubly Robust Estimation of Dynamic Peer Effects
arXiv:2608.27187v1 Announce Type: new Abstract: Peer effects are difficult to estimate when interaction graphs evolve because pre-assignment network history, dynamic peer exposure, and post-assignment network change have distinct causal roles. We introduce a controlled contrast…
11 -
-
arXiv — Machine Learning research 2d ago
Profit based evaluation of machine learning for nitrogen recommendations in winter wheat
arXiv:2608.27205v1 Announce Type: new Abstract: Nitrogen rates for winter wheat are set before the season, under unknown prices and weather. The standard UK advice does not respond to prices, yet recent price swings moved the most profitable rate by tens of kilograms per…
26 -
arXiv — Machine Learning research 2d ago
HALO: A Heterogeneity-Aware Language-Aligned IMU Foundation Model for Open-Set Human Activity Recognition
arXiv:2608.27233v1 Announce Type: new Abstract: Human Activity Recognition (HAR) using inertial measurement units (IMUs) enables a wide range of applications, yet the field still lacks a unified model that can generalize across diverse subjects, devices, and activities. Training…
29 -
arXiv — Machine Learning research 2d ago
Importance Scoring of Transformer Attention Heads in Learning Tabular Data
arXiv:2608.27241v1 Announce Type: new Abstract: Computationally demanding and opaque deep learning models can be better understood and optimized by analyzing how they transform data. While deep transformers have been widely studied in computer vision and natural language…
14 -
arXiv — Machine Learning research 2d ago
Circuit Condensation: Post-Training that Concentrates a Behavior's Causal Circuit
arXiv:2608.27254v1 Announce Type: new Abstract: One approach to mechanistic interpretability explains behavior through circuits: the components and connections that carry it. Frozen discovery often returns hundreds of edges, making them hard to inspect, compare, or verify…
32 -
arXiv — Machine Learning research 2d ago
Making Latent Evolution Explicit: Operator-Structured Transitions for World Action Models
arXiv:2608.27259v1 Announce Type: new Abstract: World Action Models (WAMs) augment robot policies by predicting how task-relevant scene states may evolve under interaction. Recent WAMs increasingly perform such prediction in latent representation spaces, avoiding full…
30 -
arXiv — Machine Learning research 2d ago
MM-Spectrum: Multimodal Multi-spectral Molecular Structural Elucidation with a Stable MoE Framework
arXiv:2608.27286v1 Announce Type: new Abstract: Inferring molecular structures from multimodal spectroscopic measurements requires integrating complementary yet highly heterogeneous signals. However, the common paradigm of directly concatenating multispectral sequences can…
27 -
-
arXiv — Machine Learning research 2d ago
Beyond Parallel Blindness: Information Floors and Model Gaps in Block Drafting
arXiv:2608.27339v1 Announce Type: new Abstract: Block drafters propose several tokens in one forward pass, before earlier target tokens are realised. Their rejection mixes two losses: missing within-block path information and imperfect modelling of observable information.…
21 -
arXiv — Machine Learning research 2d ago
Understanding Evolution Strategies for LLM Reasoning: Broader Reasoning Coverage than GRPO
arXiv:2608.27351v1 Announce Type: new Abstract: Evolution Strategies (ES) have recently emerged as a memory-efficient post-training paradigm for LLM reasoning. However, the optimization behavior of ES remains understudied, making it hard to define its advantage scope compared to…
14 -
arXiv — Machine Learning research 2d ago
A Dynamic Likelihood Approach to Filtering for Advection-Diffusion Dynamics
arXiv:2406.06837v2 Announce Type: cross Abstract: A Bayesian data assimilation scheme is formulated for advection-dominated advective and diffusive evolutionary problems, based upon the Dynamic Likelihood (DLF) approach to filtering. The DLF was developed specifically for…
34 -
arXiv — Machine Learning research 2d ago
Generative Monte Carlo Sampling for Constant-Cost Particle Transport
arXiv:2512.13965v1 Announce Type: cross Abstract: We present Generative Monte Carlo (GMC), a novel paradigm for particle transport simulation that integrates generative artificial intelligence directly into the stochastic solution of the linear Boltzmann equation. By…
9 -
arXiv — Machine Learning research 2d ago
Recipes for Steering and Scaling LLMs via Sampling
arXiv:2608.26120v1 Announce Type: cross Abstract: Large Language Models (LLMs) are probabilistic models, typically defined by an autoregressive factorization. While recent work has begun to study richer target distributions beyond the base model, the sampling strategies remain…
4 -
-
arXiv — Machine Learning research 2d ago
Graph-Based Modeling of Financial Volatility Dynamics
arXiv:2608.26127v1 Announce Type: cross Abstract: Accurate forecasting of realized volatility ($RV$) is crucial for risk management and derivatives pricing. Although the implied volatility ($IV$) surface offers rich informational content, prevailing methods that treat it as a…
18 -
arXiv — Machine Learning research 2d ago
FIRSTPASS: A Multi-Domain, Multi-Round Peer Review Dataset Grounded in Real Editorial Outcomes
arXiv:2608.26129v1 Announce Type: cross Abstract: Scientific peer review datasets have trained AI systems exclusively on Computer Science and Machine Learning venues, producing models that critique ablation studies yet have never seen a biology reviewer demand contamination…
8 -
-
-
arXiv — Machine Learning research 2d ago
Affix Cache for Diffusion Large Language Models
arXiv:2608.26140v1 Announce Type: cross Abstract: Diffusion Large Language Models (DLLMs) enable non-autoregressive decoding and bidirectional context modeling, but efficient inference remains challenging. Unlike autoregressive systems, whose key-value (KV) cache can be reused…
38 -
arXiv — Machine Learning research 2d ago
AdaThinking-E: One-Token Entropy Regulation for Adaptive Thinking
arXiv:2608.26141v1 Announce Type: cross Abstract: Multimodal large language models have demonstrated strong document reasoning capabilities by incorporating explicit thinking processes. While this capability significantly improves performance on challenging tasks, current models…
13 -
arXiv — Machine Learning research 2d ago
Methodological and Conceptual Framework for 5D Multi-Table Analysis: A Unified Approach for Complex Data Reuse
arXiv:2608.26149v1 Announce Type: cross Abstract: Multi-table learning remains a major challenge in machine learning for healthcare and other complex information systems. Relational data combine several sources of complexity, including large data volume, high-dimensional…
11 -
arXiv — Machine Learning research 2d ago
Explainable Artificial Intelligence for Customer Churn Prediction in Telecommunications: A Framework for CRM Integration
arXiv:2608.26151v1 Announce Type: cross Abstract: Subscriber attrition is a costly, persistent challenge for telecommunications providers, with monthly churn of roughly 1.9% in mature markets eroding billions in revenue annually. Predictive models can flag at-risk customers…
37 -
arXiv — Machine Learning research 2d ago
Selection Bias Correction in Retail Intelligence
arXiv:2608.26156v1 Announce Type: cross Abstract: Retail intelligence often relies on monitoring popular, high-velocity products, potentially biasing economic indicators by ignoring the "long tail" of niche items. This simulation study investigates selection bias in inflation…
36 -
-
arXiv — Machine Learning research 2d ago
ClassVision: AI-Powered Classroom Attendance System
arXiv:2608.26173v1 Announce Type: cross Abstract: Students and working professionals have to go through the attendance process every day. Traditional methods of marking attendance using pen and paper or online platforms are human-intensive and time-consuming. To address the…
31 -
arXiv — Machine Learning research 2d ago
When the Canonical Completion Is Wrong: Formalizing and Measuring the Jump in Large Language Models
arXiv:2608.26187v1 Announce Type: cross Abstract: Whether large language models (LLMs) can perform the abductive leap from evidence to a new system of axioms, commonly referred to as a jump, has recently attracted considerable debate. A prominent position holds that LLMs are…
17 -
arXiv — Machine Learning research 2d ago
Invocation-Level Reliability of Tool-Using Agents
arXiv:2608.26189v1 Announce Type: cross Abstract: Tool-using agents fail two ways: choosing the wrong tool, or forming wrong arguments, and an early failure of either kind can silently corrupt everything downstream. We measure a correct-invocation rate that separates the two,…
15 -
arXiv — Machine Learning research 2d ago
GameWAM: A World Action Model for Video Games
arXiv:2608.26200v1 Announce Type: cross Abstract: Modern video games combine first-person perception, rapid visual changes, persistent world state, and heterogeneous native controls. Existing game agents map visual and task context directly to actions but lack explicit world…
38 -
arXiv — Machine Learning research 2d ago
Fairness Invariants: A Relational Approach to Explaining and Mitigating Fairness Bugs
arXiv:2608.26209v1 Announce Type: cross Abstract: Data-driven software systems are increasingly deployed in high-stakes socio-economic domains, from criminal justice to financial lending. However, these systems often exhibit individual discrimination---unjustified disparities in…
38 -
-
arXiv — Machine Learning research 2d ago
TRACE: Retrospective Streaming Generation of Physical Fields under Sparse Structured Sensing
arXiv:2608.26219v1 Announce Type: cross Abstract: Reconstructing continuous physical fields from sparse measurements is central to scientific monitoring, inverse modeling, and digital-twin construction. Generative reconstruction has recently emerged as a promising paradigm for…
30 -
arXiv — Machine Learning research 2d ago
Prompt Sensitivity of Generative Agents: Evidence from an Epidemic Model
arXiv:2608.26221v1 Announce Type: cross Abstract: As generative AI gains traction, researchers are investigating its potential to serve as proxies for humans. From undergoing cognitive psychology experiments to experiencing an epidemic, generative agents, agents powered by…
5 -