News / #paper Tag Research papers 500 articles archived under #paper · RSS Sign in to follow arXiv — Machine Learning research 2d ago ClusterAttention: A training-free speedup of bidirectional attention arXiv:2608.26965v1 Announce Type: new Abstract: This paper introduces ClusterAttention, a general training-free speedup of bidirectional attention layers. Existing sparse attention methods either rely on structure in the input, such as order in language or spatial proximity in… 13 arXiv — Machine Learning research 2d ago TEMPLAR Wales: A georeferenced environmental and toponymic dataset of Welsh settlements arXiv:2608.26970v1 Announce Type: new Abstract: Place names provide persistent records of how landscapes have been described and organised, but their quantitative reuse requires explicit separation between mapped places, lexical annotations and environmental measurements.… 36 arXiv — Machine Learning research 2d ago Terrain signatures in Welsh settlement names arXiv:2608.26978v1 Announce Type: new Abstract: Landscapes are named, but whether names retain measurable environmental information beyond broad geographic structure is rarely tested. We analysed 3,757 Welsh settlements using a frozen, source-audited 24-element lexical… 33 arXiv — Machine Learning research 2d ago Decentralized Multitask Learning over Learned Task Graphs arXiv:2608.26989v1 Announce Type: new Abstract: This paper investigates decentralized multitask learning over networks when the underlying task relationships are unknown. While existing graph-regularized multitask frameworks typically assume a known structure, practical settings… 36 arXiv — Machine Learning research 2d ago Benchmarking_Fast_Domain_Adaptation_for_Unsupervised_Speech_Units arXiv:2608.26992v1 Announce Type: new Abstract: Representation learning has attracted great atten- tion and managed to reach good performances as a pretraining method for downstream tasks or as a first step towards unsu- pervised speech modeling. Yet, little is known about how… 4 arXiv — Machine Learning research 2d ago Disentangling Optimization Scale from Preference Scale in DPO arXiv:2608.27032v1 Announce Type: new Abstract: Direct Preference Optimization (DPO) is a widely used objective for aligning language models from preference data, with the coefficient $\beta$ commonly interpreted as controlling the KL constraint to a reference policy. We show… 25 arXiv — Machine Learning research 2d ago Performance Foundations of Parallel & Distributed Reasoning Language Models arXiv:2608.27046v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) and other RL-style post-training paradigms have been used for aligning large language models (LLMs) with reasoning standards. The resulting recent Reasoning Language Models… 12 arXiv — Machine Learning research 2d ago Soft Active Electromyography Interface for Machine Learning-Enabled Silent Speech Recognition arXiv:2608.27048v1 Announce Type: new Abstract: Silent speech recognition (SSR) provides an alternative communication pathway in the absence of audible speech. However, conventional approaches are limited by the need for constant facial attachment, privacy concerns, and unstable… 33 arXiv — Machine Learning research 2d ago Unifying Detection and Adaptation in Task-Free Continual Learning arXiv:2608.27070v1 Announce Type: new Abstract: To mitigate catastrophic forgetting in downstream continual learning (CL) for large language models (LLMs), existing methods typically constrain parameter updates or introduce task-specific adaptation modules. However, these… 34 arXiv — Machine Learning research 2d ago Emotional Preferences as Goal-Priority Regulation arXiv:2608.27072v1 Announce Type: new Abstract: A core question in decision-making for agents is whether the relative priorities of competing lower-level objectives can be determined by emotional preferences autonomously generated by higher-level goals, rather than being… 26 arXiv — Machine Learning research 2d ago Tabular Deep Learning for Algorithmic Trading: Cross-Regime Bayesian Optimisation for Equity Signal Generation arXiv:2608.27076v1 Announce Type: new Abstract: Algorithmic trading now represents a market exceeding $20 billion, where even marginal gains in signal robustness can translate into economically significant returns. Existing evaluations of equity prediction models do not… 24 arXiv — Machine Learning research 2d ago Cone Extended Rayleigh Quotients for Directed Graph Learning: Minimax Spectral Certificates, Sensitivity, and Adaptive Control arXiv:2608.27122v1 Announce Type: new Abstract: Directed graph learning naturally leads to trainable nonsymmetric propagation operators with distinct right and left spectral structures. Building on the two-sided cone Rayleigh framework for generalized pencils \[ B_\theta-\lambda… 24 arXiv — Machine Learning research 2d ago TRACE-CRC: Trajectory-Adaptive Conformal Risk Control for Multi-Step Channel State Information Prediction arXiv:2608.27124v1 Announce Type: new Abstract: Reliable prediction of time-varying channel state information (CSI) is essential for efficient wireless communication. Each CSI frame is a matrix-valued representation of the wireless channel response, and a sequence of CSI frames… 5 arXiv — Machine Learning research 2d ago Ultra Low-Power, Lightweight, Probabilistic RSS-Based Path Reconstruction: A System for Landscape-Scale Bee Tracking arXiv:2608.27152v1 Announce Type: new Abstract: Applications in fields such as movement ecology, Internet of Things or robotics share the need for systems that localize devices that are too small and power constrained to implement GNSS (Global Navigation Satellite Systems).… 37 arXiv — Machine Learning research 2d ago Inductive Correlation Clustering with Graph Neural Networks arXiv:2608.27153v1 Announce Type: new Abstract: Correlation Clustering (CC) is a natural formulation of clustering in combinatorial optimization, which uses a graph representation of the input and does not require a pre-specified number of clusters. Given $n$ objects and a… 6 arXiv — Machine Learning research 2d ago Diffusion Policies for Short-Horizon Planning in Robot Crowd Navigation arXiv:2608.27158v1 Announce Type: new Abstract: Robot crowd navigation requires safe and efficient decision-making under dense, dynamic, and multimodal human--robot interactions. Existing reinforcement-learning methods typically output a single reactive action at each timestep,… 10 arXiv — Machine Learning research 2d ago TraceBench: Controlled Evaluation of LLM Agents for Time-Series Root-Cause Attribution arXiv:2608.27182v1 Announce Type: new Abstract: LLM agents are increasingly applied to anomaly detection and root-cause analysis in time-series observations collected from real-world systems; however, their performance on these tasks has not been systematically evaluated under… 10 arXiv — Machine Learning research 2d ago When Interference Graphs Evolve: Doubly Robust Estimation of Dynamic Peer Effects arXiv:2608.27187v1 Announce Type: new Abstract: Peer effects are difficult to estimate when interaction graphs evolve because pre-assignment network history, dynamic peer exposure, and post-assignment network change have distinct causal roles. We introduce a controlled contrast… 11 arXiv — Machine Learning research 2d ago Common Geodesics Do Not Guarantee Fisher Consistency of the Structured SVM: Minimal Counterexamples and a Tree-Metric Classification arXiv:2608.27203v1 Announce Type: new Abstract: A known necessary condition for Fisher consistency of the structured support vector machine requires the task loss to be a metric for which every output triple has a common geodesic point. We show that this condition is not… 29 arXiv — Machine Learning research 2d ago Profit based evaluation of machine learning for nitrogen recommendations in winter wheat arXiv:2608.27205v1 Announce Type: new Abstract: Nitrogen rates for winter wheat are set before the season, under unknown prices and weather. The standard UK advice does not respond to prices, yet recent price swings moved the most profitable rate by tens of kilograms per… 26 arXiv — Machine Learning research 2d ago HALO: A Heterogeneity-Aware Language-Aligned IMU Foundation Model for Open-Set Human Activity Recognition arXiv:2608.27233v1 Announce Type: new Abstract: Human Activity Recognition (HAR) using inertial measurement units (IMUs) enables a wide range of applications, yet the field still lacks a unified model that can generalize across diverse subjects, devices, and activities. Training… 29 arXiv — Machine Learning research 2d ago Importance Scoring of Transformer Attention Heads in Learning Tabular Data arXiv:2608.27241v1 Announce Type: new Abstract: Computationally demanding and opaque deep learning models can be better understood and optimized by analyzing how they transform data. While deep transformers have been widely studied in computer vision and natural language… 14 arXiv — Machine Learning research 2d ago Circuit Condensation: Post-Training that Concentrates a Behavior's Causal Circuit arXiv:2608.27254v1 Announce Type: new Abstract: One approach to mechanistic interpretability explains behavior through circuits: the components and connections that carry it. Frozen discovery often returns hundreds of edges, making them hard to inspect, compare, or verify… 32 arXiv — Machine Learning research 2d ago Making Latent Evolution Explicit: Operator-Structured Transitions for World Action Models arXiv:2608.27259v1 Announce Type: new Abstract: World Action Models (WAMs) augment robot policies by predicting how task-relevant scene states may evolve under interaction. Recent WAMs increasingly perform such prediction in latent representation spaces, avoiding full… 30 arXiv — Machine Learning research 2d ago MM-Spectrum: Multimodal Multi-spectral Molecular Structural Elucidation with a Stable MoE Framework arXiv:2608.27286v1 Announce Type: new Abstract: Inferring molecular structures from multimodal spectroscopic measurements requires integrating complementary yet highly heterogeneous signals. However, the common paradigm of directly concatenating multispectral sequences can… 27 arXiv — Machine Learning research 2d ago QuantumBoostNet: A Hybrid Classical-Quantum Architecture for Enhanced Accuracy in Cardiac Ultrasound View Identification arXiv:2608.27302v1 Announce Type: new Abstract: Accurate identification of the correct view or angle in cardiac ultrasound (echocardiogram) is a critical component of cardiologic imaging. This step is essential for precise anatomical interpretation, reliable measurement, and the… 13 arXiv — Machine Learning research 2d ago Beyond Parallel Blindness: Information Floors and Model Gaps in Block Drafting arXiv:2608.27339v1 Announce Type: new Abstract: Block drafters propose several tokens in one forward pass, before earlier target tokens are realised. Their rejection mixes two losses: missing within-block path information and imperfect modelling of observable information.… 21 arXiv — Machine Learning research 2d ago Understanding Evolution Strategies for LLM Reasoning: Broader Reasoning Coverage than GRPO arXiv:2608.27351v1 Announce Type: new Abstract: Evolution Strategies (ES) have recently emerged as a memory-efficient post-training paradigm for LLM reasoning. However, the optimization behavior of ES remains understudied, making it hard to define its advantage scope compared to… 14 arXiv — Machine Learning research 2d ago A Dynamic Likelihood Approach to Filtering for Advection-Diffusion Dynamics arXiv:2406.06837v2 Announce Type: cross Abstract: A Bayesian data assimilation scheme is formulated for advection-dominated advective and diffusive evolutionary problems, based upon the Dynamic Likelihood (DLF) approach to filtering. The DLF was developed specifically for… 34 arXiv — Machine Learning research 2d ago Generative Monte Carlo Sampling for Constant-Cost Particle Transport arXiv:2512.13965v1 Announce Type: cross Abstract: We present Generative Monte Carlo (GMC), a novel paradigm for particle transport simulation that integrates generative artificial intelligence directly into the stochastic solution of the linear Boltzmann equation. By… 9 arXiv — NLP / Computation & Language research 2d ago Recipes for Steering and Scaling LLMs via Sampling arXiv:2608.26120v1 Announce Type: new Abstract: Large Language Models (LLMs) are probabilistic models, typically defined by an autoregressive factorization. While recent work has begun to study richer target distributions beyond the base model, the sampling strategies remain… 4 arXiv — NLP / Computation & Language research 2d ago Can a Model Catch Its Own Hallucinations for Free?: Label-Free Doubt Signals Hold Their Own Against a Labelled Dataset for Abstention arXiv:2608.26121v1 Announce Type: new Abstract: Large language models state false facts as fluently as true ones, yet a model often "knows" internally when it is on shaky ground: the probability it assigns to its own answer tends to dip on the facts it gets wrong. The usual way… 19 arXiv — Machine Learning research 2d ago Graph-Based Modeling of Financial Volatility Dynamics arXiv:2608.26127v1 Announce Type: cross Abstract: Accurate forecasting of realized volatility ($RV$) is crucial for risk management and derivatives pricing. Although the implied volatility ($IV$) surface offers rich informational content, prevailing methods that treat it as a… 18 arXiv — NLP / Computation & Language research 2d ago FIRSTPASS: A Multi-Domain, Multi-Round Peer Review Dataset Grounded in Real Editorial Outcomes arXiv:2608.26129v1 Announce Type: new Abstract: Scientific peer review datasets have trained AI systems exclusively on Computer Science and Machine Learning venues, producing models that critique ablation studies yet have never seen a biology reviewer demand contamination… 8 arXiv — NLP / Computation & Language research 2d ago Interpretable, Fairly Evaluated Automated L2 Speaking Assessment that Beats the Single-Human Ceiling and Why Pause Encoding Does Not Change LLM Fluency Scores arXiv:2608.26137v1 Announce Type: new Abstract: Second-language (L2) English learners can rarely rehearse speaking with a partner. Speaking is also the most anxiety-laden skill. These gaps drive a fast-growing market for automated speaking practice and scoring. But an automated… 24 arXiv — NLP / Computation & Language research 2d ago Cross-Platform Generalisation Failure in Mental Health Natural Language Processing: A Five-Axis Fairness Audit of Transformer Models on Social Media arXiv:2608.26138v1 Announce Type: new Abstract: We introduce the Cross-Platform Fairness Evaluation (CPFE) framework -- a five-axis audit protocol covering discriminative performance, calibration, statistical significance, prediction equity, and attribution stability -- and… 34 arXiv — NLP / Computation & Language research 2d ago Affix Cache for Diffusion Large Language Models arXiv:2608.26140v1 Announce Type: new Abstract: Diffusion Large Language Models (DLLMs) enable non-autoregressive decoding and bidirectional context modeling, but efficient inference remains challenging. Unlike autoregressive systems, whose key-value (KV) cache can be reused for… 38 arXiv — NLP / Computation & Language research 2d ago AdaThinking-E: One-Token Entropy Regulation for Adaptive Thinking arXiv:2608.26141v1 Announce Type: new Abstract: Multimodal large language models have demonstrated strong document reasoning capabilities by incorporating explicit thinking processes. While this capability significantly improves performance on challenging tasks, current models… 13 arXiv — Machine Learning research 2d ago Methodological and Conceptual Framework for 5D Multi-Table Analysis: A Unified Approach for Complex Data Reuse arXiv:2608.26149v1 Announce Type: cross Abstract: Multi-table learning remains a major challenge in machine learning for healthcare and other complex information systems. Relational data combine several sources of complexity, including large data volume, high-dimensional… 11 arXiv — Machine Learning research 2d ago Explainable Artificial Intelligence for Customer Churn Prediction in Telecommunications: A Framework for CRM Integration arXiv:2608.26151v1 Announce Type: cross Abstract: Subscriber attrition is a costly, persistent challenge for telecommunications providers, with monthly churn of roughly 1.9% in mature markets eroding billions in revenue annually. Predictive models can flag at-risk customers… 37 arXiv — Machine Learning research 2d ago Selection Bias Correction in Retail Intelligence arXiv:2608.26156v1 Announce Type: cross Abstract: Retail intelligence often relies on monitoring popular, high-velocity products, potentially biasing economic indicators by ignoring the "long tail" of niche items. This simulation study investigates selection bias in inflation… 36 arXiv — Machine Learning research 2d ago Refusal Is Not Robustness: Auditing Confident Fabrication in Large Language Models on a Provably Uninformative Clinical Pain Speech Transcript arXiv:2608.26167v1 Announce Type: cross Abstract: Hallucination and abstention benchmarks rarely establish that a model could not have known the correct answer, making it difficult to distinguish appropriate abstention from an unsupported prediction. Seven large language models… 32 arXiv — Machine Learning research 2d ago ClassVision: AI-Powered Classroom Attendance System arXiv:2608.26173v1 Announce Type: cross Abstract: Students and working professionals have to go through the attendance process every day. Traditional methods of marking attendance using pen and paper or online platforms are human-intensive and time-consuming. To address the… 31 arXiv — NLP / Computation & Language research 2d ago When the Canonical Completion Is Wrong: Formalizing and Measuring the Jump in Large Language Models arXiv:2608.26187v1 Announce Type: new Abstract: Whether large language models (LLMs) can perform the abductive leap from evidence to a new system of axioms, commonly referred to as a jump, has recently attracted considerable debate. A prominent position holds that LLMs are… 17 arXiv — Machine Learning research 2d ago Invocation-Level Reliability of Tool-Using Agents arXiv:2608.26189v1 Announce Type: cross Abstract: Tool-using agents fail two ways: choosing the wrong tool, or forming wrong arguments, and an early failure of either kind can silently corrupt everything downstream. We measure a correct-invocation rate that separates the two,… 15 arXiv — Machine Learning research 2d ago GameWAM: A World Action Model for Video Games arXiv:2608.26200v1 Announce Type: cross Abstract: Modern video games combine first-person perception, rapid visual changes, persistent world state, and heterogeneous native controls. Existing game agents map visual and task context directly to actions but lack explicit world… 38 arXiv — Machine Learning research 2d ago Fairness Invariants: A Relational Approach to Explaining and Mitigating Fairness Bugs arXiv:2608.26209v1 Announce Type: cross Abstract: Data-driven software systems are increasingly deployed in high-stakes socio-economic domains, from criminal justice to financial lending. However, these systems often exhibit individual discrimination---unjustified disparities in… 38 arXiv — Machine Learning research 2d ago Real-time virtual circuits for plasma shape control via neural network emulators: integration and testing in the MAST-U PCS arXiv:2608.26216v1 Announce Type: cross Abstract: The deployment of advanced, AI-enabled control algorithms in tokamak experiments requires robust integration with existing plasma control system (PCS) architectures and extensive pre-experimental validation. In this contribution,… 37 arXiv — Machine Learning research 2d ago TRACE: Retrospective Streaming Generation of Physical Fields under Sparse Structured Sensing arXiv:2608.26219v1 Announce Type: cross Abstract: Reconstructing continuous physical fields from sparse measurements is central to scientific monitoring, inverse modeling, and digital-twin construction. Generative reconstruction has recently emerged as a promising paradigm for… 30 arXiv — Machine Learning research 2d ago Prompt Sensitivity of Generative Agents: Evidence from an Epidemic Model arXiv:2608.26221v1 Announce Type: cross Abstract: As generative AI gains traction, researchers are investigating its potential to serve as proxies for humans. From undergoing cognitive psychology experiments to experiencing an epidemic, generative agents, agents powered by… 5 Page 2 of 10 · 500 articles ← Newer Older →