arXiv — Machine Learning
500 articles archived · Visit source ↗ · RSS
-
arXiv — Machine Learning research 6d ago
SPARCL: Spectral Partitioned Analytic Continual Learning
arXiv:2608.21307v1 Announce Type: new Abstract: Analytic continual learning has emerged as a strong exemplar-free alternative to gradient-based class-incremental learning because it replaces iterative optimization with closed-form ridge updates. Yet the usual forgetting…
22 -
arXiv — Machine Learning research 6d ago
Rethinking Expressivity and Efficiency in Test-Time Training
arXiv:2608.21308v1 Announce Type: new Abstract: Test-Time Training (TTT) enables long-context processing via continuous weight updates during inference, but current methods struggle to balance the expressivity of per-token update dynamics with the hardware efficiency of…
19 -
arXiv — Machine Learning research 6d ago
Time-Aware Tranformer-Based Prediction Model for AECOPD
arXiv:2608.21324v1 Announce Type: new Abstract: The rapid symptom change of Acute exacerbation of chronic obstructive pulmonary disease (AECOPD) makes it critical to have time-sensitive prediction models. However, most current machine learning models studying AECOPD use clinical…
23 -
arXiv — Machine Learning research 6d ago
Across-Design Uncertainty in Short Pricing Panels: Evidence from Simulated Price Trajectories
arXiv:2608.21334v1 Announce Type: new Abstract: Short observational pricing panels can contain many observations while offering only a small number of distinct price movements. This paper studies the inferential consequences of that distinction in a synthetic data-generating…
11 -
arXiv — Machine Learning research 6d ago
Asymmetric Capacity Allocation in Self-Refinement Pipelines
arXiv:2608.21345v1 Announce Type: new Abstract: Self-refinement, typically structured as generation, critique, and revision, is a widely adopted paradigm for improving LLM generation and serves as a core mechanism in many LLM agents. While the three stages involve different…
37 -
-
-
arXiv — Machine Learning research 6d ago
TriPLU: Bypassing the Gate with Direct Trilinear Product FFNs in Tiny Language Models
arXiv:2608.20360v1 Announce Type: cross Abstract: We study whether tiny decoder-only language models benefit from feed-forward layers that directly multiply learned feature projections. TriPLU, a Trilinear Product Linear Unit, replaces the usual gated FFN branch with a…
32 -
arXiv — Machine Learning research 6d ago
Multilingual Verifier Bias in RLVR: Benchmark, Rollout Diagnosis, and the Cross-Lingual Selection Bottleneck
arXiv:2608.20362v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) is a standard recipe for training large language models on mathematical reasoning, where an answer verifier serves as a language-neutral reward function. We show that this…
33 -
arXiv — Machine Learning research 6d ago
Harmonic Torsional Diffusion for Protein-Ligand Flexible Docking
arXiv:2608.20366v1 Announce Type: cross Abstract: Molecular docking requires reasoning jointly about ligand pose and protein flexibility. Most diffusion-based docking models predict torsional updates with generic Euclidean heads that ignore the periodic geometry of angular…
18 -
arXiv — Machine Learning research 6d ago
VA-DPO: Valence-Arousal Direct Preference Optimization for Controllable Emotion Generation in Language Models
arXiv:2608.20374v1 Announce Type: cross Abstract: How precisely can we tell a language model how to feel? Most work on emotional generation answers with a discrete label - happy, angry, sad - which cannot express a target like "mildly downcast but calm." We instead specify the…
17 -
arXiv — Machine Learning research 6d ago
TH-GNN: Heterogeneous Temporal Graph Neural Networks for LLM-Agent Shilling Attack Detection
arXiv:2608.20376v1 Announce Type: cross Abstract: LLM agents can now generate realistic shilling profiles, fluent reviews, and coherent ratings at scale, systematically defeating recommender-system defenses. Text-only detectors that flag semantic drift in review embeddings are…
27 -
arXiv — Machine Learning research 6d ago
If It Walks Like an Arbitrage: Protocol-Agnostic Detection with Decidable Structural Equivalence
arXiv:2608.20377v1 Announce Type: cross Abstract: Ethereum transactions admit a canonical structural form. Each execution trace is built into an abstract syntax tree of token transfers grouped by call-frame nesting and reduced by a convergent term rewriting system of 15 rules to…
15 -
arXiv — Machine Learning research 6d ago
Interpretable Information-Decomposed Brain Graph Learning for fMRI-based Disease Diagnosis
arXiv:2608.20380v1 Announce Type: cross Abstract: Resting-state functional magnetic resonance imaging (rs-fMRI) has enabled non-invasive mapping of functional brain interactions for computer-aided diagnosis, yet most existing approaches reduce inter-regional relationships to…
35 -
-
arXiv — Machine Learning research 6d ago
World models of environment, agent and joint agent-environment systems
arXiv:2608.20401v1 Announce Type: cross Abstract: World models are a central component of model-based reinforcement learning. They are usually discussed in terms of what variables they predict, such as observations, rewards, states, latent or information states. We argue that…
12 -
arXiv — Machine Learning research 6d ago
Robust Discovery of Coarse-Grained Continuum Equations from Microscopic Dynamics
arXiv:2608.20404v1 Announce Type: cross Abstract: The discovery of governing partial differential equations (PDEs) directly from spatiotemporal data has emerged as a powerful tool for understanding the dynamics of complex systems. In this work, we apply PDE-SINDy to well-known…
30 -
-
arXiv — Machine Learning research 6d ago
Uncertainty propagation in auto-regressive random neural network models
arXiv:2608.20483v1 Announce Type: cross Abstract: We develop analytical and particle-based methods for uncertainty propagation in random neural network models, where both the inputs and network parameters are allowed to be random. Building on the piecewise-linear structure of…
4 -
-
arXiv — Machine Learning research 6d ago
aiXamine: Unified Black-Box Evaluation of Cross-Dimensional Trade-offs in LLM Safety, Security, and Privacy
arXiv:2608.20554v1 Announce Type: cross Abstract: The critical failure modes in deployed large language models (LLMs) are cross-dimensional: a model can score 99.3 in safety alignment while refusing one in three benign queries, or improve across every capability metric while…
29 -
arXiv — Machine Learning research 6d ago
Learning Prostate Anatomy at Test Time for Cancer Detection in Micro-Ultrasound
arXiv:2608.20557v1 Announce Type: cross Abstract: Domain shift across clinical centers using different imaging hardware or acquisition protocols remains a fundamental barrier to deploying deep learning models for prostate cancer (PCa) detection. Existing test-time adaptation…
14 -
arXiv — Machine Learning research 6d ago
Consistency Models for Fast MRI Reconstruction Using Regularization by Denoising
arXiv:2608.20561v1 Announce Type: cross Abstract: Diffusion models (DMs) have emerged as powerful generative priors for MRI reconstruction with promising results. Yet DM-based methods require extensive iterative refinement, limiting their practical deployment. Consistency models…
4 -
arXiv — Machine Learning research 6d ago
Conditional-Independence-Regularized Distributional Autoencoders for Mixed-Type Data
arXiv:2608.20562v1 Announce Type: cross Abstract: Mixed-type data containing both numerical and categorical variables arise in many scientific and real-world applications. Existing representation learning and generative modeling approaches typically focus either on…
30 -
arXiv — Machine Learning research 6d ago
FlavourBench: Ranking Frontier Language Models with Executable Culinary Ground Truth
arXiv:2608.20574v1 Announce Type: cross Abstract: Open-ended language-model benchmarks usually inherit a judge: a human preference panel, another model, or a brittle exact-match key. We introduce FlavourBench, an automated benchmark in which a versioned culinary system supplies…
35 -
arXiv — Machine Learning research 6d ago
Keyed Provenance Watermarking with Complementary Lattice-Based Secure Aggregation for Federated Learning
arXiv:2608.20580v1 Announce Type: cross Abstract: Federated learning (FL) is vulnerable to multi-level attacks. However, existing methods address them separately, leaving FL exposed to data leakage, unauthorized reuse, and malicious gradient manipulation. In this work, we…
26 -
-
arXiv — Machine Learning research 6d ago
Dual-Cache Latent Space Communication between Heterogeneous Language Models
arXiv:2608.20617v1 Announce Type: cross Abstract: Multi-agent LLM systems split work across models, so answering often requires knowledge that sits in another agent's context: a Sharer has encoded information that a Receiver needs to complete its task. They usually communicate…
5 -
arXiv — Machine Learning research 6d ago
Minimax Optimality of Score-Entropy Discrete Diffusion
arXiv:2608.20635v1 Announce Type: cross Abstract: Discrete diffusion models have demonstrated strong performance across a range of datasets, including natural language data and graph-structured data. Among many variants, score-entropy discrete diffusion (SEDD) has achieved…
13 -
arXiv — Machine Learning research 6d ago
MIL-BERT: Classification of Arbitrarily Large Text with Performance and Explanatory Guarantees
arXiv:2608.20636v1 Announce Type: cross Abstract: Many text classification decisions are viable based on constituent excerpts alone. Taking inspiration from the field of multiple instance learning, we present an algorithm for training a neural network to classify text by…
10 -
arXiv — Machine Learning research 6d ago
Predicting Resource Efficient Hamiltonian Decomposition for Continuous-Time Quantum Walk Simulations
arXiv:2608.20660v1 Announce Type: cross Abstract: Simulating a continuous-time quantum walk (CTQW) on a graph in the circuit model of quantum computing requires decomposing its Hamiltonian into terms that can be Trotterized into hardware-native gates. We consider two such…
15 -
arXiv — Machine Learning research 6d ago
Amplifying the imaging power of digital sky surveys with space telescopes data and generative AI
arXiv:2608.20666v1 Announce Type: cross Abstract: While Digital sky surveys provide excellent throughput of image data and can cover a large footprint, their imaging power is normally inferior to that of space-based telescopes. Space-based telescopes, on the other hand, provide…
31 -
arXiv — Machine Learning research 6d ago
Temporal Validity on Real Software Histories: Eliminating Stale-Fact Errors in Code-Assistant Memory over GitHub Fixes
arXiv:2608.20685v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) has no model of time: when a fact changes across a coding session - a function is renamed, an endpoint moves, a dependency is bumped - RAG retrieves both the old and new value with…
17 -
arXiv — Machine Learning research 6d ago
CDRL: Certification-Driven Reinforcement Learning for Neutrino Flavor Model Discovery
arXiv:2608.20686v1 Announce Type: cross Abstract: Many scientific discovery problems require searching combinatorial hypothesis spaces under complex domain constraints. Reinforcement learning (RL) offers a promising approach, but existing methods rely on scalar rewards that…
26 -
arXiv — Machine Learning research 6d ago
PSK at WMT 2026 MIST: Task-Specialized QLoRA Adapters for Multilingual Summarization and Question Answering
arXiv:2608.20757v1 Announce Type: cross Abstract: We describe the PSK submission to the WMT 2026 Multilingual Instruction Shared Task. Our system uses the 3.35B-parameter Tiny Aya Global model with three QLoRA adapters, one for each task. The adapters are trained on multilingual…
26 -
arXiv — Machine Learning research 6d ago
Rethinking Demonstration Unlearning in Imitation Learning for Robotics
arXiv:2608.20784v1 Announce Type: cross Abstract: Imitation learning for robotics depends on human demonstrations, some of which people may later ask to remove. Retraining without them is the natural reference, but its cost grows with policy and dataset scale, motivating cheaper…
12 -
arXiv — Machine Learning research 6d ago
CubicSplat: Differentiable Vector Graphics via Error-Bounded Forward Relaxation
arXiv:2608.20803v1 Announce Type: cross Abstract: Vector graphics are prized for their resolution independence, compact storage, and direct editability, making differentiable optimization of their parametric primitives an attractive goal. Yet classical rasterization is…
31 -
arXiv — Machine Learning research 6d ago
Neuro-Geospatial Modelling of EEG Affective States Using Literature-Informed Environmental Context
arXiv:2608.20807v1 Announce Type: cross Abstract: Environmental exposures such as air pollution and greenness have been associated with affective and cognitive outcomes, but EEG and environmental datasets are rarely jointly georeferenced. We investigate whether…
30 -
arXiv — Machine Learning research 6d ago
Fine-tuning LLMs for Tourist Trajectory Prediction using Field Experiment Data
arXiv:2608.20830v1 Announce Type: cross Abstract: Evaluating mobility interventions at tourist destinations requires predicting visitor behavior under varying conditions. Traditional methods struggle because tourist decisions depend heavily on context like weather and fatigue,…
18 -
arXiv — Machine Learning research 6d ago
SAC-Copula: Quality-Preserving Watermarking for Diffusion Language Models via Smooth Correlated Gumbel Fields
arXiv:2608.20839v1 Announce Type: cross Abstract: Watermarking diffusion language models (DLMs) requires mechanisms compatible with iterative parallel unmasking rather than autoregressive decoding. Existing sampling-based watermarking methods typically inject position-wise…
14 -
-
arXiv — Machine Learning research 6d ago
ReCurveflow: A Flow Matching Framework that Learns Curved Reaction Trajectories to Predict Transition State Geometries
arXiv:2608.20869v1 Announce Type: cross Abstract: Predicting transition states (TS) in chemical reactions is crucial, as they provide insights into reaction mechanisms. Recent work on TS prediction have focused on flow matching supervised on straight linear paths that do not…
20 -
arXiv — Machine Learning research 6d ago
EviRank: Structured Relevance Evidence for Multimodal Image Re-ranking
arXiv:2608.20886v1 Announce Type: cross Abstract: Real-world image search queries are multimodal and compositional: ``find this shirt in pink'' specifies an entity to retain, an attribute to modify, and context to ignore. Yet existing re-rankers either compress such multifaceted…
21 -
arXiv — Machine Learning research 6d ago
Quantization-Aware Healing: A Practical Recipe for Recovering Compressed, 4-Bit LLMs
arXiv:2608.20953v1 Announce Type: cross Abstract: Serving large language models cheaply increasingly means shipping models that are both structurally compressed to a fraction of their parameters and quantized to 4 bits. Together these steps degrade reasoning, mathematics,…
5 -
arXiv — Machine Learning research 6d ago
TreeWY: Speculative Verification for Gated DeltaNet Hybrids
arXiv:2608.20961v1 Announce Type: cross Abstract: Modern open models are hybrids: most layers are linear-attention (Gated DeltaNet, GDN) layers carrying a small fixed-size recurrent state instead of a growing key-value (KV) cache. This makes ordinary decoding memory-efficient,…
33 -
arXiv — Machine Learning research 6d ago
Training DeepFilterNet with Accurate Room Acoustic Simulations Improves Single-Channel Speech Enhancement
arXiv:2608.20971v1 Announce Type: cross Abstract: We investigate how the realism of synthetic room impulse response (RIR) datasets affects the training of DeepFilterNet3 for single-channel speech enhancement. We compare a DNS4 image-source-method (ISM) RIR dataset with a…
6 -
-
arXiv — Machine Learning research 6d ago
COMET: Contrastive Motion-Enhanced Temporal Reasoning for Video Multimodal Large Language Models
arXiv:2608.21030v1 Announce Type: cross Abstract: Video multimodal large language models have advanced significantly, yet fine-grained motion-temporal understanding remains fragile. The core bottleneck is not only sparse frame sampling, but also the lack of a complete temporal…
37 -
arXiv — Machine Learning research 6d ago
AudioWorldSim: Realistic Binaural Audio Datasets For World Models
arXiv:2608.21075v1 Announce Type: cross Abstract: This technical report presents AudioWorldSim, an open-source platform designed to generate realistic binaural audio datasets and advance research in audio-based machine learning, particularly world models. Built as a custom…
37 -
arXiv — Machine Learning research 6d ago
Llama-Mobile: Efficient 2.7-Bit Quantization of VLMs
arXiv:2608.21134v1 Announce Type: cross Abstract: Deploying vision-language models (VLMs) on mobile devices is challenging due to their significant memory and compute requirements. We present a framework for quantizing VLMs for efficient inference on resource-constrained…
18