arXiv · Machine Learning · 13 Aug 2026 · paper
Anchor-based pointwise LLM reranking scores each candidate against a shared reference passage to recover cross-document context at pointwise cost. We study when this actually helps, using G…
arXiv · Machine Learning · 13 Aug 2026 · paper
Harness plays a critical role in large language model agent performance, and building a high-performing harness requires substantial expert effort. Therefore, recent research has increasing…
arXiv · Machine Learning · 13 Aug 2026 · paper
Image-source-method (ISM)-based room impulse response (RIR) simulation is a useful and physically interpretable tool for acoustic scene modeling, but full-order ISM becomes computationally…
arXiv · Machine Learning · 13 Aug 2026 · paper
Although mixture-of-experts, MoE, models have been increasingly adopted to scale large language models with moderate computation cost, it remains challenging to deploy MoE inference over re…
arXiv · Machine Learning · 13 Aug 2026 · paper
Large kernel depthwise convolutions achieve strong performance but suffer from significant degradation as kernel size grows due to irregular memory access from gather-based computation; whi…
arXiv · Machine Learning · 13 Aug 2026 · paper
Most repository-level code translation and validation techniques have been evaluated on a single source-target programming language (PL) pair, owing to the complex engineering effort requir…
arXiv · Machine Learning · 13 Aug 2026 · paper
This work establishes a rigorous bridge between infinite-dimensional delay dynamics and finite-dimensional Koopman learning, with explicit and interpretable error guarantees. While Koopman…
arXiv · Machine Learning · 13 Aug 2026 · paper
Multi-agent reinforcement learning (MARL) provides a promising solution for cooperative target tracking in networks of autonomous underwater vehicles (AUVs). However, existing methods still…
arXiv · Machine Learning · 13 Aug 2026 · paper
Constrained optimization in high-dimensional black-box settings is difficult due to expensive evaluations, the lack of gradient information, and complex feasibility regions. In this work, w…
arXiv · Machine Learning · 13 Aug 2026 · paper
Graph Neural Networks (GNNs) deliver strong performance on graph tasks, but their accuracy drops significantly under out-of-distribution (OOD) scenarios. Under distribution shifts, GNNs oft…
arXiv · Machine Learning · 13 Aug 2026 · paper
Large-scale homogeneous detectors with optical readouts are widely used in particle detection, with Cherenkov and scintillator neutrino detectors as prominent examples. Analyses in experime…
arXiv · Machine Learning · 13 Aug 2026 · paper
The classical kernel ridge regression problem aims to find the best fit for the output $Y$ as a function of the input data $X\in \mathbb{R}^d$, with a fixed choice of regularization term im…
arXiv · Machine Learning · 13 Aug 2026 · paper
We study Heterogeneous Transfer Learning (HTL) for high-dimensional regression with differing feature sets. Such feature mismatch arises when some variables available in a data-rich source…
arXiv · Machine Learning · 13 Aug 2026 · paper
In clinical applications, neural networks must focus on and highlight the most important parts of an input image. Soft-Attention mechanism enables a neural network toachieve this goal. This…
arXiv · Machine Learning · 13 Aug 2026 · paper
What if judges already behave like algorithms? As artificial intelligence and algorithms are deployed in many settings, including the judicial system, many have debated whether judges shoul…
arXiv · Machine Learning · 13 Aug 2026 · paper
Low-precision formats usually optimize scalar fidelity while inheriting conventional product arithmetic. We introduce CurveFP, a block-scaled family that distributes magnitudes across inter…
arXiv · Machine Learning · 13 Aug 2026 · paper
Estimating counterfactual outcomes over time from longitudinal observational data is central to clinical decision support. Existing methods rely on domain confusion -- adversarial training…
arXiv · Machine Learning · 13 Aug 2026 · paper
Nonlinear state estimation requires sequentially fusing model-based predictions with noisy measurements. Under imperfect dynamics and unknown, time-varying noise statistics, this fusion can…
arXiv · Machine Learning · 13 Aug 2026 · paper
Credit risk detection, particularly mitigating individual fraud, is crucial for maintaining the stability of digital financial ecosystems. Accurately identifying credit fraud among billions…
arXiv · Machine Learning · 13 Aug 2026 · paper
Bayesian optimization (BO) has become the standard tool for sample-efficient optimization and owes its efficiency to uncertainty-aware search driven by generic statistical priors. Richer do…
arXiv · Machine Learning · 13 Aug 2026 · paper
The Forward-Forward algorithm trains each layer locally, so that a scalar goodness - the sum of squared activations - is high on real inputs and low on contrastive ones. Under an explicit g…
arXiv · Machine Learning · 13 Aug 2026 · paper
We introduce Quantum Port-Hamiltonian Neural Networks (Q-pHNNs), parameterised quantum circuits that learn classical dynamics in a structure-preserving manner. The framework rests on the Is…
arXiv · Machine Learning · 13 Aug 2026 · paper
Learned representations are central to modern machine learning and are typically evaluated through predictive performance, robustness, uncertainty estimation, and generalization. However, a…
arXiv · Machine Learning · 13 Aug 2026 · paper
Representation learning is central to modern machine learning, yet most research focuses on optimizing representations after a representational framework has been selected. Less attention i…
arXiv · Machine Learning · 13 Aug 2026 · paper
Generating objects with specific symmetries is essential in various real-world scenarios. However, adapting existing 2D continuous representations to enforce planar group symmetry remains a…
arXiv · Machine Learning · 13 Aug 2026 · paper
Graph anomaly detection methods aim to distinguish anomalous nodes. While prior methods characterize anomalies through increased variation in the spectral energy distributions, they overloo…
arXiv · Machine Learning · 13 Aug 2026 · paper
A fundamental dichotomy in the theory of classification sets smoothness against statistical efficiency: smooth surrogate losses such as the logistic loss enable fast $O(1/T)$ optimization b…
arXiv · Machine Learning · 13 Aug 2026 · paper
Aligning Large Language Models (LLMs) with human intent, whether through explicit reward modeling or direct methods such as DPO, fundamentally relies on minimizing a surrogate loss as a pro…
arXiv · Machine Learning · 13 Aug 2026 · paper
Learning algorithms can be significantly improved by routing complex or uncertain inputs to specialized experts, balancing accuracy with computational cost. This approach, known as learning…
arXiv · Machine Learning · 13 Aug 2026 · paper
Scaling laws are used to plan multi-million-dollar training runs, but fitting those laws can itself cost millions. In modern large-scale workflows, assembling a sufficiently informative set…
arXiv · Machine Learning · 13 Aug 2026 · paper
Despite deep learning models running well-defined mathematical functions, we lack a formal mathematical framework for describing model architectures. Ad-hoc notation, diagrams, and pseudoco…
arXiv · Machine Learning · 13 Aug 2026 · paper
Physics-Informed Neural Networks (PINNs) for high-dimensional and high-order partial differential equations (PDEs) are primarily constrained by the $\mathcal{O}(d^k)$ spatial derivative com…
arXiv · Machine Learning · 13 Aug 2026 · paper
Training large-scale generative models is resource-intensive and relies heavily on heuristic dataset weighting. We address two fundamental questions: Can we train Large Language Models (LLM…
arXiv · Machine Learning · 13 Aug 2026 · paper
We propose a federated learning framework for the calibration of parametric insurance indices under heterogeneous renewable energy production losses. Producers locally model their losses us…
arXiv · Machine Learning · 13 Aug 2026 · paper
Designing generalizable control policies that operate reliably under changing conditions is essential for robust network services in modern digital infrastructure. Yet network control remai…
arXiv · Machine Learning · 13 Aug 2026 · paper
Time series forecasting drives operational decisions under tight latency budgets, and autoregressive time series foundation models (TSFMs) increasingly deliver the most accurate forecasts.…
arXiv · Machine Learning · 13 Aug 2026 · paper
Low-Rank Adaptation (LoRA) has become a popular technique for parameter-efficient fine-tuning of large language models (LLMs). In many real-world scenarios, multiple adapters are loaded sim…
arXiv · Machine Learning · 13 Aug 2026 · paper
Edge intelligence enables low-latency inference via compact on-device models, but assuring reliability remains challenging. We study edge-cloud cascades that must preserve conditional cover…
arXiv · Machine Learning · 13 Aug 2026 · paper
Dementia disorders such as Alzheimer's disease (AD) and frontotemporal dementia (FTD) exhibit overlapping electrophysiological signatures in electroencephalography (EEG) that challenge accu…
arXiv · Machine Learning · 13 Aug 2026 · paper
Recently reconstruction-based deep models have been widely used for time series anomaly detection, but as their capacity and generalization capability increase, these models tend to over-ge…
arXiv · Machine Learning · 13 Aug 2026 · paper
This article presents an adaptive mean shift algorithm in which every parameter used at a point is derived from that point's own distance distribution. The distance distribution from a poin…
arXiv · Machine Learning · 13 Aug 2026 · paper
Many sequential decision problems offer qualitatively different ways of influencing the environment: some interventions act immediately, whereas others induce persistent effects that contin…
arXiv · Machine Learning · 13 Aug 2026 · paper
We propose a new ternary spiking neuron model to improve the representation capacity of binary spiking neurons in deep Q-learning. Although a ternary neuron model has recently been introduc…
arXiv · Machine Learning · 13 Aug 2026 · paper
Solving systems of polynomial equations, particularly those with finitely many solutions, is a crucial challenge across many scientific fields. Traditional methods like Gr\"obner and Border…
arXiv · Machine Learning · 13 Aug 2026 · paper
Learning new tasks by leveraging prior experience is a fundamental trait of intelligent systems. While Model-Agnostic Meta-Learning (MAML) is a leading approach, it suffers from significant…
arXiv · Machine Learning · 13 Aug 2026 · paper
Deep learning models, such as wide neural networks, can be conceptualized as nonlinear dynamical physical systems characterized by a multitude of interacting degrees of freedom. Such system…
arXiv · Machine Learning · 13 Aug 2026 · paper
Financial volatility is regime dependent, yet incorporating regime information into neural networks can also destabilize training. This paper asks where such information should enter a neur…
arXiv · Machine Learning · 13 Aug 2026 · paper
Autonomous rehabilitation systems must not only recognize human motion but also provide structured feedback to support users without continuous therapist supervision. This paper presents a…
arXiv · Machine Learning · 13 Aug 2026 · paper
Over the past decade, many test adequacy metrics have been proposed for deep learning that characterize test dataset adequacy from different perspectives, e.g., neuron activation behavior,…
arXiv · Machine Learning · 13 Aug 2026 · paper
Boosted decision trees (BDTs) are widely used in latency-critical applications, but efficient hardware deployment remains challenging. Existing designs often rely on uniform or manually tun…
arXiv · Machine Learning · 13 Aug 2026 · paper
A novel advective Fisher-Rao metric is introduced for optimization tasks on paths of probability measures governed by the continuity equation. This metric is shown to lead to optimal descen…
arXiv · Machine Learning · 13 Aug 2026 · paper
Vision-language models, such as contrastive language-image pre-training (CLIP)-based approaches, have reached state-of-the-art (SOTA) results in medical artificial intelligence. However, re…
arXiv · Machine Learning · 13 Aug 2026 · paper
We invert the typical formulation of sketch generation: instead of drawing strokes in order, we predict a 2D field that defines the order in which strokes are drawn. We use a pretrained lat…
arXiv · Machine Learning · 13 Aug 2026 · paper
Acceleration for deterministic root-finding problems has been extensively studied in recent years; specifically, the anchor-based, or Halpern-type methods achieve optimal convergence rates…
arXiv · Machine Learning · 13 Aug 2026 · paper
Bregman proximal stochastic gradient (BPSG) methods bring variance-reduced composite optimization to objectives whose geometry is poorly captured by Euclidean smoothness. Their performance,…
arXiv · Machine Learning · 13 Aug 2026 · paper
Identifying individual learning styles optimizes pedagogical efficacy. While traditional questionnaires are structured, behavioral tracking methods require prolonged interaction log accumul…
arXiv · Machine Learning · 13 Aug 2026 · paper
Cashew production is a widespread economic activity in Guinea-Bissau, as well as other countries in West Africa. However, unregulated cashew production can be directly associated with incre…
arXiv · Machine Learning · 13 Aug 2026 · paper
The robust treatment of environmental and operational variability (EOV) is an open challenge in population-based structural health monitoring (PBSHM). The difficulty is compounded in the ca…
arXiv · Machine Learning · 13 Aug 2026 · paper
Air traffic controllers perceive traffic complexity through the radar display, suggesting that a computer vision model operating on the same imagery may provide a natural architecture for m…
arXiv · Machine Learning · 13 Aug 2026 · paper
Edge-deployed vision systems in target recognition, surveillance, autonomous vehicles, and drone domains require hierarchical inference pipelines where a detection model identifies objects…