Scout score 22.7 title matches inverse, modeling, systems / watch term operator learning / category weight cs.LG +0.8
Modeling spatiotemporal dynamical systems governed by partial differential equations (PDEs) poses two major challenges: it either requires expensive physics-based simulators that entail iterative numerical solving at high computational cost, or it depends on abundant training data, yet purely data-driven models often generalize poorly to downstream dynamic operating conditions. We propose DEFT, a frequency-domain data sampling method that identifies the dominant Fourier modes of a physical system and...
Connection to your workThe local scout matched this paper to your profile through inverse, modeling, systems, data, operator learning.
Next action: Read the introduction and main theorem statements, focusing on inverse, modeling, systems.
Scout score 19.98 title matches uncertainty quantification, uncertainty, quantification / watch term uncertainty quantification / category weight cs.LG +0.8
Reliable clinical deployment of machine learning requires models that know when they are likely to fail, particularly for subgroups underrepresented in training data. A common case is pediatric care, where models trained on adult cohorts can silently under-perform on children with no indication that something has gone wrong.
Connection to your workThe local scout matched this paper to your profile through uncertainty quantification, uncertainty, quantification, learning, training.
Next action: Read the introduction and main theorem statements, focusing on uncertainty quantification, uncertainty, quantification.
Scout score 17.18 title matches learning, uncertainty / watch term uncertainty quantification / category weight cs.LG +0.8
Deep learning models have emerged as the standard computational tool for a wide range of applications in genomics. Yet, uncertainty quantification (UQ) -- and more specifically, the reliability of different uncertainty estimates in this domain -- has received little systematic attention.
Connection to your workThe local scout matched this paper to your profile through learning, uncertainty, uncertainty quantification, bayesian, neural.
Next action: Read the introduction and main theorem statements, focusing on learning, uncertainty, uncertainty quantification.
Scout score 14.5 title matches behavioral / watch term drift / category weight cs.LG +0.8
Benchmark leaderboards summarize how well a language model performs, but not how its behavior relates to that of other models or changes across generations. We characterize the output behavior of 32 models from six families using their responses to a shared bank of 10{,}000 prompts.
Connection to your workThe local scout matched this paper to your profile through behavioral, drift, behavior, training, response.
Next action: Read the introduction and main theorem statements, focusing on behavioral, drift, behavior.
Inverse reinforcement learning (IRL) aims to recover a reward function under which the resulting policy reproduces the behavior observed in expert demonstrations. A natural approach is to formulate IRL as a bilevel optimization problem, in which the inner level corresponds to policy optimization under the learned reward and the outer level measures the discrepancy between the induced policy and expert data.
Connection to your workThe local scout matched this paper to your profile through learning, inverse, control, stochastic, behavior.
Next action: Read the introduction and main theorem statements, focusing on learning, inverse, control.
Haiteng Wang, Yunfei Zhu, Tao Wang, Yikang Li, Jiabao Dong, Xiaoge Zhang et al.
Scout score 12.42 title matches systems, data / category weight cs.LG +0.8 / RL ranker +2.4 from 500 reward events
Industrial time-series signals, such as turbine temperature and rotational speed in aero-engines, are essential for monitoring the health and operational status of complex dynamical systems. However, collecting such data is often limited by harsh environments (e.g., high temperature and high pressure) and the high cost of experimental testing.
Connection to your workThe local scout matched this paper to your profile through systems, data, training, high, processes.
Next action: Read the introduction and main theorem statements, focusing on systems, data, training.
Constant-stepsize temporal-difference (TD) learning is attractive for policy evaluation, but inference from a single Markov trajectory must account for serial dependence and a stepsize-dependent stationary target. For fixed-stepsize linear TD, we establish a functional central limit theorem whose covariance retains the multiplicative component induced by the random TD matrix and the stationary iterate error.
Connection to your workThe local scout matched this paper to your profile through learning, inference, behavior, dependence.
Next action: Read the introduction and main theorem statements, focusing on learning, inference, behavior.
Clustering is a fundamental class of data analysis techniques with the most important representatives being centroid-based methods like $k$-means. Such methods are strongly connected to quantization problems, which aim to approximate general probability measures with discrete ones.
Connection to your workThe local scout matched this paper to your profile through optimal, neural, modeling, networks, data.
Next action: Read the introduction and main theorem statements, focusing on optimal, neural, modeling.
Robert Bitterling, Christian Nettersheim, Jörn Hees, Michael Rademacher
Scout score 10.46 title matches prediction / category weight cs.LG +0.8 / RL ranker +2.3 from 500 reward events
Low Power Wide Area Networks like LoRa are increasingly deployed for smart city applications, requiring accurate path loss prediction for effective network planning. Traditional (empirical) propagation models often exhibit limited accuracy in these scenarios.
Connection to your workThe local scout matched this paper to your profile through prediction, learning, training, machine, networks.
Next action: Read the introduction and main theorem statements, focusing on prediction, learning, training.
We study a Restart POMDP (Partially Observable Markov Decision Process) on a general Borel state space, where the controller either lets the hidden state evolve unobserved or restarts the system and observes the new state. Exploiting a sufficient-statistic representation consisting of the last observed state and the elapsed time since restart, we reduce the problem to a fully observed MDP.
Connection to your workThe local scout matched this paper to your profile through optimal, when.
Next action: Read the introduction and main theorem statements, focusing on optimal, when.
Spectral clustering methods for network data are commonly based on a few matrix representations, such as the adjacency matrix and the symmetric Laplacian. We study a continuum of degree-normalized spectral embeddings that includes these commonly used choices as special cases.
Connection to your workThe local scout matched this paper to your profile through graphs, stochastic, uncertainty, data, probability.
Next action: Read the introduction and main theorem statements, focusing on graphs, stochastic, uncertainty.
Ruirui Wang, Yanke Li, Manuel Günther, Diego Paez-Granados
Scout score 9.31 title matches learning / category weight cs.LG +0.8 / RL ranker +1.8 from 500 reward events
Healthcare data, such as Intensive Care Unit (ICU) records, comprise heterogeneous multivariate time series sampled at irregular intervals with pervasive missingness. However, clinical applications demand predictive models that are both accurate and interpretable.
Connection to your workThe local scout matched this paper to your profile through learning, graphs, training, neural, data.
Next action: Read the introduction and main theorem statements, focusing on learning, graphs, training.
Scout score 8.93 title matches training / category weight cs.LG +0.8 / RL ranker +1.4 from 500 reward events
Large-scale neural recommender systems are typically trained with a softmax cross-entropy objective over the full item vocabulary. For a typical large number of possible items $K$, the final classification layer dominates memory, requiring $O(nK)$ logits and gradients to materialize for a batch of $n$ examples.
Connection to your workThe local scout matched this paper to your profile through training, neural, systems, number, science.
Next action: Read the introduction and main theorem statements, focusing on training, neural, systems.
Scout score 8.85 title matches probability / category weight cs.LG +0.8 / RL ranker +1.3 from 500 reward events
The attention mechanism forms the foundation of many modern AI models such as the Transformer. In one subclass of problems where attention is used, inputs and outputs are bound to the probability simplex so that all outputs sum to one.
Connection to your workThe local scout matched this paper to your profile through probability, inverse, stochastic, machine, boundary.
Next action: Read the introduction and main theorem statements, focusing on probability, inverse, stochastic.
Nigel Bastian Cendra, Abdelhamid Ezzerg, Fernando Julio Cendra, Jeremias Knoblauch, Jakob Zeitler
Scout score 8.48 title matches bayesian, neural / category weight cs.LG +0.8 / RL ranker +1.3 from 500 reward events
Gradient-free post-training has emerged as a compelling alternative to gradient-based optimization for large language models (LLMs), but existing approaches remain costly. We ask whether structured search can identify a strong single expert under a modest evaluation budget.
Connection to your workThe local scout matched this paper to your profile through bayesian, neural, training.
Next action: Read the introduction and main theorem statements, focusing on bayesian, neural, training.
Modern reinforcement learning (RL) post-training pipelines for large language models (LLMs) increasingly combine rollout workloads across multiple domains and feedback paradigms. Prefix-aware routing improves inference efficiency through cache reuse and load balancing, but it does not control how heterogeneous rollout sessions compete for KV-cache capacity.
Connection to your workThe local scout matched this paper to your profile through learning, control, training, mean, human.
Next action: Read the introduction and main theorem statements, focusing on learning, control, training.
Worst-case multiclass bounds do not become smaller when the best classifier is already nearly correct: what is missing is an optimistic rate, a guarantee whose fluctuation scales with the oracle risk itself. For a class of Natarajan dimension $d_N$ and Daniely-Shalev-Shwartz dimension $d_{DS}$, the optimal excess risk is known at the two endpoints ($d_{DS}/n$ realizable, $\sqrt{d_N/n}+d_{DS}/n$ agnostic [HMZ24, CEH+26, Pab26]) and open in between.
Connection to your workThe local scout matched this paper to your profile through learning, optimal, when.
Next action: Skim the abstract, introduction, and conclusion for learning, optimal, when.
Scout score 7.48 title matches learning / category weight cs.LG +0.8 / RL ranker +1.4 from 500 reward events
Instant delivery platforms have become a critical component of urban logistics, increasingly relying on crowdsourced couriers to fulfill highly dynamic orders. In real-world systems, couriers are not exclusive to a single platform and may concurrently serve multiple platforms, while each platform can only observe its own orders and couriers' interactions due to privacy and operational constraints.
Connection to your workThe local scout matched this paper to your profile through learning, systems, high, when.
Next action: Skim the abstract, introduction, and conclusion for learning, systems, high.
Kiran Madhusudhanan, Christian Klötergens, Lars Schmidt-Thieme, Vijaya Krishna Yalavarthi
Scout score 7.3 title matches mean / category weight cs.LG +0.8 / RL ranker +1.2 from 500 reward events
Probabilistic forecasting plays an essential role in risk-sensitive decision-making, particularly in long-horizon settings. However, existing approaches often face a fundamental trade-off between distributional flexibility and accurate mean prediction.
Connection to your workThe local scout matched this paper to your profile through mean, uncertainty, when, prediction.
Next action: Skim the abstract, introduction, and conclusion for mean, uncertainty, when.
Scout score 7.23 title matches training / category weight cs.LG +0.8 / RL ranker +1.1 from 500 reward events
In LLM pre-training, synchronization propagates rank-local stalls, slowdowns, and numerical errors into job-wide symptoms, obscuring their origin. Existing diagnosis often relies on in-process monitors that cannot report after the trainer blocks or terminates, or on post-mortem logs that preserve only synchronized symptoms; offline health tests lose the workload and operating conditions that triggered the failure.
Connection to your workThe local scout matched this paper to your profile through training, compact, data, when.
Next action: Skim the abstract, introduction, and conclusion for training, compact, data.
Viktoria Schuster, Sana Tonekaboni, Caroline Uhler
Scout score 7.08 title matches data / category weight cs.LG +0.8 / RL ranker +1.0 from 500 reward events
Determining the complexity, or Intrinsic Dimension (ID), of data is fundamental to efficient and interpretable representation learning. This is particularly challenging in multi-modal settings when trying to learn disentangled representations for shared and private information.
Connection to your workThe local scout matched this paper to your profile through data, learning, applied, when.
Next action: Skim the abstract, introduction, and conclusion for data, learning, applied.
The Naive Bayes (NB) classifier remains a standard choice for categorical data, yet its widely used smoothing rules, such as Laplace, Lidstone, Krichevsky-Trofimov, and the $m$-estimate, all prescribe a fixed smoothing strength that ignores feature cardinality, sample size, and class imbalance, inducing a non-vanishing bias on modern high-cardinality tabular data. We propose hierarchical empirical-Bayes Naive Bayes (HEB-NB), in which each class-feature conditional probability is smoothed by a Dirichlet prior whose...
Connection to your workThe local scout matched this paper to your profile through weighting, dependence, high, data, probability.
Next action: Skim the abstract, introduction, and conclusion for weighting, dependence, high.
We systematically investigate finite-difference (FD) derivative computation in Physics-Informed Neural Networks (PINNs) as an alternative to automatic differentiation (AD). On three benchmark PDEs we show that, with a properly calibrated step size, FD matches AD in accuracy on every problem while running faster across the full tested batch-size range and using substantially less GPU memory, and that a stochastic variant we propose outperforms AD on a stationary problem.
Connection to your workThe local scout matched this paper to your profile through stochastic, pdes, neural, networks.
Next action: Skim the abstract, introduction, and conclusion for stochastic, pdes, neural.
On-policy distillation (OPD) applies token-level teacher supervision to student-generated trajectories, but this supervision is not always reliable. Existing methods use local confidence or teacher-student agreement to weight, filter, or truncate the sampled trajectory.
Connection to your workThe local scout matched this paper to your profile through training, mathematics, mean, high, probability.
Next action: Skim the abstract, introduction, and conclusion for training, mathematics, mean.