When Does Scale-Invariant Optimization Become Unstable? An Exact Schedule Law with Weight Decay
Normalization renders large parts of neural networks effectively scale invariant, inducing a hidden feedback l
- 用途
- 回帰
- 難易度
- Hard
- コスト
- High
「regression」の検索結果
71 件Normalization renders large parts of neural networks effectively scale invariant, inducing a hidden feedback l
The local dark matter density determines the strength of the signal expected in direct-detection experiments,
A growing body of work establishes that large language models are not mere statistical memorizers, but are cap
Sparsity is a powerful structural resource in optimization and statistics. We develop frameworks for leveragin
In this paper, we consider the scalar-on-function linear regression model under a realistic sampling scheme in
Large language models have collapsed the cost of producing lexically elaborate prose, and whether peer reviewe
Industrial process monitoring is fundamental to the safety and economic performance of modern process plants.
Neural networks acquire internal representations through learning. In this work, we formulate stochastic gradi
This paper introduces rlaopt, a PyTorch-based package for large-scale optimization and scientific computing us
Monocular depth estimation is a ubiquitous yet highly ill-posed computer vision task, with downstream applicat
Food waste in the restaurant sector poses a substantial challenge to environmental sustainability and economic
Understanding patient heterogeneity is key to improving prognostic modeling in traumatic brain injury (TBI). U
Wallet reputation scores decide who receives an airdrop, who can borrow, and who enters an allowlist across de
Purpose: Accurate CT protocol selection is critical for diagnostic quality and patient safety, yet the current
Standard semi-supervised learning (SSL) typically relies on labelled and unlabelled data sharing a common marg
We study the local errors of classical machine-learning surrogate models, which approximate the time evolution
We study the training dynamics of multiclass logistic regression on high-dimensional Gaussian mixture models w
Physics-informed neural networks (PINNs) struggle on PDEs whose governing physics varies across the domain. We
Kolmogorov-Arnold Networks (KANs) replace the fixed activation functions and linear weights of Multi-Layer Per
We convert black-box clinical prediction models for tabular data into standalone nomograms that can be audited
Statistical post-processing improves ensemble weather forecasts, but generating calibrated predictions at loca
Biomedical tables often combine thousands of measured variables with only tens or hundreds of labelled samples
This study investigates the classification of individuals as healthy or at risk of Parkinson's disease using m
Learning with group invariances is central to many scientific and geometric learning problems, yet its computa
Existing causal-inference benchmarks for LLMs mostly score method descriptions or whether generated code runs,
Deep learning models for CT scan analysis are often limited by the scarcity of precise pixel-level annotations
Deep learning models have been increasingly applied to Time Series Forecasting (TSF) in recent years. Transfor
Diffusion policies offer a powerful and expressive parameterization for continuous control. Yet, their integra
Trust-Hub reuses host circuits: several files differ mainly in the inserted Trojan. When gates from sibling va
The optimal transport (OT) map provides a geometric transformation for aligning probability distributions and
An LLM judge evaluates outputs at scale. Experts should label only where it is least sure.
In function-on-function regression, the coefficient surface $β(s,t)$ may exhibit complex support structure---f
We study variance-preserving diffusion of the response in mixed linear regression (MLR) with unknown mixing we
Late Gadolinium Enhancement (LGE) on cardiac magnetic resonance is a key marker of myocardial scar, but its li
Dynamical symbolic regression methods identify governing differential equations from noisy data, balancing int
We study regression under bounded responses in terms of excess mean squared error. When the comparator class i
Autonomous robot navigation failures differ not only in categorical severity but also in the physical context
Building on the pioneering paper of Kearns, Roth, and Ryu (SODA'26), we study information aggregation in a net
Rollcast is a probabilistic forecasting method for univariate time series that combines a compact set of rolli
Multi-Task semantic communication (SemCom) prioritizes simultaneous execution of multiple tasks over bit-accur
Constant optimization refines the numerical coefficients of candidate expressions in tree-based genetic progra
Several classical machine-learning methods, such as KRRs and SVRs, are both computationally and analytically t
We study when and how momentum improves large-batch training in the one-pass regime, using power-law kernel re
A fundamental quantity in machine learning is the optimal performance achievable by any model on a given task.
Generative modeling directly on geometric manifolds can avoid errors introduced by flattening non-Euclidean da
Stochastic gradient Markov chain Monte Carlo (SGMCMC) methods enable scalable Bayesian inference, but their pe
Symbolic Regression (SR) seeks to find succinct mathematical expressions that represent the fundamental relati
We develop a marginal coordinate test for regression with Euclidean predictors and a random-object response in
The performance of artificial intelligence (AI) and machine learning (ML) models degrades when the problem the
We study kernel ridge regression under anisotropic Gaussian data, where the input covariance decays as a power
Recurrent fast-weight memories and selective state-space models compress an expanding context into a fixed-siz
Kernels measure similarity or correlation in tasks such as regression and classification. The Gaussian kernel,
Randomized sketch-and-solve algorithms accelerate overconstrained $\ell_2$ regression by replacing the input w
Kolmogorov-Arnold Networks (KANs) replace fixed activations in deep architectures with learnable univariate ed
We address the simultaneous prediction of multiple high-dimensional physical fields governed by linear equalit
Exact Bayes prediction enjoys fast predictive regret guarantees, but exact posterior updating or representatio
Random feature methods provide a scalable approximation to kernel ridge regression (KRR), but the regularizati
Functional data analysis is an important statistical field that treats data as random functions. In practice,
ES-HyperNEAT evolves substrate topology through adaptive quadtree subdivision; to our knowledge, no implementa
Population optimizers such as CMA-ES, DE, and multi-objective evolutionary algorithms drive search mainly thro
Reservoir computing (RC) couples a fixed recurrent dynamical system with a trained lightweight readout, but th
The study employed an Artificial Neural Network in combination with the optimized Adaptive Moment Estimation (
We investigate whether agentic artificial intelligence can automate parts of the process of designing genetic
Evolutionary feature construction has shown strong promise in symbolic regression by automatically discovering
Genetic Programming Symbolic Regression (GPSR) generates mathematical expressions to model input-output relati
Symbolic regression provides analytical expressions, but it is usually applied one output at a time. This is l
回避可能な計算資源である雑音を扱い、学習と推論を行うための神経回路モデル(NNN)を提案。このモデルが学習し推論するための、伝達の反対方向の重みトランスポートの問題を回避できた。
Parent selection significantly affects exploration, exploitation, and complexity control in genetic programmin
Gene regulatory network modeling often requires balancing predictive accuracy and mechanistic interpretability
We present a regression-based approach to Arabic dialect geolocation that models dialectal variation as a cont
fMRIデータから視覚情報を解釈するために、スパイクニューラルネットワークを用いた方法を提案し、fMRIデータから視覚情報を解釈する検証を行う。