Вход на сайт

Проиндексировано 95213572 новости

Найдите то, что Вас интересует

Последние поступления

будьте всегда теме

В Турции застрелили члена правящей партии и двух членов его семьи

Дата публикации: 26-07-2026 20:34:46

В турецкой провинции Батман были застрелены член руководства районного отделения правящей партии, его сын и племянник.

Классификация: Международные. Схожих патентов: 0. Схожих новостей: 9. Тональность: 0. Информативность: 6.64. Источник: aif.ru.

Раскрыто послужившее началом противостояния Зеленского и Европы событие

Дата публикации: 26-07-2026 20:34:33

Классификация: Международные. Схожих патентов: 0. Схожих новостей: 9. Тональность: 0. Информативность: 10. Источник: lenta.ru.

Россиянин Илюмжинов не вошел в список кандидатов в президенты FIDE

Дата публикации: 26-07-2026 20:33:25

Россиянин Кирсан Илюмжинов не вошел в список кандидатов на пост президента Международной шахматной организации (FIDE). Он возглавлял организацию с 1995 по 2018 год.

Классификация: Информация. Схожих патентов: 0. Схожих новостей: 9. Тональность: 0. Информативность: 9.08. Источник: www.kommersant.ru.

Three paramedics, one civilian wounded in drone attack in DPR’s Gorlovka

Дата публикации: 26-07-2026 20:31:37

Mayor Ivan Prikhodko said an ambulance car was attacked deliberately

Классификация: . Схожих патентов: 0. Схожих новостей: 9. Тональность: 0. Информативность: 9.07. Источник: tass.com.

В Донецке провели акцию памяти о погибших от украинской агрессии детях ДНР и ЛНР

Дата публикации: 26-07-2026 20:31:30

Мероприятие состоялось у Донбасс оперы

Классификация: Общество. Схожих патентов: 0. Схожих новостей: 9. Тональность: 0. Информативность: 10. Источник: www.itar-tass.com.

Юрий Гаврилов: Сосудистые центры в Москве находятся на расстоянии вытянутой руки

Дата публикации: 26-07-2026 20:30:38

Сердечно-сосудистые заболевания по-прежнему уносят больше всего жизней во всем мире. Между тем в российской столице все идет к тому, чтобы переломить эту тенденцию. Лишь за последние пятнадцать лет число спасенных пациентов с инфарктом миокарда увеличилось более чем в два раза, а смертность от инсульта снизилась на 15 процентов. Такие данные привел недавно в соцсетях мэр Москвы Сергей Собянин. Глава города связывает это прежде всего с созданием в Москве единых инфарктных и инсультных сетей на базе стационаров. Что реально изменилось с их появлением, корреспонденту "РГ" рассказал Юрий Гаврилов, руководитель регионального сосудистого центра городской клинической больницы им. Буянова, одного из двух десятков ...

Классификация: Москва. Схожих патентов: 0. Схожих новостей: 9. Тональность: 0. Информативность: 9.77. Источник: rg.ru.

"Война с Европой": В Турции прокомментировали резонансное решение Зеленского

Дата публикации: 26-07-2026 20:29:50

Турецкое издание dikGazete пишет, что отставка Михаила Федорова с поста министра обороны Украины стала началом открытого противостояния между Владимиром Зеленским и Европой

Классификация: Международные. Схожих патентов: 0. Схожих новостей: 10. Тональность: 0. Информативность: 8.75. Источник: www.mk.ru.

Трамп разместил в соцсети Truth Social фотографии с отсылкой к выборам 2028 года

Дата публикации: 26-07-2026 20:28:24


Президент США Дональд Трамп опубликовал в своей соцсети Truth Social несколько изображений с надписью «2028», отсылающей к следующим президентским выборам. Читать далее

Классификация: Международные. Схожих патентов: 0. Схожих новостей: 9. Тональность: 0. Информативность: 11.79. Источник: russian.rt.com.

Журова заявила, что уровень мирового фигурного катания снижается

Дата публикации: 26-07-2026 20:27:52


Олимпийская чемпионка и депутат Госдумы Светлана Журова считает, что уровень мирового фигурного катания снижается. Читать далее

Классификация: Спорт. Схожих патентов: 0. Схожих новостей: 9. Тональность: 0. Информативность: 4.62. Источник: russian.rt.com.

В Калифорнии один человек погиб при стрельбе на вечеринке

Дата публикации: 26-07-2026 20:26:09

Один человек погиб, шесть пострадали в результате стрельбы на вечеринке в американском штате Калифорния, сообщает офис шерифа округа Санта-Клара.

Классификация: Происшествия. Схожих патентов: 0. Схожих новостей: 9. Тональность: 0. Информативность: 8.3. Источник: ria.ru.

Learning from Similar Linear Representations: Adaptivity, Minimaxity, and Robustness

Дата публикации: 26-07-2026 20:25:00


Representation multi-task learning (MTL) has achieved tremendous success in practice. However, the theoretical understanding of these methods is still lacking. Most existing theoretical works focus on cases where all tasks share the same representation, and claim that MTL almost always improves performance. Nevertheless, as the number of tasks grows, assuming all tasks share the same representation is unrealistic. Furthermore, empirical findings often indicate that a shared representation does not necessarily improve single-task learning performance. In this paper, we aim to understand how to learn from tasks with similar but not exactly the same linear representations, while dealing with outlier tasks. ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Multiple Instance Verification

Дата публикации: 26-07-2026 20:25:00


We explore multiple instance verification, a problem setting in which a query instance is verified against a bag of target instances with heterogeneous, unknown relevancy. We show that naive adaptations of attention-based multiple instance learning (MIL) methods and standard verification methods like Siamese neural networks are unsuitable for this setting: directly combining state-of-the-art (SOTA) MIL methods and Siamese networks is shown to be no better, and sometimes significantly worse, than a simple baseline model. Postulating that this may be caused by the failure of the representation of the target bag to incorporate the query instance, we introduce a new pooling ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

EF21 with Bells & Whistles: Six Algorithmic Extensions of Modern Error Feedback

Дата публикации: 26-07-2026 20:25:00


First proposed by Seide (2014) as a heuristic, error feedback (EF) is a very popular mechanism for enforcing convergence of distributed gradient-based optimization methods enhanced with communication compression strategies based on the application of contractive compression operators. However, existing theory of EF relies on very strong assumptions (e.g., bounded gradients), and provides pessimistic convergence rates (e.g., while the best known rate for EF in the smooth nonconvex regime, and when full gradients are compressed, is $O(1/T^{2/3})$, the rate of gradient descent in the same regime is $O(1/T)$). Recently, Richtàrik et al. (2021) proposed a new error feedback mechanism, EF21, based ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Model-free Change-Point Detection Using AUC of a Classifier

Дата публикации: 26-07-2026 20:25:00


In contemporary data analysis, it is increasingly common to work with non-stationary complex data sets. These data sets typically extend beyond the classical low-dimensional Euclidean space, making it challenging to detect shifts in their distribution without relying on strong structural assumptions. This paper proposes a novel offline change-point detection method that leverages classifiers developed in the statistics and machine learning community. With suitable data splitting, the test statistic is constructed through sequential computation of the Area Under the Curve (AUC) of a classifier, which is trained on data segments on both ends of the sequence. It is shown that the ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

A New Random Reshuffling Method for Nonsmooth Nonconvex Finite-sum Optimization

Дата публикации: 26-07-2026 20:25:00


Random reshuffling techniques are prevalent in large-scale applications, such as training neural networks. While the convergence and acceleration effects of random reshuffling-type methods are fairly well understood in the smooth setting, much less studies seem available in the nonsmooth case. In this work, we design a new normal map-based proximal random reshuffling (norm-PRR) method for nonsmooth nonconvex finite-sum problems. We show that norm-PRR achieves the iteration complexity ${\cal O}(n^{-1/3}T^{-2/3})$ where $n$ denotes the number of component functions $f(\cdot,i)$ and $T$ counts the total number of iterations. This improves the currently known complexity bounds for this class of problems by a ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Learning with Linear Function Approximations in Mean-Field Control

Дата публикации: 26-07-2026 20:25:00


The paper focuses on mean-field type multi-agent control problems with finite state and action spaces where the dynamics and cost structures are symmetric and homogeneous, and are affected by the distribution of the agents. A standard solution method for these problems is to consider the infinite population limit as an approximation and use symmetric solutions of the limit problem to achieve near optimality. The control policies, and in particular the dynamics, depend on the population distribution in the finite population setting, or the marginal distribution of the state variable of a representative agent for the infinite population setting. Hence, learning ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

On the Convergence of Projected Policy Gradient for Any Constant Step Sizes

Дата публикации: 26-07-2026 20:25:00


Projected policy gradient (PPG) is a basic policy optimization method in reinforcement learning. Given access to exact policy evaluations, previous studies have established the sublinear convergence of PPG for sufficiently small step sizes based on the smoothness and the gradient domination properties of the value function. However, as the step size goes to infinity, PPG reduces to the classic policy iteration method, which suggests the convergence of PPG even for large step sizes. In this paper, we fill this gap and show that PPG admits a sublinear convergence for any constant step sizes. Due to the existence of the state-wise ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Linear Separation Capacity of Self-Supervised Representation Learning

Дата публикации: 26-07-2026 20:25:00


Recent advances in self-supervised learning have highlighted the efficacy of data augmentation in learning data representation from unlabeled data. Training a linear model atop these enhanced representations can yield an adept classifier. Despite the remarkable empirical performance, the underlying mechanisms that enable data augmentation to unravel nonlinear data structures into linearly separable representations remain elusive. This paper seeks to bridge this gap by investigating under what conditions learned representations can linearly separate manifolds when data is drawn from a multi-manifold model. Our investigation reveals that data augmentation offers additional information beyond observed data and can thus improve the information-theoretic optimal ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

On Non-asymptotic Theory of Recurrent Neural Networks in Temporal Point Processes

Дата публикации: 26-07-2026 20:25:00


Temporal point process (TPP) is an important tool for modeling and predicting irregularly timed events across various domains. Recently, the recurrent neural network (RNN)-based TPPs have shown practical advantages over traditional parametric TPP models. However, in the current literature, it remains nascent in understanding neural TPPs from theoretical viewpoints. In this paper, we establish the excess risk bounds of RNN-TPPs under many well-known TPP settings. We especially show that an RNN-TPP with no more than four layers can achieve vanishing generalization errors. Our technical contributions include the characterization of the complexity of the multi-layer RNN class, the construction of $\tanh$ ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Frontiers to the learning of nonparametric hidden Markov models

Дата публикации: 26-07-2026 20:25:00


Hidden Markov models (HMMs) are flexible tools for clustering dependent data coming from unknown populations, allowing nonparametric modelling of the population densities. Identifiability fails when the data is in fact independent and identically distributed (i.i.d.), and we study the frontier between learnable and unlearnable two-state nonparametric HMMs. Learning the parameters of the HMM requires solving a nonlinear inverse problem whose difficulty depends not only on the smoothnesses of the populations but also on the distance to the i.i.d. boundary of the parameter set. The latter difficulty is mostly ignored in the literature in favour of assumptions precluding nearly independent data. ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

WEFE: A Python Library for Measuring and Mitigating Bias in Word Embeddings

Дата публикации: 26-07-2026 20:25:00


Word embeddings, which are a mapping of words into continuous vectors, are widely used in modern Natural Language Processing (NLP) systems. However, they are prone to inherit stereotypical social biases from the corpus on which they are built.
The research community has focused on two main tasks to address this problem: 1) how to measure these biases, and 2) how to mitigate them.
Word Embedding Fairness Evaluation (WEFE) is an open source library that implements many fairness metrics and mitigation methods in a unified framework. It also provides a standard interface for designing new ones.
The software follows the object-oriented paradigm ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Regularized Rényi Divergence Minimization through Bregman Proximal Gradient Algorithms

Дата публикации: 26-07-2026 20:25:00


We study the variational inference problem of minimizing a regularized Rényi divergence over an exponential family. We propose to solve this problem with a Bregman proximal gradient algorithm. We propose a sampling-based algorithm to cover the black-box setting, corresponding to a stochastic Bregman proximal gradient algorithm with biased gradient estimator. We show that the resulting algorithms can be seen as relaxed moment-matching algorithms with an additional proximal step. Using Bregman updates instead of Euclidean ones allows us to exploit the geometry of our approximate model. We prove strong convergence guarantees for both our deterministic and stochastic algorithms using this viewpoint, ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Score-Based Diffusion Models in Function Space

Дата публикации: 26-07-2026 20:25:00


Diffusion models have recently emerged as a powerful framework for generative modeling. They consist of a forward process that perturbs input data with Gaussian white noise and a reverse process that learns a score function to generate samples by denoising. Despite their tremendous success, they are mostly formulated on finite-dimensional spaces, e.g., Euclidean, limiting their applications to many domains where the data has a functional form, such as in scientific computing and 3D geometric data analysis. This work introduces a mathematically rigorous framework called Denoising Diffusion Operators (DDOs) for training diffusion models in function space. In DDOs, the forward process ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Simplex Constrained Sparse Optimization via Tail Screening

Дата публикации: 26-07-2026 20:25:00


We consider the probabilistic simplex-constrained sparse recovery problem. The commonly used Lasso-type penalty for promoting sparsity is ineffective in this context since it is a constant within the simplex. Despite this challenge, fortunately, simplex constraint itself brings a self-regularization property, i.e., the empirical risk minimizer without any sparsity-promoting procedure obtains the usual Lasso-type estimation error. Moreover, we analyze the iterates of a projected gradient descent method and show its convergence to the ground truth sparse solution in the geometric rate until a satisfied statistical precision is attained. Although the estimation error is statistically optimal, the resulting solution is usually more ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Density Estimation Using the Perceptron

Дата публикации: 26-07-2026 20:25:00


We propose a new density estimation algorithm. Given
$n$ i.i.d. observations from a distribution belonging to a class
of densities on $\mathbb{R}^d$, our estimator outputs any density in the class whose “perceptron
discrepancy” with the empirical distribution is at most $O(\sqrt{d/n})$.
The perceptron discrepancy is defined as the largest
difference in mass two distribution place on any halfspace. It is shown that
this estimator achieves the expected total variation distance to the truth that is almost
minimax optimal over the class of densities with bounded Sobolev norm and Gaussian
mixtures. This suggests that the regularity of the prior distribution could be an
explanation for the efficiency of the ubiquitous ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Extending Temperature Scaling with Homogenizing Maps

Дата публикации: 26-07-2026 20:25:00


As machine learning models continue to grow more complex, poor calibration significantly limits the reliability of their predictions. Temperature scaling learns a single temperature parameter to scale the output logits, and despite its simplicity, remains one of the most effective post-hoc recalibration methods. We identify one of temperature scaling's defining attributes, that it increases the uncertainty of the predictions in a manner that we term homogenization, and propose to learn the optimal recalibration mapping from a larger class of functions that satisfies this property. We demonstrate the advantage of our method over temperature scaling in both calibration and out-of-distribution detection. ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Distribution Estimation under the Infinity Norm

Дата публикации: 26-07-2026 20:25:00


We present novel bounds for estimating discrete probability distributions under the $\ell_\infty$ norm. These are nearly optimal in various precise senses, including a kind of instance-optimality. Our data-dependent convergence guarantees for the maximum likelihood estimator significantly improve upon the currently known results. A variety of techniques are utilized and innovated upon, including Chernoff-type inequalities and empirical Bernstein bounds. We illustrate our results in synthetic and real-world experiments. Finally, we apply our proposed framework to a basic selective inference problem, where we estimate the most frequent probabilities in a sample.

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

System Neural Diversity: Measuring Behavioral Heterogeneity in Multi-Agent Learning

Дата публикации: 26-07-2026 20:25:00


Evolutionary science provides evidence that diversity confers resilience in natural systems. Yet, traditional multi-agent reinforcement learning techniques commonly enforce homogeneity to increase training sample efficiency. When a system of learning agents is not constrained to homogeneous policies, individuals may develop diverse behaviors, resulting in emergent complementarity that benefits the system. Despite this, there is a surprising lack of tools that quantify behavioral diversity. Such techniques would pave the way towards understanding the impact of diversity in collective artificial intelligence and enabling its control. In this paper, we introduce System Neural Diversity (SND): a measure of behavioral heterogeneity in multi-agent systems. ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Nonparametric Regression on Random Geometric Graphs Sampled from Submanifolds

Дата публикации: 26-07-2026 20:25:00


We consider the nonparametric regression problem when the covariates are located on an unknown compact submanifold of a Euclidean space. Under defining a random geometric graph structure over the covariates we analyse the asymptotic frequentist behaviour of the posterior distribution arising from Bayesian priors designed through random basis expansion in the graph Laplacian eigenbasis. Under Hölder smoothness assumption on the regression function and the density of the covariates over the submanifold, we prove that the posterior contraction rates of such methods are minimax optimal (up to logarithmic factors) for any positive smoothness index.

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Autoencoders in Function Space

Дата публикации: 26-07-2026 20:25:00


Autoencoders have found widespread application in both their original deterministic form and in their variational formulation (VAEs). In scientific applications and in image processing it is often of interest to consider data that are viewed as functions; while discretisation (of differential equations arising in the sciences) or pixellation (of images) renders problems finite dimensional in practice, conceiving first of algorithms that operate on functions, and only then discretising or pixellating, leads to better algorithms that smoothly operate between resolutions. In this paper function-space versions of the autoencoder (FAE) and variational autoencoder (FVAE) are introduced, analysed, and deployed. Well-definedness of the ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

EMaP: Explainable AI with Manifold-based Perturbations

Дата публикации: 26-07-2026 20:25:00


In the last few years, many explanation methods based on the perturbations of input data have been introduced to shed light on the predictions generated by black-box models. The goal of this work is to introduce a novel perturbation scheme so that more faithful and robust explanations can be obtained. Our study focuses on the impact of perturbing directions on the data topology. We show that perturbing along the orthogonal directions of the input manifold better preserves the data topology, both in the worst-case analysis of the discrete Gromov-Hausdorff distance and in the average-case analysis via persistent homology. From those ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Asymptotic Inference for Multi-Stage Stationary Treatment Policy with Variable Selection

Дата публикации: 26-07-2026 20:25:00


Dynamic treatment regimes or policies are a sequence of decision functions over multiple stages that are tailored to individual features. One important class of treatment policies in practice, namely multi-stage stationary treatment policies, prescribes treatment assignment probabilities using the same decision function across stages, where the decision is based on the same set of features consisting of time-evolving variables (e.g., routinely collected disease biomarkers). Although there has been extensive literature on constructing valid inference for the value function associated with dynamic treatment policies, little work has focused on the policies themselves, especially in the presence of high-dimensional features. We aim ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Frequentist Guarantees of Distributed (Non)-Bayesian Inference

Дата публикации: 26-07-2026 20:25:00


We establish frequentist properties, i.e., posterior consistency, asymptotic normality, and posterior contraction rates, for the distributed (non-)Bayesian inference problem for a set of agents connected over a network. These results are motivated by the need to analyze large, decentralized datasets, where distributed (non)-Bayesian inference has become a critical research area across multiple fields, including statistics, machine learning, and economics. Our results show that, under appropriate assumptions on the communication graph, distributed (non)-Bayesian inference retains parametric efficiency while enhancing robustness in uncertainty quantification. We also explore the trade-off between statistical efficiency and communication efficiency by examining how the design and size ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Boosting Causal Additive Models

Дата публикации: 26-07-2026 20:25:00


We present a boosting-based method to learn additive Structural Equation Models (SEMs) from observational data, with a focus on the theoretical aspects of determining the causal order among variables. We introduce a family of score functions based on arbitrary regression techniques, for which we establish sufficient conditions that guarantee consistent identification of the true causal ordering. Our analysis reveals that boosting with early stopping meets these criteria and thus offers a consistent score function for causal orderings. To address the challenges posed by high-dimensional data sets, we adapt our approach through a component-wise gradient descent in the space of additive ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Contextual Bandits with Stage-wise Constraints

Дата публикации: 26-07-2026 20:25:00


We study contextual bandits in the presence of a stage-wise constraint when the constraint must be satisfied both with high probability and in expectation. We start with the linear case where both the reward function and the stage-wise constraint (cost function) are linear. In each of the high probability and in expectation settings, we propose an upper-confidence bound algorithm for the problem and prove a $T$-round regret bound for it. We also prove a lower-bound for this constrained problem, show how our algorithms and analyses can be extended to multiple constraints, and provide simulations to validate our theoretical results. In ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Data-Driven Performance Guarantees for Classical and Learned Optimizers

Дата публикации: 26-07-2026 20:25:00


We introduce a data-driven approach to analyze the performance of continuous optimization algorithms using generalization guarantees from statistical learning theory. We study classical and learned optimizers to solve families of parametric optimization problems. We build generalization guarantees for classical optimizers, using a sample convergence bound, and for learned optimizers, using the Probably Approximately Correct (PAC)-Bayes framework. To train learned optimizers, we use a gradient-based algorithm to directly minimize the PAC-Bayes upper bound. Numerical experiments in signal processing, control, and meta-learning showcase the ability of our framework to provide strong generalization guarantees for both classical and learned optimizers given a fixed ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Enhanced Feature Learning via Regularisation: Integrating Neural Networks and Kernel Methods

Дата публикации: 26-07-2026 20:25:00


We propose a new method for feature learning and function estimation in supervised learning via regularised empirical risk minimisation. Our approach considers functions as expectations of Sobolev functions over all possible one-dimensional projections of the data. This framework is similar to kernel ridge regression, where the kernel is E_w(k(B)(wx, wx')), with k(B)(a, b) := min(|a|, |b|)1_{ab>0} the Brownian kernel, and the distribution of the projections w is learnt. This can also be viewed as an infinite-width one-hidden layer neural network, optimising the first layer’s weights through gradient descent and explicitly adjusting the non-linearity and weights of the second layer. We ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Interpretable Global Minima of Deep ReLU Neural Networks on Sequentially Separable Data

Дата публикации: 26-07-2026 20:25:00


We explicitly construct zero loss neural network classifiers. We write the weight matrices and bias vectors in terms of cumulative parameters, which determine truncation maps acting recursively on input space. The configurations for the training data considered are $(i)$ sufficiently small, well separated clusters corresponding to each class, and $(ii)$ equivalence classes which are sequentially linearly separable. In the best case, for $Q$ classes of data in $\mathbb{R}^{M}$, global minimizers can be described with $Q(M+2)$ parameters.

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Best Linear Unbiased Estimate from Privatized Contingency Tables

Дата публикации: 26-07-2026 20:25:00


In differential privacy (DP) mechanisms, it can be beneficial to release "redundant" outputs, where some quantities can be estimated in multiple ways by combining different privatized values. Indeed, the DP 2020 Decennial Census products published by the U.S. Census Bureau consist of such redundant noisy counts. When redundancy is present, the DP output can be improved by enforcing self-consistency (i.e., estimators obtained using different noisy counts result in the same value), and we show that the minimum variance processing is a linear projection. However, standard projection algorithms require excessive computation and memory, making them impractical for large-scale applications such as ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

High-Rank Irreducible Cartesian Tensor Decomposition and Bases of Equivariant Spaces

Дата публикации: 26-07-2026 20:25:00


Irreducible Cartesian tensors (ICTs) play a crucial role in the design of equivariant graph neural networks, as well as in theoretical chemistry and chemical physics. Meanwhile, the design space of available linear operations on tensors that preserve symmetry presents a significant challenge. The ICT decomposition and a basis of this equivariant space are difficult to obtain for high-rank tensors. After decades of research, Bonvicini (2024) has recently achieved an explicit ICT decomposition for $n=5$ with factorial time/space complexity. In this work we, for the first time, obtain decomposition matrices for ICTs up to rank $n=9$ with reduced and affordable complexity, ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Fast Algorithm for Constrained Linear Inverse Problems

Дата публикации: 26-07-2026 20:25:00


We consider the constrained Linear Inverse Problem (LIP), where a certain atomic norm (like the $\ell_1 $ norm) is minimized subject to a quadratic constraint. Typically, such cost functions are non-differentiable, which makes them not amenable to the fast optimization methods existing in practice. We propose two equivalent reformulations of the constrained LIP with improved convex regularity: (i) a smooth convex minimization problem, and (ii) a strongly convex min-max problem. These problems could be solved by applying existing acceleration-based convex optimization methods which provide better $ O \left( \frac{1}{k^2} \right)$ theoretical convergence guarantee, improving upon the current best rate of ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Four Axiomatic Characterizations of the Integrated Gradients Attribution Method

Дата публикации: 26-07-2026 20:25:00


Deep neural networks have produced significant progress among machine learning models in terms of accuracy and functionality, but their inner workings are still largely unknown. Attribution methods seek to shine a light on these "black box" models by indicating how much each input contributed to a model's outputs. The Integrated Gradients (IG) method is a state of the art baseline attribution method in the axiomatic vein, meaning it is designed to conform to particular principles of attributions. We present four axiomatic characterizations of IG, establishing IG as the unique method satisfying four different sets of axioms.

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Bagged Regularized k-Distances for Anomaly Detection

Дата публикации: 26-07-2026 20:25:00


We consider the paradigm of unsupervised anomaly detection, which involves the identification of anomalies within a dataset in the absence of labeled examples. Though distance-based methods are top-performing for unsupervised anomaly detection, they suffer heavily from the sensitivity to the choice of the number of the nearest neighbors. In this paper, we propose a new distance-based algorithm called bagged regularized $k$-distances for anomaly detection (BRDAD), converting the unsupervised anomaly detection problem into a convex optimization problem. Our BRDAD algorithm selects the weights by minimizing the surrogate risk, i.e., the finite sample bound of the empirical risk of the bagged weighted ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Assumption-lean and data-adaptive post-prediction inference

Дата публикации: 26-07-2026 20:25:00


A primary challenge facing modern scientific research is the limited availability of gold-standard data, which can be costly, labor-intensive, or invasive to obtain. With the rapid development of machine learning (ML), scientists can now employ ML algorithms to predict gold-standard outcomes using variables that are easier to obtain. However, these predicted outcomes are often used directly in subsequent statistical analyses, ignoring imprecision and heterogeneity introduced by the prediction procedure. This will likely result in false positive findings and invalid scientific conclusions. In this work, we introduce PoSt-Prediction Adaptive inference (PSPA) that allows valid and powerful inference based on ML-predicted data. ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

“What is Different Between These Datasets?” A Framework for Explaining Data Distribution Shifts

Дата публикации: 26-07-2026 20:25:00


The performance of machine learning models relies heavily on the quality of input data, yet real-world applications often face significant data-related challenges. A common issue arises when curating training data or deploying models: two datasets from the same domain may exhibit differing distributions. While many techniques exist for detecting such distribution shifts, there is a lack of comprehensive methods to explain these differences in a human-understandable way beyond opaque quantitative metrics. To bridge this gap, we propose a versatile framework of interpretable methods for comparing datasets. Using a variety of case studies, we demonstrate the effectiveness of our approach across ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Generative Adversarial Networks: Dynamics

Дата публикации: 26-07-2026 20:25:00


We study quantitatively the overparametrization limit of the original Wasserstein-GAN algorithm. Effectively, we show that the algorithm is a stochastic discretization of a system of continuity equations for the parameter distributions of the generator and discriminator. We show that parameter clipping to satisfy the Lipschitz condition in the algorithm induces a discontinuous vector field in the mean field dynamics, which gives rise to blow-up in finite time of the mean field dynamics. We look into a specific toy example that shows that all solutions to the mean field equations converge in the long time limit to time periodic solutions, this ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Hierarchical Decision Making Based on Structural Information Principles

Дата публикации: 26-07-2026 20:25:00


Hierarchical Reinforcement Learning (HRL) is a promising approach for managing task complexity across multiple levels of abstraction and accelerating long-horizon agent exploration. However, the effectiveness of hierarchical policies heavily depends on prior knowledge and manual assumptions about skill definitions and task decomposition. In this paper, we propose a novel Structural Information principles-based framework, namely SIDM, for hierarchical Decision Making in both single-agent and multi-agent scenarios. Central to our work is the utilization of structural information embedded in the decision-making process to adaptively and dynamically discover and learn hierarchical policies through environmental abstractions. Specifically, we present an abstraction mechanism that processes ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Early Alignment in Two-Layer Networks Training is a Two-Edged Sword

Дата публикации: 26-07-2026 20:25:00


Training neural networks with first order optimisation methods is at the core of the empirical success of deep learning. The scale of initialisation is a crucial factor, as small initialisations are generally associated to a feature learning regime, for which gradient descent is implicitly biased towards simple solutions. This work provides a general and quantitative description of the early alignment phase, originally introduced by Maennel et al. (2018). For small initialisation and one hidden ReLU layer networks, the early stage of the training dynamics leads to an alignment of the neurons towards key directions. This alignment induces a sparse representation ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Imprecise Multi-Armed Bandits: Representing Irreducible Uncertainty as a Zero-Sum Game

Дата публикации: 26-07-2026 20:25:00


We introduce a novel multi-armed bandit framework, where each arm is associated with a fixed unknown credal set over the space of outcomes (which can be richer than just the reward). The arm-to-credal-set correspondence comes from a known class of hypotheses. We then define a notion of regret corresponding to the lower prevision defined by these credal sets. Equivalently, the setting can be regarded as a two-player zero-sum game, where, on each round, the agent chooses an arm and the adversary chooses the distribution over outcomes from a set of options associated with this arm. The regret is defined with ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Optimizing Return Distributions with Distributional Dynamic Programming

Дата публикации: 26-07-2026 20:25:00


We introduce distributional dynamic programming (DP) methods for optimizing statistical functionals of the return distribution, with standard reinforcement learning as a special case. Previous distributional DP methods could optimize the same class of expected utilities as classic DP. To go beyond, we combine distributional DP with stock augmentation, a technique previously introduced for classic DP in the context of risk-sensitive RL, where the MDP state is augmented with a statistic of the rewards obtained since the first time step. We find that a number of recently studied problems can be formulated as stock-augmented return distribution optimization, and we show that ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Exponential Family Graphical Models: Correlated Replicates and Unmeasured Confounders, with Applications to fMRI Data

Дата публикации: 26-07-2026 20:25:00


Graphical models have been used extensively for modeling brain connectivity networks. However, unmeasured confounders and correlations among measurements are often overlooked during model fitting, which may lead to spurious scientific discoveries. Motivated by functional magnetic resonance imaging (fMRI) studies, we propose a novel method for constructing brain connectivity networks with correlated replicates and latent effects. In a typical fMRI study, each participant is scanned and fMRI measurements are collected across a period of time. In many cases, subjects may have different states of mind that cannot be measured during the brain scan: for instance, some subjects may be awake during ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Last-iterate Convergence of Shuffling Momentum Gradient Method under the Kurdyka-Lojasiewicz Inequality

Дата публикации: 26-07-2026 20:25:00


Shuffling gradient algorithms are extensively used to solve finite-sum optimization problems in machine learning. However, their theoretical properties still need to be further explored, especially the last-iterate convergence in the non-convex setting. In this paper, we study the last-iterate convergence behavior of shuffling momentum gradient (SMG) method, a shuffling gradient algorithm with momentum. Specifically, we focus on the non-convex scenario and provide theoretical guarantees under arbitrary shuffling strategies. For non-convex objectives, we achieve the convergence of gradient norms at the last-iterate, showing that every accumulation point of the iterative sequence is a stationary point of the non-convex problem. Our analysis ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Physics-informed Kernel Learning

Дата публикации: 26-07-2026 20:25:00


Physics-informed machine learning typically integrates physical priors into the learning process by minimizing a loss function that includes both a data-driven term and a partial differential equation (PDE) regularization. Building on the formulation of the problem as a kernel regression task, we use Fourier methods to approximate the associated kernel, and propose a tractable estimator that minimizes the physics-informed risk function. We refer to this approach as physics-informed kernel learning (PIKL). This framework provides theoretical guarantees, enabling the quantification of the physical prior’s impact on convergence speed. We demonstrate the numerical performance of the PIKL estimator through simulations, both in ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

BitNet: 1-bit Pre-training for Large Language Models

Дата публикации: 26-07-2026 20:25:00


The increasing size of large language models (LLMs) has posed challenges for deployment and raised concerns about environmental impact due to high energy consumption. Previous research typically applies quantization after pre-training. While these methods avoid the need for model retraining, they often cause notable accuracy loss at extremely low bit-widths. In this work, we explore the feasibility and scalability of 1-bit pre-training. We introduce BitNet b1 and BitNet b1.58, the scalable and stable 1-bit Transformer architecture designed for LLMs. Specifically, we introduce BitLinear as a drop-in replacement of the nn.Linear layer in order to train 1-bit weights from scratch. Experimental ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Modelling Populations of Interaction Networks via Distance Metrics

Дата публикации: 26-07-2026 20:25:00


Network data arises through the observation of relational information between a collection of entities, for example, friendships (relations) amongst a sample of people (entities). Traditionally, statistical models of such data have been developed to analyse a single network, that is, a single collection of entities and relations. More recently, attention has shifted to analysing samples of networks. A driving force has been the analysis of connectome data, arising in neuroscience applications, where a single network is observed for each patient in a study. These models typically assume, within each network, the entities are the units of observation, that is, more ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Actor-Critic learning for mean-field control in continuous time

Дата публикации: 26-07-2026 20:25:00


We study policy gradient for mean-field control in continuous time in a reinforcement learning setting. By considering randomised policies with entropy regularisation, we derive a gradient expectation representation of the value function, which is amenable to actor-critic type algorithms, where the value functions and the policies are learnt alternately based on observation samples of the state and model-free estimation of the population state distribution, either by offline or online learning. In the linear-quadratic mean-field framework, we obtain an exact parametrisation of the actor and critic functions defined on the Wasserstein space. Finally, we illustrate the results of our algorithms with ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Optimal Sample Selection Through Uncertainty Estimation and Its Application in Deep Learning

Дата публикации: 26-07-2026 20:25:00


Modern deep learning heavily relies on large labeled datasets, which often comse with high costs in terms of both manual labeling and computational resources. To mitigate these challenges, researchers have explored the use of informative subset selection techniques. In this study, we present a theoretically optimal solution for addressing both sampling with and without labels within the context of linear softmax regression. Our proposed method, COPS (unCertainty based OPtimal Sub-sampling), is designed to minimize the expected loss of a model trained on subsampled data. Unlike existing approaches that rely on explicit calculations of the inverse covariance matrix, which are not ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Transformers from Diffusion: A Unified Framework for Neural Message Passing

Дата публикации: 26-07-2026 20:25:00


Learning representations for structured data with certain geometries (e.g., observed or unobserved) is a fundamental challenge, wherein message passing neural networks (MPNNs) have become a de facto class of model solutions. In this paper, inspired by physical systems, we propose an energy-constrained diffusion model, which integrates the inductive bias of diffusion on manifolds with layer-wise constraints of energy minimization. We identify that the diffusion operators have a one-to-one correspondence with the energy functions implicitly descended by the diffusion process, and the finite-difference iteration for solving the energy-constrained diffusion system induces the propagation layers of various types of MPNNs operating on ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Categorical Semantics of Compositional Reinforcement Learning

Дата публикации: 26-07-2026 20:25:00


Compositional knowledge representations in reinforcement learning (RL) facilitate modular, interpretable, and safe task specifications. However, generating compositional models requires the characterization of minimal assumptions for the robustness of the compositionality feature, especially in the case of functional decompositions. Using a categorical point of view, we develop a knowledge representation framework for a compositional theory of RL. Our approach relies on the theoretical study of the category $\mathsf{MDP}$, whose objects are Markov decision processes (MDPs) acting as models of tasks. The categorical semantics models the compositionality of tasks through the application of pushout operations akin to combining puzzle pieces. As a ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

On the O(sqrt(d)/T^(1/4)) Convergence Rate of RMSProp and Its Momentum Extension Measured by l_1 Norm

Дата публикации: 26-07-2026 20:25:00


Although adaptive gradient methods have been extensively used in deep learning, their convergence rates proved in the literature are all slower than that of SGD, particularly with respect to their dependence on the dimension. This paper considers the classical RMSProp and its momentum extension and establishes the convergence rate of $\frac{1}{T}\sum_{k=1}^TE\left[||\nabla f(\mathbf{x}^k)||_1\right]\leq O(\frac{\sqrt{d}C}{T^{1/4}})$ measured by $\ell_1$ norm without the bounded gradient assumption, where $d$ is the dimension of the optimization variable, $T$ is the iteration number, and $C$ is a constant identical to that appeared in the optimal convergence rate of SGD. Our convergence rate matches the lower bound with ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Score-Aware Policy-Gradient and Performance Guarantees using Local Lyapunov Stability

Дата публикации: 26-07-2026 20:25:00


In this paper, we introduce a policy-gradient method for model-based reinforcement learning (RL) that exploits a type of stationary distributions commonly obtained from Markov decision processes (MDPs) in stochastic networks, queueing systems, and statistical mechanics. Specifically, when the stationary distribution of the MDP belongs to an exponential family that is parametrized by policy parameters, we can improve existing policy gradient methods for average-reward RL. Our key identification is a family of gradient estimators, called score-aware gradient estimators (SAGEs), that enable policy gradient estimation without relying on value-function estimation in the aforementioned setting. We show that SAGE-based policy-gradient locally converges, and ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

PREMAP: A Unifying PREiMage APproximation Framework for Neural Networks

Дата публикации: 26-07-2026 20:25:00


Most methods for neural network verification focus on bounding the image, i.e., set of outputs for a given input set. This can be used to, for example, check the robustness of neural network predictions to bounded perturbations of an input. However, verifying properties concerning the preimage, i.e., the set of inputs satisfying an output property, requires abstractions in the input space. We present a general framework for preimage abstraction that produces under- and over-approximations of any polyhedral output set. Our framework employs cheap parameterised linear relaxations of the neural network, together with an anytime refinement procedure that iteratively partitions the ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Characterizing Dynamical Stability of Stochastic Gradient Descent in Overparameterized Learning

Дата публикации: 26-07-2026 20:25:00


For overparameterized optimization tasks, such as those found in modern machine learning, global minima are generally not unique. In order to understand generalization in these settings, it is vital to study to which minimum an optimization algorithm converges. The possibility of having minima that are unstable under the dynamics imposed by the optimization algorithm limits the potential minima that the algorithm can find. In this paper, we characterize the global minima that are dynamically stable/unstable for both deterministic and stochastic gradient descent (SGD). In particular, we introduce a characteristic Lyapunov exponent that depends on the local dynamics around a global ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Optimal and Efficient Algorithms for Decentralized Online Convex Optimization

Дата публикации: 26-07-2026 20:25:00


We investigate decentralized online convex optimization (D-OCO), in which a set of local learners are required to minimize a sequence of global loss functions using only local computations and communications. Previous studies have established $O(n^{5/4}\rho^{-1/2}\sqrt{T})$ and ${O}(n^{3/2}\rho^{-1}\log T)$ regret bounds for convex and strongly convex functions respectively, where $n$ is the number of local learners, $\rho

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Minimax Optimal Deep Neural Network Classifiers Under Smooth Decision Boundary

Дата публикации: 26-07-2026 20:25:00


Deep learning has gained huge empirical successes in large-scale classification problems. In contrast, there is a lack of statistical understanding about deep learning methods, particularly in the minimax optimality perspective. For instance, in the classical smooth decision boundary setting, existing deep neural network (DNN) approaches are rate-suboptimal, and it remains elusive how to construct minimax optimal DNN classifiers. Moreover, it is interesting to explore whether DNN classifiers can circumvent the "curse of dimensionality" in handling high-dimensional data. The contributions of this paper are two-fold. First, based on a localized margin framework, we discover the source of suboptimality of existing DNN ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Randomly Projected Convex Clustering Model: Motivation, Realization, and Cluster Recovery Guarantees

Дата публикации: 26-07-2026 20:25:00


In this paper, we propose a randomly projected convex clustering model for clustering a collection of $n$ high dimensional data points in $\mathbb{R}^d$ with $K$ hidden clusters. Compared to the convex clustering model for clustering original data with dimension $d$, we prove that, under some mild conditions, the perfect recovery of the cluster membership assignments of the convex clustering model, if exists, can be preserved by the randomly projected convex clustering model with embedding dimension $m = O(\epsilon^{-2}\log(n))$, where $\epsilon > 0$ is some given parameter. We further prove that the embedding dimension can be improved to be $O(\epsilon^{-2}\log(K))$, which ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Finite Expression Method for Solving High-Dimensional Partial Differential Equations

Дата публикации: 26-07-2026 20:25:00


Designing efficient and accurate numerical solvers for high-dimensional partial differential equations (PDEs) remains a challenging and important topic in computational science and engineering, mainly due to the "curse of dimensionality" in designing numerical schemes that scale in dimension. This paper introduces a new methodology that seeks an approximate PDE solution in the space of functions with finitely many analytic expressions and, hence, this methodology is named the finite expression method (FEX). It is proved in approximation theory that FEX can avoid the curse of dimensionality. As a proof of concept, a deep reinforcement learning method is proposed to implement FEX ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Diffeomorphism-based feature learning using Poincaré inequalities on augmented input space

Дата публикации: 26-07-2026 20:25:00


We propose a gradient-enhanced algorithm for high-dimensional function approximation.
The algorithm proceeds in two steps: firstly, we reduce the input dimension by learning the relevant input features from gradient evaluations, and secondly, we regress the function output against the pre-learned features. To ensure theoretical guarantees, we construct the feature map as the first components of a diffeomorphism, which we learn by minimizing an error bound obtained using Poincaré Inequality applied either in the input space or in the feature space. This leads to two different strategies, which we compare both theoretically and numerically and relate to existing methods in the literature.
In ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Deep Variational Multivariate Information Bottleneck - A Framework for Variational Losses

Дата публикации: 26-07-2026 20:25:00


Variational dimensionality reduction methods are widely used for their accuracy, generative capabilities, and robustness. We introduce a unifying framework that generalizes both such as traditional and state-of-the-art methods. The framework is based on an interpretation of the multivariate information bottleneck, trading off the information preserved in an encoder graph (defining what to compress) against that in a decoder graph (defining a generative model for data). Using this approach, we rederive existing methods, including the deep variational information bottleneck, variational autoencoders, and deep multiview information bottleneck. We naturally extend the deep variational CCA (DVCCA) family to beta-DVCCA and introduce a new ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Conditional Wasserstein Distances with Applications in Bayesian OT Flow Matching

Дата публикации: 26-07-2026 20:25:00


In inverse problems, many conditional generative models approximate the posterior measure by minimizing a distance between the joint measure and its learned approximation. While this approach also controls the distance between the posterior measures in the case of the Kullback–Leibler divergence, the same in general does not hold true for the Wasserstein distance. In this paper, we introduce a conditional Wasserstein distance via a set of restricted couplings that equals the expected Wasserstein distance of the posteriors. Interestingly, the dual formulation of the conditional Wasserstein-1 distance resembles losses in the conditional Wasserstein GAN literature in a quite natural way. We ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

ClimSim-Online: A Large Multi-Scale Dataset and Framework for Hybrid Physics-ML Climate Emulation

Дата публикации: 26-07-2026 20:25:00


Modern climate projections lack adequate spatial and temporal resolution due to computational constraints, leading to inaccuracies in representing critical processes like thunderstorms that occur on the sub-resolution scale. Hybrid methods combining physics with machine learning (ML) offer faster, higher fidelity climate simulations by outsourcing compute-hungry, high-resolution simulations to ML emulators. However, these hybrid physics-ML simulations require domain-specific data and workflows that have been inaccessible to many ML experts. This paper is an extended version of our NeurIPS award-winning ClimSim dataset paper. The ClimSim dataset includes 5.7 billion pairs of multivariate input/output vectors spanning ten years at high temporal resolution, capturing ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Deep Generative Models: Complexity, Dimensionality, and Approximation

Дата публикации: 26-07-2026 20:25:00


Generative networks have shown remarkable success in learning complex data distributions, particularly in generating high-dimensional data from lower-dimensional inputs. While this capability is well-documented empirically, its theoretical underpinning remains unclear. One common theoretical explanation appeals to the widely accepted manifold hypothesis, which suggests that many real-world datasets, such as images and signals, often possess intrinsic low-dimensional geometric structures. Under this manifold hypothesis, it is widely believed that to approximate a distribution on a $d$-dimensional Riemannian manifold, the latent dimension needs to be at least $d$ or $d+1$. In this work, we show that this requirement on the latent dimension is ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Fine-grained Analysis and Faster Algorithms for Iteratively Solving Linear Systems

Дата публикации: 26-07-2026 20:25:00


Despite being a key bottleneck in many machine learning tasks, the cost of solving large linear systems has proven challenging to quantify due to problem-dependent quantities such as condition numbers.
To tackle this, we
consider a fine-grained notion of complexity for solving linear systems, which is motivated by applications where the data exhibits low-dimensional structure, including spiked covariance models and kernel machines, and when the linear system is explicitly regularized, such as ridge regression.

Concretely, let $\kappa_\ell$ be the ratio between the $\ell$th largest and the smallest singular value of $n\times n$ matrix $A$.
We give a stochastic algorithm based ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

On the Ability of Deep Networks to Learn Symmetries from Data: A Neural Kernel Theory

Дата публикации: 26-07-2026 20:25:00


Symmetries (transformations by group actions) are present in many datasets, and leveraging them holds considerable promise for improving predictions in machine learning. In this work, we aim to understand when and how deep networks---with standard architectures trained in a standard, supervised way---learn symmetries from data. Inspired by real-world scenarios, we study a classification paradigm where data symmetries are only partially observed during training: some classes include all transformations of a cyclic group, while others---only a subset. We ask: under which conditions will deep networks correctly classify the partially sampled classes?
In the infinite-width limit, where neural networks behave like kernel machines, ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Dynamic Bayesian Learning for Spatiotemporal Mechanistic Models

Дата публикации: 26-07-2026 20:25:00


We develop an approach for Bayesian learning of spatiotemporal dynamical mechanistic models. Such learning consists of statistical emulation of the mechanistic system that can efficiently interpolate the output of the system from arbitrary inputs. The emulated learner can then be used to train the system from noisy data achieved by melding information from observed data with the emulated mechanistic system. This joint melding of mechanistic systems employ hierarchical state-space models with Gaussian process regression. Assuming the dynamical system is controlled by a finite collection of inputs, Gaussian process regression learns the effect of these parameters through a number of training ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Latent Process Models for Functional Network Data

Дата публикации: 26-07-2026 20:25:00


Network data are often sampled with auxiliary information or collected through the observation of a complex system over time, leading to multiple network snapshots indexed by a continuous variable. Many methods in statistical network analysis are traditionally designed for a single network, and can be applied to an aggregated network in this setting, but that approach can miss important functional structure. Here we develop an approach to estimating the expected network explicitly as a function of a continuous index, be it time or another indexing variable. We parameterize the network expectation through low dimensional latent processes, whose components we represent ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Losing Momentum in Continuous-time Stochastic Optimisation

Дата публикации: 26-07-2026 20:25:00


The training of modern machine learning models often consists in solving high-dimensional non-convex optimisation problems that are subject to large-scale data. In this context, momentum-based stochastic optimisation algorithms have become particularly widespread. The stochasticity arises from data subsampling which reduces computational cost. Both, momentum and stochasticity help the algorithm to converge globally. In this work, we propose and analyse a continuous-time model for stochastic gradient descent with momentum. This model is a piecewise-deterministic Markov process that represents the optimiser by an underdamped dynamical system and the data subsampling through a stochastic switching. We investigate longtime limits, the subsampling-to-no-subsampling limit, and ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

skglm: Improving scikit-learn for Regularized Generalized Linear Models

Дата публикации: 26-07-2026 20:25:00


We introduce skglm, an open-source Python package for regularized Generalized Linear Models. Thanks to its composable nature, it supports combining datafits, penalties, and solvers to fit a wide range of models, many of them not included in scikit-learn (e.g. Group Lasso and variants). It uses state-of-the-art algorithms to solve problems involving high-dimensional datasets, providing large speed-ups compared to existing implementations. It is fully compliant with the scikit-learn API and acts as a drop-in replacement for its estimators. Finally, it abides by the standards of open source development and is integrated in the scikit-learn-contrib GitHub organization.

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Randomization Can Reduce Both Bias and Variance: A Case Study in Random Forests

Дата публикации: 26-07-2026 20:25:00


We study the often overlooked phenomenon, first noted in Breiman (2001), that random forests appear to reduce bias compared to bagging. Motivated by an interesting paper by Mentch and Zhou (2020), where the authors explain the success of random forests in low signal-to-noise ratio (SNR) settings through regularization, we explore how random forests can capture patterns in the data that bagging ensembles fail to capture. We empirically demonstrate that in the presence of such patterns, random forests reduce bias along with variance and can increasingly outperform bagging ensembles when SNR is high. Our observations offer insights into the real-world success ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Sample Complexity of the Linear Quadratic Regulator: A Reinforcement Learning Lens

Дата публикации: 26-07-2026 20:25:00


We provide the first known algorithm that provably achieves $\varepsilon$-optimality within $\widetilde{O}(1/\varepsilon)$ function evaluations for the discounted discrete-time linear quadratic regulator problem with unknown parameters, without relying on two-point gradient estimates. These estimates are known to be unrealistic in many settings, as they depend on using the exact same initialization, which is to be selected randomly, for two different policies. Our results substantially improve upon the existing literature outside the realm of two-point gradient estimates, which either leads to $\widetilde{O}(1/\varepsilon^2)$ rates or heavily relies on stability assumptions.

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Universal Online Convex Optimization Meets Second-order Bounds

Дата публикации: 26-07-2026 20:25:00


Recently, several universal methods have been proposed for online convex optimization, and attain minimax rates for multiple types of convex functions simultaneously. However, they need to design and optimize one surrogate loss for each type of functions, making it difficult to exploit the structure of the problem and utilize existing algorithms. In this paper, we propose a simple strategy for universal online convex optimization, which avoids these limitations. The key idea is to construct a set of experts to process the original online functions, and deploy a meta-algorithm over the linearized losses to aggregate predictions from experts. Specifically, the meta-algorithm ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Classification in the high dimensional Anisotropic mixture framework: A new take on Robust Interpolation

Дата публикации: 26-07-2026 20:25:00


We study the classification problem under the two-component anisotropic sub-Gaussian mixture model in high dimensions and in the non-asymptotic setting. First, we derive lower bounds and matching upper bounds for the minimax risk of classification in this framework. We also show that in the high-dimensional regime, the linear discriminant analysis classifier turns out to be sub-optimal in the minimax sense. Next, we give precise characterization of the risk of classifiers based on solutions of $\ell_2$-regularized least squares problem. We deduce that the interpolating solutions may outperform the regularized classifiers under mild assumptions on the covariance structure of the noise, and ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Bayesian Scalar-on-Image Regression with a Spatially Varying Single-layer Neural Network Prior

Дата публикации: 26-07-2026 20:25:00


Deep neural networks (DNN) have been widely used in scalar-on-image regression to predict an outcome variable from imaging predictors. However, training DNN typically requires large sample sizes for accurate prediction, and the resulting models often lack interpretability. In this work, we propose a novel Bayesian nonlinear scalar-on-image regression framework with a spatially varying single-layer neural network (SV-NN) prior. The SV-NN is constructed using a single hidden layer neural network with its weights generated by the soft-thresholded Gaussian process. Our framework enables the selection of interpretable image regions while achieving high prediction accuracy with limited training samples. The SV-NN offers large ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

On Model Identification and Out-of-Sample Prediction of PCR with Applications to Synthetic Controls

Дата публикации: 26-07-2026 20:25:00


We analyze principal component regression (PCR) in a high-dimensional error-in-variables setting with fixed design. Under suitable conditions, we show that PCR consistently identifies the unique model with minimum $\ell_2$-norm. These results enable us to establish non-asymptotic out-of-sample prediction guarantees that improve upon the best known rates. In the course of our analysis, we introduce a natural linear algebraic condition between the in- and out-of-sample covariates, which allows us to avoid distributional assumptions for out-of-sample predictions. Our simulations illustrate the importance of this condition for generalization, even under covariate shifts. Accordingly, we construct a hypothesis test to check when this condition ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Sparse SVM with Hard-Margin Loss: a Newton-Augmented Lagrangian Method in Reduced Dimensions

Дата публикации: 26-07-2026 20:25:00


The hard-margin loss function has been at the core of the support vector machine research from the very beginning due to its generalization capability. On the other hand, the cardinality constraint has been widely used for feature selection, leading to sparse solutions. This paper studies the sparse SVM with the hard-margin loss that integrates the virtues of both worlds, resulting in one of the most challenging models to solve. We cast the problem as a composite optimization with the cardinality constraint. We characterize its local minimizers in terms of pseudo KKT point that well captures the combinatorial structure of the ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Quantifying the Effectiveness of Linear Preconditioning in Markov Chain Monte Carlo

Дата публикации: 26-07-2026 20:25:00


We study linear preconditioning in Markov chain Monte Carlo. We consider the class of well-conditioned distributions, for which several mixing time bounds depend on the condition number $\kappa$. First we show that well-conditioned distributions exist for which $\kappa$ can be arbitrarily large and yet no linear preconditioner can reduce it. We then impose two sets of extra assumptions under which a linear preconditioner can significantly reduce $\kappa$. For the random walk Metropolis we further provide upper and lower bounds on the spectral gap with tight $1/\kappa$ dependence. This allows us to give conditions under which linear preconditioning can provably increase ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Degree of Interference: A General Framework For Causal Inference Under Interference

Дата публикации: 26-07-2026 20:25:00


One core assumption typically adopted for valid causal inference is that of no interference between experimental units, i.e., the outcome of an experimental unit is unaffected by the treatments assigned to other experimental units. This assumption can be violated in real-life experiments, which significantly complicates the task of causal inference. As the number of potential outcomes increases, it becomes challenging to disentangle direct treatment effects from “spillover” effects. Current methodologies are lacking, as they cannot handle arbitrary, unknown interference structures to permit inference on causal estimands. We present a general framework to address the limitations of existing approaches. Our framework ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Maximum Causal Entropy IRL in Mean-Field Games and GNEP Framework for Forward RL

Дата публикации: 26-07-2026 20:25:00


This paper explores the use of Maximum Causal Entropy Inverse Reinforcement Learning (IRL) within the context of discrete-time stationary Mean-Field Games (MFGs) characterized by finite state spaces and an infinite-horizon, discounted-reward setting. Although the resulting optimization problem is non-convex with respect to policies, we reformulate it as a convex optimization problem in terms of state-action occupation measures by leveraging the linear programming framework of Markov Decision Processes. Based on this convex reformulation, we introduce a gradient descent algorithm with a guaranteed convergence rate to efficiently compute the optimal solution. Moreover, we develop a new method that conceptualizes the MFG problem ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

Posterior and Variational Inference for Deep Neural Networks with Heavy-Tailed Weights

Дата публикации: 26-07-2026 20:25:00


We consider deep neural networks in a Bayesian framework with a prior distribution sampling the network weights at random. Following a recent idea of Agapiou and Castillo (2024), who show that heavy-tailed prior distributions achieve automatic adaptation to smoothness, we introduce a simple Bayesian deep learning prior based on heavy-tailed weights and ReLU activation. We show that the corresponding posterior distribution achieves near-optimal minimax contraction rates, simultaneously adaptive to both intrinsic dimension and smoothness of the underlying function, in a variety of contexts including nonparametric regression, geometric data and Besov spaces. While most works so far need a form of ...

Классификация: . Схожих патентов: 0. Схожих новостей: 0. Тональность: 0. Информативность: 0. Источник: jmlr.org.

В Ростовской области вернули электричество 60% оставшихся без света абонентов

Дата публикации: 26-07-2026 20:24:06

Непогода повредила линии электропередачи

Классификация: Общество. Схожих патентов: 0. Схожих новостей: 9. Тональность: 0. Информативность: 17.14. Источник: www.itar-tass.com.

Найдено тело художницы, которую унесли потоки воды в Пермском крае

Дата публикации: 26-07-2026 20:22:05

Найдено тело девушки, которую 10 июля в пермском селе Кын унесли потоки воды. Об этом сообщили РИА Новости в краевом СУ СКР.

Классификация: Происшествия. Схожих патентов: 0. Схожих новостей: 9. Тональность: 0. Информативность: 7.71. Источник: rg.ru.

"Война с Европой". СМИ пришли в ужас от дерзкого решения Зеленского

Дата публикации: 26-07-2026 20:20:39

Отставка Михаила Федорова с поста главы Минобороны Украины стала началом открытого противостояния между Владимиром Зеленским и Европой, пишет турецкое издание dikGazete.

Классификация: Международные. Схожих патентов: 0. Схожих новостей: 10. Тональность: 0. Информативность: 8.79. Источник: ria.ru.

Russia’s Ilyumzhinov not to vie for FIDE presidency

Дата публикации: 26-07-2026 20:20:37

The candidates are Jan Henric Buettner and Wadim Rosenstein of Germany, and Timur Turlov of Kazakhstan

Классификация: Информация. Схожих патентов: 0. Схожих новостей: 9. Тональность: 0. Информативность: 8.4. Источник: tass.com.

Медведев иронично отреагировал на угрозы Каллас

Дата публикации: 26-07-2026 20:20:29

Зампред Совбеза России Дмитрий Медведев иронично отреагировал на заявление главы евродипломатии Каи Каллас о возможном запрете на въезд в Евросоюз для ветеранов специальной военной операции (СВО). Свой комментарий он опубликовал в канале ...

Классификация: Политика. Схожих патентов: 0. Схожих новостей: 9. Тональность: 0. Информативность: 5.82. Источник: www.gazeta.ru.

Экипаж МКС с иркутянином на борту успешно приземлился в Казахстане.

Дата публикации: 26-07-2026 20:20:20

В Казахстане поисково-спасательная группа Роскосмоса встретила [кого/что], после чего специалисты провели необходимые процедуры, сообщается в соцсетях госкорпорации.

Классификация: Космос. Схожих патентов: 0. Схожих новостей: 9. Тональность: 1. Информативность: 11.19. Источник: vk.com.

SA-202 или Apollo 2 . AS-202 (также известный как SA-202 ...

Дата публикации: 26-07-2026 20:20:02

SA-202 или Apollo 2 .
AS-202 (также известный как SA-202 или Apollo 2 ) был вторым беспилотным суборбитальным испытательным полетом серийного командно-служебного модуля Block I Apollo, запущенного с помощью ракеты-носителя Saturn IB. Он был запущен 25 августа 1966 года и стал первым полетом, в ходе которого были продемонстрированы система наведения космического корабля, навигационная система управления и топливные элементы.
На фотографиях:
- Запуск AS-202;
- AS-202 CM-011 выставлен на авианосце USS Hornet.
- Интерьер AS-202 CM-011
- марка Румынии на марке Аполлон-2 или SA-202

Классификация: Космос. Схожих патентов: 0. Схожих новостей: 10. Тональность: 0. Информативность: 11.33. Источник: vk.com.

СМИ выяснили, почему США позволяют некоторым иранским ракетам достигать цели

Дата публикации: 26-07-2026 20:19:27

Как сообщает NBC News со ссылкой на два высокопоставленных источника в администрации США, американское военное командование сознательно выбирает, какие иранские ракеты и беспилотники сбивать, позволяя некоторым атакам достигать целей

Классификация: Международные. Схожих патентов: 0. Схожих новостей: 9. Тональность: 0. Информативность: 7.6. Источник: www.mk.ru.

В аэропорту Пензы ввели временные ограничения на полёты

Дата публикации: 26-07-2026 20:19:06


Временные ограничения на приём и выпуск воздушных судов введены в аэропорту Пензы. Об этом сообщила Росавиация. Читать далее

Классификация: Общество. Схожих патентов: 0. Схожих новостей: 10. Тональность: 0. Информативность: 6.05. Источник: russian.rt.com.

Отечественный гигабит по всей Земле: Бюро 1440 запустило второй пакет ...

Дата публикации: 26-07-2026 20:18:15

Отечественный гигабит по всей Земле: Бюро 1440 запустило второй пакет спутников
«Российский Starlink» отчитался об успешном запуске 19 июля. Количество выведенных на орбиту спутников не сообщается, однако все они успешно отделились от ракета-носителя и принимают команды с ЦУП. В ближайшее время они пройдут тесты и займут свои орбиты.
Первый запуск спутников проекта «Рассвет» состоялся в марте 2026 года – тогда на орбиту вывели 16 аппаратов с поддержкой 5G NTN (Non-Terrestrial Networks) и межспутниковой лазерной связью, а также плазменными двигателями для маневрирования. Для базового покрытия по всей Земле требуется 300 спутников, для получения гигабита – 950.
[club42956320|ISAKIN - техноблогер]

Классификация: Космос. Схожих патентов: 0. Схожих новостей: 9. Тональность: 1. Информативность: 11.4. Источник: vk.com.

Москва поможет регионам с внедрением успешных проектов

Дата публикации: 26-07-2026 20:18:04


Москва готова делиться опытом с другими субъектами РФ, заявил замглавы департамента инвестиционной и промышленной политики Владимир Петраков. Восемь регионов завершили разработку индивидуальных дорожных карт и начинают их реализацию. Это итог программы правительства Москвы, АСИ, РАНХиГС и АНО «Общественный капитал», запущенной в 2026 году.

Классификация: Национальные проекты. Схожих патентов: 0. Схожих новостей: 9. Тональность: 0. Информативность: 9.9. Источник: www.ferra.ru.