Showing 1–2 of 2 results for author: Taupin, J
-
Improved Trial and Error Learning for Random Games
Authors:
Jérôme Taupin,
Xavier Leturc,
Christophe J. Le Martret
Abstract:
When a game involves many agents or when communication between agents is not possible, it is useful to resort to distributed learning where each agent acts in complete autonomy without any information on the other agents' situations. Perturbation-based algorithms have already been used for such tasks. We propose some improvements based on practical observations to improve the performance of these…
▽ More
When a game involves many agents or when communication between agents is not possible, it is useful to resort to distributed learning where each agent acts in complete autonomy without any information on the other agents' situations. Perturbation-based algorithms have already been used for such tasks. We propose some improvements based on practical observations to improve the performance of these algorithms. We show that the introduction of these changes preserves their theoretical convergence properties towards states that maximize the average reward and improve them in the case where optimal states exist. Moreover, we show that these algorithms can be made robust to the addition of randomness to the rewards, achieving similar convergence guarantees. Finally, we discuss the possibility for the perturbation factor of the algorithm to decrease during the learning process, akin to simulated annealing processes.
△ Less
Submitted 23 September, 2025;
originally announced September 2025.
-
Fermat Distance-to-Measure: a robust Fermat-like metric
Authors:
Jérôme Taupin,
Frédéric Chazal
Abstract:
Given a probability measure with density, Fermat distances and density-driven metrics are conformal transformation of the Euclidean metric that shrink distances in high density areas and enlarge distances in low density areas. Although they have been widely studied and have shown to be useful in various machine learning tasks, they are limited to measures with density (with respect to Lebesgue mea…
▽ More
Given a probability measure with density, Fermat distances and density-driven metrics are conformal transformation of the Euclidean metric that shrink distances in high density areas and enlarge distances in low density areas. Although they have been widely studied and have shown to be useful in various machine learning tasks, they are limited to measures with density (with respect to Lebesgue measure, or volume form on manifold). In this paper, by replacing the density with the Distance-to-Measure, we introduce a new metric, the Fermat Distance-to-Measure, defined for any probability measure in R^d. We derive strong stability properties for the Fermat Distance-to-Measure with respect to the measure and propose an estimator from random sampling of the measure, featuring an explicit bound on its convergence speed.
△ Less
Submitted 3 April, 2025;
originally announced April 2025.