ML-MTP is a series of seminars aiming to federate the Machine Learning community around Montpellier, hosting talks from local researchers as well as national and international experts.
Inria Montpellier, St-Priest Campus, Building 2, Room 167
Paul-Gauthier Noé LIS – CNRS/ Aix Marseille Université
While calibration of probabilistic predictions has been widely studied, we will rather discuss calibration of likelihood functions. This has been studied, especially in biometrics, in cases with only two exhaustive and mutually exclusive hypotheses (or classes): where likelihood functions can be written as log-likelihood-ratios (LLRs). After defining calibration for LLRs and its connection with the concept of weight-of-evidence, I will present the idempotence property and its associated constraint on the distribution of the LLRs. Although these results have been known for decades, they have been limited to the binary case. In this talk, we will see how the Aitchison geometry of the simplex allows us to extend these results to cases with more than two hypotheses. To be more precise, it recovers, in a vector form, the additive form of the Bayes’ rule; extending therefore the LLR and the weight-of-evidence to any number of hypotheses. Especially, we will extend the definition of calibration, the idempotence, and the constraint on the distribution of likelihood functions to this multiple hypotheses and multiclass counterpart of the LLR: the isometric-log-ratio transformed likelihood function. Even if this work is mainly conceptual, we will discuss one application to machine learning by presenting a non-linear discriminant analysis where the discriminant components form a calibrated likelihood function over the classes, improving therefore the interpretability and the reliability of the method.
18
Juin
2026
11
Juin
2026
17
Fév
2026
10
Fév
2026
03
Fév
2026
14
Jan
2026
14
Jan
2026
13
Jan
2026
13
Nov
2025
13
Nov
2025
04
Nov
2025
23
Oct
2025
09
Oct
2025
02
Oct
2025
03
Juil
2025
03
Juil
2025
24
Juin
2025
19
Juin
2025
12
Juin
2025
10
Juin
2025
27
Mai
2025
19
Mai
2025
16
Mai
2025
05
Mai
2025
14
Avr
2025
10
Avr
2025
07
Avr
2025
04
Avr
2025
03
Avr
2025
01
Avr
2025
31
Mar
2025
28
Mar
2025
26
Mar
2025
25
Mar
2025
13
Mar
2025
13
Mar
2025
10
Mar
2025
06
Mar
2025
24
Fév
2025
17
Fév
2025
10
Fév
2025
03
Fév
2025
23
Jan
2025
13
Déc
2024
12
Déc
2024
09
Déc
2024
02
Déc
2024
21
Nov
2024
19
Nov
2024
18
Nov
2024
14
Nov
2024
12
Nov
2024
21
Oct
2024
17
Oct
2024
08
Oct
2024
01
Oct
2024
24
Sep
2024
19
Sep
2024
17
Sep
2024
16
Sep
2024
12
Sep
2024
10
Sep
2024
10
Sep
2024
05
Sep
2024
27
Juin
2024
13
Juin
2024
23
Mai
2024
21
Mai
2024
16
Mai
2024
02
Mai
2024
29
Avr
2024
25
Avr
2024
22
Avr
2024
11
Avr
2024
02
Avr
2024
21
Mar
2024
21
Mar
2024
18
Mar
2024
11
Mar
2024
07
Mar
2024
29
Fév
2024
26
Fév
2024
12
Déc
2023
04
Déc
2023
30
Nov
2023
21
Nov
2023
09
Nov
2023
12
Oct
2023
28
Sep
2023
28
Sep
2023
28
Sep
2023
28
Sep
2023
15
Sep
2023
15
Sep
2023
15
Sep
2023
15
Sep
2023
14
Sep
2023
15
Juin
2023
08
Juin
2023
01
Juin
2023
25
Mai
2023
27
Avr
2023
20
Avr
2023
30
Mar
2023
23
Mar
2023
16
Mar
2023
23
Fév
2023
16
Fév
2023
09
Fév
2023
08
Déc
2022
01
Déc
2022
24
Nov
2022
10
Nov
2022
20
Oct
2022
13
Oct
2022
29
Sep
2022
22
Sep
2022
15
Sep
2022
09
Juin
2022
12
Mai
2022
28
Avr
2022
21
Avr
2022
14
Avr
2022
07
Avr
2022
24
Fév
2022
17
Fév
2022
10
Fév
2022
03
Fév
2022
27
Jan
2022
13
Jan
2022
06
Jan
2022
09
Déc
2021
02
Déc
2021
25
Nov
2021
28
Oct
2021
14
Oct
2021
23
Sep
2021
01
Juil
2021
17
Juin
2021
20
Mai
2021
29
Avr
2021
15
Avr
2021
01
Avr
2021
01
Avr
2021
18
Mar
2021
04
Mar
2021
18
Fév
2021
04
Fév
2021
21
Jan
2021
07
Jan
2021
Inria Montpellier, St-Priest Campus, Building 2, Room 167
Paul-Gauthier Noé LIS – CNRS/ Aix Marseille Université
While calibration of probabilistic predictions has been widely studied, we will rather discuss calibration of likelihood functions. This has been studied, especially in biometrics, in cases with only two exhaustive and mutually exclusive hypotheses (or classes): where likelihood functions can be written as log-likelihood-ratios (LLRs). After defining calibration for LLRs and its connection with the concept of weight-of-evidence, I will present the idempotence property and its associated constraint on the distribution of the LLRs. Although these results have been known for decades, they have been limited to the binary case. In this talk, we will see how the Aitchison geometry of the simplex allows us to extend these results to cases with more than two hypotheses. To be more precise, it recovers, in a vector form, the additive form of the Bayes’ rule; extending therefore the LLR and the weight-of-evidence to any number of hypotheses. Especially, we will extend the definition of calibration, the idempotence, and the constraint on the distribution of likelihood functions to this multiple hypotheses and multiclass counterpart of the LLR: the isometric-log-ratio transformed likelihood function. Even if this work is mainly conceptual, we will discuss one application to machine learning by presenting a non-linear discriminant analysis where the discriminant components form a calibrated likelihood function over the classes, improving therefore the interpretability and the reliability of the method.
Inria Montpellier, St-Priest Campus, Building 5, Room 03.124
Pierre-Alexandre Mattei Inria (Maasai)
Ensemble methods combine predictions from various statistical learning models. Their most famous representatives are random forests or deep ensembles. This talk will center around the question: « How many models should I aggregate? »
We will see that the answer depends on the chosen performance metric. Specifically, in the case of convex losses (such as cross-entropy in classification or mean squared error in regression), the error is a decreasing function of the number of models. In the case of non-convex losses (such as classification error in classification or the Fréchet Inception distance in generative modelling), things are more nuanced, and the error can sometimes be non-monotonic.
These results will be illustrated with examples of neural network ensembles, both for classification and generative modelling. This work is notably based on the papers:
– Are Ensembles Getting Better All the Time? (with Damien Garreau), JMLR 2025
– When Are Two Scores Better Than One? Investigating Ensembles of Diffusion Models (with Raphaël Razafindralambo, Rémy Sun, Frédéric Precioso, and Damien Garreau), TMLR 2026
– Beyond Mixtures and Products for Ensemble Aggregation: A Likelihood Perspective on Generalized Means (with Raphaël Razafindralambo, Rémy Sun, Frédéric Precioso, and Damien Garreau), arXiv 2026