Archive for arXiv

Bayesian Adversarial Privacy [v2]

Posted in Books, Statistics, University life with tags , , , , , , , , , , , , , , , , , , , , , , , , on September 11, 2026 by xi'an

We have just reposted our paper Bayesian Adversarial Privacy on arXiv to reflect the revision we wrote in the past months, to address the (quite sensible) comments from the reviewers. Interestingly the discussants of my Akaike lecture made similar points. The main changes are in explaining more clearly the nature of the combined loss, with Antoine coming up with the use of illuminating R-U map representations, in enlarging the references to other approaches, in stressing that Eve was an Alice’s construct rather than a genuine adversary, but still integrating the case of “multiple Eves”, in mellowing our criticisms of DP, and in expanding the conclusion with limitations and extensions subsections.

arXiv, not AIrXiv!

Posted in Books, University life with tags , , , , , , , on July 18, 2026 by xi'an


As reported in Nature of 28 May, arXiv is now enforcing a ban of researchers (who hardly qualify as “authors”!) submitting articles produced by LLMs with demonstrably insufficient human input, e.g. those with hallucinated (fake) references—easy to detect. The ban is one-year long and, besides, subsequent submissions must first be accepted by a genuine academic journal. Be warned!

comments from Bob

Posted in Books, pictures, Statistics, University life with tags , , , , , , , , , , on April 10, 2026 by xi'an

Bob replied to my short post with further items of information that I find worth sharing:

Thanks for the kind post, Christian. It’s amusing to be the subject of one of these posts given how many of them I’ve read about other people. I really appreciate your summaries. And thanks to everyone in the audience for all the great feedback during and after the talk. Here’s a link to my slides.

One of your students or postdocs mentioned an approach that does continuous adaptation on some kind of polynomial schedule that is provably correct, but I didn’t manage to write down the author/reference or the name of the person who recommended it. If you happen to know what that is, I’d be grateful for the reference.

I would also like to follow up on the Robert & Andrieu paper you mention, but I could not find the exact reference on your Google Scholar page. The closest match I can find is:

Controlled MCMC for optimal sampling. 2001. C Andrieu, CP Robert. INSEE.

Section 1.3 is titled “Criteria for local adaptation.” The section cites two things. The first is Haario et al.’s (1999) sliding window approach, for which HMC moves too fast to be useful locally. The second is the multiple try approach of Liu et al. (2000) and the delayed rejection approach Tierny and Mira (1999). We applied delayed rejection to HMC step size adaptation in a couple of papers before developing GIST (Modi, Barnett and Carpenter in Bayesian Analysis; Turok, Modi, and Carpenter in AISTATS); these mirror our second GIST paper and third GIST paper in doing the step size adaptation for a whole trajectory and at each leapfrog step. The GIST approach is easier to understand, easier to describe mathematically, easier to implement, and is more efficient.

The nice part about GIST compared to Riemannian HMC is that we do not need to do any volume adjustments (which must be autodiffed through), which are cubic, and we do not need an implicit integrator, which is incredibly fussy to tune. The tradeoff is the we require reversibility of the adaptation, which I think is going to be tricky with varying curvature. Of course, we can’t afford to compute Hessian matrices in high dimensions, but we could manage Hessian-vector products if we could figure out how to use just those and we could also manage low-rank plus diagonal approximations or sketches as described in the Nutpie paper.

We’ve arXived the Nutpie paper since the talk:

Preconditioning HMC by minimizing Fisher divergence. arXiv. 2026. Seyboldt, Carlsen, and Carpenter.

The WALNUTS paper has been accepted by JMLR, but currently only the arXiv version is available:

The within-orbit adaptive leapfrog no-U-turn sampler. 2026. Nawaf Bou-Rabee, Bob Carpenter, Tore Selland Kleppe, Sifan Liu. 2025 arXiv; 2026 to appear JMLR.

Working with Nawaf and Tore has made all the difference in the world on this—it’s not something I could have done by myself. Sifan’s the one who came up with the nice characterization of NUTS and Nawaf’s done a number of additional things like providing mixing time bounds for NUTS (with Milo Marsden, who’s sadly no longer with us—he’s gone into finance).

Furthermore, you can adjust the U-turn criterion from 180 degrees to whatever you want to control how much of a full orbit you get. Those tend to be even more wasteful of iterations, though—this is what the plot from the expected integration time of NUTS is supposed to show, but it was confusing in the talk.

The approach you took with Wu Chengye to randomize number of leapfrog steps made a deep impression on me. It’s also wasteful in leapfrog steps because any number of steps greater than or less than about 1/4 of an orbit is wasteful either in computation or because it leads to more diffusive sampling. You can see that it is roughly as gradient efficient as NUTS in a 1000-dimensional standard normal. Interestingly, it’s worse than NUTS for parameter estimates and better for squared parameter estimates, which is overall a win. Nawaf has also published on randomized HMC. I think we could turn down NUTS U-turn criterion below 180 degrees to get something similar with NUTS, but I haven’t tried it.

One important property of your randomized approach is that it is much much easier to code efficiently for GPUs than NUTS, because the conditionals in NUTS are hard to execute in SIMD fashion. There’s a very nice introduction to this problem by Sountsov, Carroll, and Hoffman, in their paper “Running Markov Chain Monte Carlo on Modern Hardware and Software,” which is out on arXiv and also going into the next edition of the Handbook of MCMC). The thing to read about how to code NUTS on GPU is Dance, Glaser, Orbanz, and Adams’s paper, “Efficiently Vectorized MCMC on Modern Accelerators,” which is on arXiv and ICML 2025.

You can also randomize step size to vary the integration time and avoid harmonics, e.g.,

Randomized Hamiltonian Monte Carlo. 2017. Bou-Rabee and Sanz-Serna. Annals of Applied Probability.

escaping the dark side of the Moon

Posted in Books, pictures, Statistics, University life with tags , , , , , , , , , , , on March 18, 2026 by xi'an

Sub-Cauchy Sampling: Escaping the Dark Side of the Moon was recently posted on arXiv by Sebastiano Grazzi (Warwick), Sifan Liu, Gareth O. Roberts (Warwick), and Jun Yang. With an hommage to Pink Floyd’s 1973 album both Gareth and I listened to at the time. (This was for sure my first Pink Floyd album!)

This highly original work is a sequel to the stereographic projection paper by Yang, Latuszýnski and Roberts (which was itself vaguely connected to our unpublished origami sampler). As in the stereographic projection method, the Euclidean space supporting the target is turned into a spherical cap of a hyper-sphere, referred to as the complement of the dark side of the Moon (or its bright side), and defined with respect to an observer ο who was at the north pole in the original method. The proposed MCMC algorithm, the Sub-Cauchy Projection Sampler (SCS), is a random-walk-type Metropolis algorithm on the bright side and it gets its name from being uniformly ergodic for sub-Cauchy targets. An explanation for this massive achievement is that points at infinity in the Euclidean space are now mapped to the (d − 1)-dimensional boundary of the dark side rather than at the north pole of the hypersphere. Meaning that the push-forward density may remain bounded. (The random walk on the bright side involves projections for proposals ending on the dark side, while keeping the target intact.) There are several calibration parameters to the algorithm that can be tuned by variational arguments (and the goal of getting near a uniform distribution over the bright side), since optimal acceptance rates no longer apply.

OWABI⁷, 25 March 2026: Robust Simulation Based Inference (10am EST time)

Posted in Books, Statistics, University life with tags , , , , , , , , , , , , , , , on March 9, 2026 by xi'an

Speaker:  Larry Wasserman (Carnegie Mellon University)

Title: Robust Simulation Based Inference
Abstract: Simulation-Based Inference (SBI) is an approach to statistical inference where simulations from an assumed model are used to construct estimators and confidence sets. SBI is often used when the likelihood is intractable and to construct confidence sets that do not rely on asymptotic methods or regularity conditions. Traditional SBI methods assume that the model is correct, but, as always, this can lead to invalid inference when the model is misspecified. This paper introduces robust methods that allow for valid frequentist inference in the presence of model misspecification. We propose a framework where the target of inference is a projection parameter that minimizes a discrepancy between the true distribution and the assumed model. The method guarantees valid inference, even when the model is incorrectly specified and even if the standard regularity conditions fail. Alternatively, we introduce model expansion through exponential tilting as another way to account for model misspecification. We also develop an SBI based goodness-of-fit test to detect model misspecification. Finally, we propose two ideas that are useful in the SBI framework beyond robust inference: an SBI based method to obtain closed form approximations of intractable models and an active learning approach to more efficiently sample the parameter space.
Keywords: Exponential tilting, model misspecification, robust inference, simulation based inference, valid inference.
Reference: Lorenzo Tomaselli, Valérie Ventura, Larry Wasserman. Robust Simulation Based Inference. Preprint at ArXiv:2508.02404