Index
The page numbers follow Appendix B of the PDF and link to the corresponding web location.
A¶
absorbed log-scaling: 152
absorptive matching: 163
action density: 300
additively homogeneous map: 160
ADMM: 298
adversarial
affine
affine-difference cost: 58
affine-invariant metric: 40
Alexandrov solution: 24
algorithmic convergence: 155
alternating
angular covering: 190
antisymmetric flux: 353
antisymmetric velocity: 307
Assouad lemma: 182
atom-preserving map: 75
atomic measure: 306
coordinate ascent: 89
augmented Lagrangian: 315
B¶
Baker-Campbell-Hausdorff formula: 288
Bakry–Émery criterion: 339
balanced interpolation: 312
balanced OT: 139
Banach fixed-point theorem: 206
Banach space: 12
Banach-space embedding: 32
barrier
barycentric
Beckmann
belief propagation: 255
Benamou–Brenier: 135, 137, 272–275, 293, 295, 296, 299, 301, 304, 312
Berry-Esseen
Besov smoothness: 183
Besov space: 184
Bessel function: 110
bias-variance decomposition: 191
bilevel problem: 261
bilinear cost: 264
Birkhoff
block-coordinate
ascent: 244
block-coordinate descent: 257
block-coordinate method: 226
block-quotient norm: 165
block-stationary point: 257
Bobkov-Ledoux formula: 31
Bochner spectral measure: 111
Bochner theorem: 194
Boltzmann entropy: 117
Boltzmann equation: 362
bounded confidence: 384
box constraint: 259
Bregman
Brenier
Brenier–Strassen projection: 270
bridge
Bures
Burg entropy: 114
C¶
c-concavity: 134
c-convexity: 29
c-splitting set: 250
capacity-constrained OT feasibility: 259
capacity-constrained Sinkhorn: 259
capped transportation polytope: 259
Catalan number: 7
causality: 382
CDF formula: 31
centered measure: 65
centroidal Voronoi: 100
centroidal Voronoi tessellation: 98
chain graph: 254
characteristic kernel: 321
Choi matrix: 288
Cholesky factorization: 142
circle
classical convexity: 347
clustering: 384
co-motion function: 251
co-motion map: 251
column constraint: 162
commutator: 288
compact sublevel set: 201
compactness: 308
complete graph: 254
complete metric space: 12
completely positive cone: 196
completely positive kernel: 196
completely positive matrix: 196
complex
composition constraint: 325
computer vision: 57
concave cost: 2
concave envelope: 146
concavity
conditional
conditionally negative definite kernel: 151
conditionally positive kernel: 109
cone
congruence normalization: 288
conic lifting: 203
conjugate moment measure: 382
consensus: 384
Dobrushin contraction: 385
contact set: 29
continuity equation: 137, 272, 274, 275, 293, 295–298, 311, 316, 318, 352, 365–367, 369, 374, 377, 378, 389
continuous
continuous DTW: 290
continuous-depth neural network: 382
contraction ratio: 169
contractive Gaussian projection: 396
controlled transport: 388
convergence
convergence-determining class: 176
convex
convexity
convolution: 75
convolutional Sinkhorn: 142
copositive certificate: 198
copositive matrix: 196
cost
cost-sectional curvature: 29
counterexample: 251
coupling measure: 65
covariance
Cramér condition: 178
Cramer-von Mises distance: 32
created mass: 200
cross-covariance: 393
Csiszár divergence: 328
cumulant
fourth: 178
cycle graph: 254
cyclic
cylindrical Gaussian: 246
D¶
Dacorogna–Moser
damped iteration: 241
damped second-order equation: 360
de la Vallée–Poussin criterion: 177
Delaunay graph: 106
demographic parity: 237
denoising score matching: 365
dense linear system: 119
density
depth evolution: 382
depth variable: 358
destroyed mass: 200
detailed balance: 355
deterministic interpolant: 365
diagonal coupling: 269
diagonal dominance: 162
diagonal isotonicity: 161
diameter
contraction: 385
diffusion model: 74, 75, 365, 369, 372
Gaussian noising: 128
diffusive map: 75
dimension dependence: 182
Dirac mass: 11, 14, 18, 31, 36, 43, 63, 69, 75, 238, 256, 316, 319, 325, 366
Dirac measure: 250
directional derivative: 261
Dirichlet energy: 320
discrepancy minimization: 373
discrete
disintegration: 12, 34, 65, 85, 137, 204, 233, 250, 266–268, 367, 374
displacement convexity: 23, 26, 327, 337, 344, 346, 379, 380
distance profile: 283
distance residual: 282
distributional robustness: 77
divergence-free field: 389
Dobrushin coefficient: 384
domain adaptation: 57
Doob transform: 136
doubly nonnegative cone: 196
doubly nonnegative matrix: 196
doubly positive kernel: 196
doubly stochastic normalization: 388
Douglas-Rachford algorithm: 298
drifting
dual
duality
Dudley entropy: 187
Dvoretzky-Kiefer-Wolfowitz inequality: 190
dyadic cube: 182
dyadic partition: 180
Dykstra algorithm: 259
dynamic
dynamic programming: 289
dynamic time warping: 289
dynamical plan: 234
E¶
earth mover’s distance: 57
edge flux: 310
edge imbalance: 33
Edgeworth expansion: 178
effective cost: 257
effective dimension: 195
effective rank: 195
electron density: 251
elliptical law: 39
elliptically contoured distribution: 39
embedding
alignment: 224
empirical
energy
energy-dissipation identity: 331
entropic
entropy: 380
$$-scaling: 92
equal marginal: 251
essential continuity: 379
Euclidean group: 224
Eulerian
explicit Euler scheme: 360
extreme minimizer: 49
F¶
factored coupling: 256
fairness: 237
farthest-point sampling: 363
fast diffusion equation: 350
feature cost: 283
feature sketch: 193
feature term: 284
Fenchel inequality: 135
Fenchel-Young loss: 263
FFT: 144
fiber: 234
fiberwise optimal plan: 234
fiberwise transport: 233
fill-in edge: 254
filtered back-projection: 248
finite metric length: 331
finite-sample bias: 374
finite-state transport: 309
first variation: 147, 261, 262, 287, 316, 317, 332, 345, 347, 359, 377, 378, 388, 392, 393, 395
fixed-point
Gaussian barycenter: 241
flattening map: 245
flow
Fokker–Planck
Fokker–Planck particle closure: 377
Fortet iteration: 158
Fourier features: 193
Fourier random features: 194
Fourier transform: 247
Fourier-slice theorem: 248
fractional diffusion: 354
fractional heat equation: 354
fractional PDE: 353
Frechet inception distance: 111
Frobenius inner product: 40
Frobenius norm: 230
frozen surrogate: 378
G¶
Gamma convergence: 128
GAN: 115
gauge direction: 161
gauge fixing: 173
gauge invariance: 162
Gaussian
blurring: 128
complex Sinkhorn: 153
convex order: 270
denoiser: 369
kernel: 194
kernel attention: 383
kernel rank: 195
mean: 390
mixture: 3, 93, 99, 121, 122, 125, 147, 200, 239, 258, 260, 277, 370, 372
one-dimensional: 394
optimal map: 31
preserving flow: 395
push-forward: 37
scale mixture: 198
Sinkhorn divergence: 171
Sinkhorn flow: 174
Sinkhorn problem: 174
sliced Wasserstein: 218
unbalanced: 207
Gaussian Hellinger-Kantorovich: 207
Gelbrich theorem: 397
generalized
generalized quantile: 34
generalized Wasserstein flow: 348
generator manifold: 374
geodesic: 307
Gibbs
Girsanov theorem: 136
global
global invariance: 224
gradient
graph
graphical model: 253
Grassmann manifold: 220
gravity interaction: 362
greedy sweep: 47
gross substitutes: 161
ground metric learning: 66
group action: 223
growth term: 311
H¶
Hadamard differentiability: 193
Hall theorem: 50
Hamilton-Jacobi-Bellman equation: 314
hard congestion: 314
Hausdorff measure: 21
heavy-tailed noise: 355
Hegselmann–Krause model: 384
Hellinger
Hellinger-Kantorovich: 312
Helmholtz decomposition: 390
hidden convexity: 380
Hilbert
Hilbert space: 330
Hilbertian
Hilbertian tangent norm: 300
Hilbertian Wasserstein space: 32
histogram: 4, 11, 32, 34, 63, 65, 118, 152, 165, 200, 237, 272, 278
Hoeffding inequality: 195
Holder curve: 331
Hölder inequality: 231
holomorphic continuation: 152
holomorphic implicit function theorem: 153
homogeneous action: 302
homogeneous Sobolev norm: 151
homogenization: 203
Horn matrix: 198
hyperbolic geodesic: 27
hyperbolic metric: 36
hypersurface: 21
I¶
ice-cream cone: 40
imitation learning: 116
in-context map: 383
incompressible Euler: 255
increasing rearrangement: 59
induced width: 254
inertia: 362
inertial particle method: 361
infinite-depth limit: 382
infinite-dimensional linear programming: 56
integer multiplicity: 52
interaction energy: 337
interaction kernel: 362
intermediate measure: 256
internal energy: 327
interpolant: 372
intrinsic length metric: 217
inverse isotonicity: 161
inverse OT gap loss: 263
inverse square root: 288
irreducibility: 308
isometric embedding: 284
isotonic regression: 59
iterative closest point: 224
J¶
Jacobi iteration: 161
Jensen inequality: 70, 113, 124, 177, 185, 188, 268, 296, 305
JKO scheme: 139, 316, 317, 319, 325, 348, 352, 355, 356, 360, 395, 397
joint
Jordan decomposition: 13
jump kernel: 306
junction tree: 254
K¶
Kantorovich
Kantorovich–Rubinstein
kernel
kernelized self-interaction: 321
kernelized Stein discrepancy: 375
kinetic action: 137
kinetic equation: 362
kinetic mean-field limit: 361
KL
KL barycenter: 241
Kolmogorov forward equation: 355
Kolmogorov-Smirnov distance: 32
Kullback-Leibler divergence: 36, 40, 113, 123, 124, 126, 139, 335
L¶
L2 attention: 383
L2 self-attention: 383
Lagrange multiplier: 324
Lagrangian
Lagrangian velocity: 361
Laguerre cell: 18, 84, 93–95, 100
discrete: 89
Langevin
Laplace method: 173
Laplace-Beltrami operator: 141
large-temperature collapse: 149
large-temperature limit: 151
latent
latent measure: 256
lattice distribution: 178
Lavenant criterion: 397
law over measures: 245
layer-cake formula: 32
layerwise matching: 359
least squares
Radon reconstruction: 247
least-square
length metric: 217
length space: 300
Lévy process: 354
linear
linear-time Sinkhorn: 193
linearized Wasserstein metric: 341
Liouville equation: 361
Lipschitz
Lloyd
local distance distribution: 278
local matching indicator: 2
local metric slope: 318
local profile: 278
local tangent action: 300
localization: 24
log-concave
log-concave generator: 382
log-concavity: 339
log-domain Sinkhorn: 118
log-scaling: 161
logarithmic error: 195
logarithmic kernel approximation: 195
look-ahead point: 360
Lorentz cone: 40
loss factor: 163
low-rank
lower semicontinuity: 57, 112, 113, 151, 201, 202, 245, 267, 272, 308, 312
lower semicontinuous relaxation: 302
lower-dimensional set: 21
Lyapunov equation: 389
M¶
M-matrix: 162
Maas distance: 309
many-token regime: 76
map estimation: 189
map extrapolation: 187
marginal
market clearing: 161
Markov
averaging: 384
masking: 382
mass
mass flux: 314
matching
matrix
matrix-vector iteration: 119
matrix-vector product: 194
max-flow min-cut theorem: 259
maximal correlation: 380
maximal density constraint: 323
maximum mean discrepancy: 108–111, 115, 116, 180, 184, 185, 228, 321, 326, 327, 342, 362, 374, 390, 392
McCann
mean
Wasserstein decomposition: 65
mean field game: 313
mean shift: 383
mean-field momentum: 361
mean-preserving splitting: 268
mean-shift PDE: 383
measure
measure convolution: 72
measure-function duality: 81
measure-preserving
measure-to-measure map: 74
measure-valued Radon transform: 213
mechanical energy: 361
mesh size: 142
message passing: 254
metric
minibatch: 373
minibatch bias: 373
minibatch noise: 355
minimal velocity: 318
minimax lower bound: 182
minimax optimality: 182
minimizer
sparse: 46
minimizing movement scheme: 316
Minkowski problem: 379
mixed Hessian: 28
MMD-GAN: 373
modulo isometry: 280
modulus of continuity: 85
moment
convergence: 73
moment condition: 176
moment energy: 336
momentum variable: 300
Monge
Monge gap: 58
Monge–Ampère
monotone
Moore-Penrose pseudo-inverse: 247
multi-marginal
multi-omics: 284
multi-species
multi-valued subdifferential: 21
multinomial distribution: 192
multiplicative scaling: 140
multiscale coupling: 180
Muon polar factor: 352
mutual information: 126
N¶
Nash equilibrium: 313
natural gradient: 374
natural language processing: 66
near-linear Sinkhorn: 195
nearest neighbor: 224
negative Sobolev norm: 342
Nesterov acceleration: 360
network simplex: 106
neural
Newton dynamics: 362
Newtonian dynamics: 362
Newton–Schulz iteration: 353
node feature: 284
noising path: 372
noising schedule: 370
noisy gradient descent: 347
non-convex domain: 25
non-convexity: 330
non-Gaussian limit: 193
nonconvex
optimization: 237
nonlinear flattening: 245
nonlocal
nonlocal Wasserstein
nonnegative factorization: 196
nonnegative rank: 256
nonnegative transport matrix: 51
norm
normal approximation: 177
normal cone: 323
normalized dynamics: 350
normalized SGD: 351
north-west corner
Nystrom approximation: 195
O¶
ODE limit: 360
one-dimensional
one-form: 163
operator bound: 231
operator norm: 13
operator-valued coupling: 288
opinion dynamics: 384
optimal coupling: 29–31, 46, 51, 52, 56, 57, 61, 62, 67, 68, 80, 81, 86, 97, 238, 243, 252, 261, 263, 266, 280, 281, 296, 325
optimal face: 192
optimal plan: 20, 28, 44, 46, 48, 62, 83, 86, 121, 145, 266, 280, 325
optimal transport map
optimality
orbit space: 224
order interval: 161
order-constrained coupling: 269
OT map estimation: 187
OT on trees: 33
OU bridge: 372
outside option: 163
overshooting bridge: 372
P¶
pair space: 307
pair-space action: 306
pairwise distance: 279
parabolic optimal transport: 172
parameter space: 373
partial matching: 8
particle method: 75
particle splitting: 306
particle Wasserstein descent: 360
path action: 300
path entropy: 291
path metric: 217
path-space problem: 136
Pearson divergence: 114
penalized minimization oracle: 349
perfect matching: 50
permutation matrix: 48
perspective recession: 272
perturbation response: 374
phase space: 361
phase-space density: 364
phase-space equation: 363
phase-space measure: 361
piecewise affine map: 190
PL convergence: 333
plan
PMO: 349
Poincaré disk: 27
Poisson equation: 142
polar factor: 353
polar formula: 231
Polyak momentum: 360
Polyak-Lojasiewicz inequality: 333
population law: 313
Portmanteau theorem: 57
positive
potential energy: 336
potential game: 313
power divergence: 328
power-law jump kernel: 354
Prékopa
primal-dual
method: 89
principal value: 354
probability integral transform: 29
probability measure: 11, 12, 14, 17, 20, 32, 35, 54, 57, 63, 71, 78, 82, 108–110, 114, 148, 213, 217, 226, 236, 249, 276, 279, 282, 295, 296, 321, 346
probability path: 366
Procrustes
product
product Wasserstein metric: 325
projected formulation: 230
projected measure: 213
projection direction: 213
projection-free attention: 383
projection-robust Wasserstein: 220
projective
propagation of chaos: 322
proximal
pullback: 15
Q¶
quadratic
quantile: 29
quantization
quantum
quotient
R¶
Rademacher bound: 187
radial coupling: 22
radial generator: 39
radial measure: 21
Radon barycenter: 247
Radon pseudo-inverse: 247
Radon sinogram: 247
Radon structure: 215
Radon Wasserstein barycenter: 247
Radon Wasserstein distance: 214
Radon-Nikodym derivative: 367
ramp filter: 248
random measure: 245
random walk: 310
rank: 46
rank constraint: 230
rank-deficient covariance: 40
rate
raw second moment: 38
reaction term: 356
reaction–transport action: 311
reaction-transport geometry: 311
rectifiable set: 21
reference
reflow: 366
regression loss: 365
regular conditional distribution: 33
regular vector field: 306
regularization
Monge gap: 59
reinforcement learning: 116
relative
rescaled convolution: 72
residual attention: 383
residual block: 358
residual network: 359
ResNet: 357
resolvent: 142
reverse formulation: 203
reverse KL divergence: 114
reversible jump kernel: 306
reversible Markov chain: 309
Riccati equation: 174
Ricci curvature: 337
Riemann sum: 71
Riemannian metric: 300
Riemannian tensor: 300
Riesz potential: 109
rigid registration: 224
rigid update: 226
robust
row constraint: 162
running-intersection property: 254
S¶
Sakoe–Chiba band: 289
sampling: 74
scalar bridge: 372
scaling
Schoenberg theorem: 198
Schrödinger
score-based generative modeling: 365
score-SDE: 365
second-moment quotient: 38
second-order limit: 193
sectional curvature: 29
self Sinkhorn: 151
self-concordance: 119
self-corrected field: 378
self-distance: 191
semi-coupling: 203
semi-discrete
discrete: 88
semi-relaxed problem: 85
semidefinite program: 285
semidefinite programming: 184
separable space: 12
separator: 255
Shampoo optimizer: 351
Shannon-Boltzmann entropy: 117
shape space: 223
signed
simplex: 310
simplex network: 53
Sinkhorn
closed-form coupling: 127
continuous dual iteration: 134
continuous iteration: 126
continuous limit: 173
coupling: 138
divergence: 117, 149–151, 171, 172, 180, 185, 193, 322, 374, 375, 393, 394
full-cycle map: 168
GAN: 373
kernel: 194
linear-time: 193
potential: 173
scaling: 117, 119, 121, 139, 145, 149, 156, 158, 161, 162, 169, 206, 288, 377
sketching: 193
statistical convergence: 176
update: 287
slack variable: 8
sliced barycenter: 247
sliced length metric: 217
sliced tangent norm: 215
sliced Wasserstein
slope
small-jump limit: 309
smooth density: 184
smooth OT estimation: 184
smoothed plug-in estimator: 183
smoothing estimator: 183
smoothness: 183
Sobolev norm: 151
soft
soft-DTW: 290
expected alignment: 292
softmax attention: 199
source term: 311
sparse linear program: 106
sparsity: 46
spectral
spectral Wasserstein
spherical average: 213
spherical cap: 190
spherical geodesic: 27
squared local action: 300
stable process: 354
stationary condition: 347
stationary density: 346
stationary potential: 173
statistical bias: 193
Stein force: 375
Stein geometry: 305
Stein method: 177
Stein operator: 375
step-size normalization: 353
Stiefel manifold: 219
stochastic
Strang splitting: 287
Strassen theorem: 269
strict
strictly correlated electrons: 251
strong law of large numbers: 176
strong MTW condition: 29
strong upper gradient: 318
structured data: 278
subgaussian measure: 185
subgradient: 262
subgradient inequality: 347
sublinear convergence: 335
subsolution: 161
subspace
subspace-sliced Wasserstein: 219
substitutability: 161
sum of squares: 184
sum-of-squares relaxation: 184
sum-product algorithm: 254
supergradient: 262
supersolution: 161
SVGD: 375
symmetric plan: 251
synthetic Ricci curvature: 337
T¶
tail index: 355
tangent coordinate: 246
tangent norm: 300
teacher kernel mean: 321
temporal noise: 365
tensor product coupling: 54
Thompson metric: 206
time change: 372
time reparametrization: 372
time series
alignment: 289
token limit: 382
Toland duality: 381
tomography: 248
topology
total unimodularity: 52
total variation: 11, 13, 63, 71, 72, 74, 107, 108, 113, 114, 200, 201, 310
totally positive kernel: 196
trace
trace-class operator: 288
transfer learning: 57
transport
transport-entropy inequality: 343
transportation polytope: 43, 44, 47, 52–54, 118, 119, 124, 156, 165, 259, 278
tree
tree-sliced Wasserstein: 33
tree-Wasserstein distance: 33
treewidth: 254
triangle inequality: 6, 19, 20, 24, 41, 64, 65, 67, 103, 166, 182, 185, 205, 214, 231, 279–281, 308
triangular map: 33
trivial coupling: 366
truncated marginal: 212
two-sample statistic: 192
two-sample testing: 111
U¶
V¶
Varadhan formula: 141
variable elimination: 254
variance: 190
variational
mean field game: 313
variational dual formula: 115
velocity covariance: 304
Vlasov equation: 362
von Neumann entropy: 286
W¶
Waddington landscape: 318
Waddington-OT: 66
warping path: 289
Wasserstein
barycenter: 139, 228, 236, 237, 239, 243, 245, 246, 252, 253
convergence: 176
coordinate: 227
distance: 32, 43, 63–65, 67, 68, 70, 72–74, 77, 114, 115, 177, 200, 205, 213, 214, 229, 230, 232, 276–279, 282, 295, 299, 303, 332, 380, 396
empirical rate: 180
flow: 332
flow of discrepancy: 374
formula: 31
GAN: 373
Gaussian geometry: 394
gradient: 75, 295, 316–318, 321, 323, 332, 345, 347, 360, 365, 375, 377, 378, 388–390, 395, 397
gradient flow: 23, 100, 147, 295, 316–318, 320, 321, 323, 332, 335, 347, 357, 359, 365, 374, 375, 388, 397
infinity distance: 78
infinity robustness: 78
Kurdyka–Łojasiewicz inequality: 335
Lipschitz functional: 75
minimax rate: 182
p-action: 303
p-distance: 26
space CLT: 246
squared distance: 330
training: 345
wavelet estimator: 184
weak
weight clipping: 116
weighted
wide stencil: 25
word mover’s distance: 66