Application of an unbalanced optimal transport distance and a mixed L1/Wasserstein distance to full waveform inversion
Bibliographic record
Abstract
Full waveform inversion (FWI) is an important and popular technique in subsurface earth property estimation. However, using the least-squares norm in the misfit function often leads to the local minimum solution of the optimization problem, and this phenomenon can be explained with the cycle-skipping artifact. Several methods that apply optimal transport distances to mitigate the cycle-skipping artifact have been proposed recently. The optimal transport distance is designed to compare two probability measures. To overcome the mass equality limit, we introduce an unbalanced optimal transport (UOT) distance with KullbackLeibler divergence to balance the mass difference. Also, a mixed L1/Wasserstein distance is constructed that can preserve the convex properties with respect to shift, dilation, and amplitude change operation. An entropy regularization approach and scaling algorithms are used to compute the distance and the gradient efficiently. Two strategies of normalization methods that transform the seismic signals into non-negative functions are provided. Numerical examples are provided to demonstrate the efficiency and effectiveness of the new method.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.004 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.001 | 0.001 |
| Scholarly communication | 0.001 | 0.002 |
| Open science | 0.001 | 0.002 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".