REVIEW 3 major objections 4 minor 32 references
Solving Self-calibration of ALMA Data with an Optimization Method
T0 review · 3 major / 4 minor · reviewed 2026-08-11 · deepseek-v4-flash
Pith's one-line read This paper turns the gain-correction step of radio-interferometry self-calibration into a regularized optimization problem and, combined with its RML imaging method, reformulates the entire self-calibration loop as a single optimization…
desk verdict Well-built joint self-calibration/imaging framework, but the empirical case rests on hand-set gain variance targets and no baseline comparison; still deserves serious refereeing. read the letter →
The pith
A machine-rendered reading of the paper's core claim, the machinery that carries it, and where it could break.
The reading
What carries the argument
The load-bearing object is the combined regularized objective of equation (14). Its gain term $S_{\mu_1,\mu_2}(g)$ is what turns self-calibration into optimization: it pulls adjacent-in-time gains toward one another through weights $w_{\alpha l}=1/(t_l-t_{l-1})$, with $\mu_1$ penalizing changes in the complex gain and $\mu_2$ penalizing only amplitude changes, while the normalization constraint prevents all gains from collapsing to zero. Because the joint objective is non-convex, the algorithm alternates a convex image step (PRIISM with $\ell^1$ plus total squared variation, solved by MFISTA using a non-uniform FFT) and a gain step solved by the duplicate-variable quadratic surrogate in Appendix 1, whose closed-form update is refined by increasing the coupling parameter $\rho$. The four regularization weights are chosen by alternating Bayesian optimization: the image weights are selected so the reconstructed image's power lies inside the covering u-v ellipsoid and the visibility residuals are uniform over u-v distance, while the gain weights are selected to bring the phase and amplitude standard deviations $(\sigma_{\rm ph},\sigma_{\rm amp})$ to hand-set targets $(\sigma^*_{\rm ph},\sigma^*_{\rm amp})$.
What would settle it
Simulate an ALMA-like observation with a known sky image and known time-varying antenna gains, run the proposed method with the paper's hand-set variance targets, and compare the recovered gains and image with the truth: the central claim fails if the recovered image is no better than the one with gains fixed to unity, or if the ranking of the two hand-set target choices reverses the reported sharpening.
Extended reading notes
Core claim
The paper's central claim is that self-calibration can be reduced to the joint minimization of equation (14), $L_{\tilde v}(x,g) + R_{\lambda_1,\lambda_2}(x) + S_{\mu_1,\mu_2}(g)$, over an image $x\ge 0$ and complex antenna gains $g$ with $\sum_l |g_{\alpha l}|=N_\alpha$. Here $L_{\tilde v}$ is the weighted visibility likelihood, $R$ combines the $\ell^1$ norm and total squared variation to favor sparse, smooth images, and $S$ penalizes squared time differences of the complex gains and of their amplitudes. The joint problem is non-convex, so the paper solves it by alternating the PRIISM image update and a gain update obtained from a quadratic surrogate that has the closed-form solution $\hat g_{\alpha l}=r_{\alpha l} b_{\alpha l}/|b_{\alpha l}|$ from Appendix 1. The authors report that on HL Tau, SDP.81, and HD 142527 the reconstructed images appear sharper with higher peak intensities than without self-calibration, and the estimated gains change smoothly in time.
Load-bearing premise
The whole procedure leans on the hand-set target gain variances $(\sigma^*_{\rm ph},\sigma^*_{\rm amp})$ that pick the regularization strengths $\mu_1,\mu_2$; the paper sets these targets by hand and says an objective way to derive them from ALMA data is future work.
Editorial extensions
If this is right
- The traditional hand-tuned loop of CLEAN plus separate gain solutions is replaceable by alternating a convex image update and a closed-form gain update within one objective.
- The temporal-smoothness penalties on gains encode the physical expectation that atmospheric phase and amplitude errors evolve continuously, so estimated gains become smooth time series rather than piecewise constant solution intervals.
- On all three data sets the method produces sharper images and higher peak intensities than no self-calibration, with the degree of correction controlled by the target gain variances.
- Because the formulation separates the objective from the solver, other image regularizers or gain priors can be dropped into equation (14) without redesigning the self-calibration logic.
- The announced public release of the self-calibration module alongside PRIISM would let ALMA users apply RML self-calibration without building their own gain solvers.
Reading between the lines
- A natural next step is an objective, data-driven rule for the target gain variances; linking them to weather diagnostics such as phase-monitor RMS or precipitable-water-vapor measurements would test whether the hand-set values are replaceable.
- A direct head-to-head comparison with CLEAN-based self-calibration on identical data would separate how much of the sharpening comes from the gain correction and how much from the RML image prior.
- The same alternating scheme could extend to polarization, multiband, and spectral-line imaging, since those enter only through the data-fidelity term and the regularizers.
Signed reviews
Editorial analysis
A structured set of objections, weighed in public.
Referee Report
Summary. The paper reformulates radio-interferometric self-calibration as a joint regularized optimization over the image and per-antenna time-dependent complex gains (Eq. 14). The image subproblem is solved with the authors' PRIISM RML method (Eq. 10), and the gain subproblem is solved with a closed-form update derived through variable splitting (Appendix 1, Eqs. A1–A5). The four regularization parameters are selected by Bayesian optimization using two image-domain criteria C1 and C2 (Eq. 16) and gain variance targets (Eq. 17). The method is demonstrated on three ALMA data sets (HL Tau, SDP.81, HD 142527), each with 'small' and 'large' hand-set gain variance targets. The reported results are qualitative: reconstructed images are described as sharper with higher peak intensity than without self-calibration, the estimated gains vary smoothly in time, and achieved values of C1, C2, sigma_ph, and sigma_amp are tabulated.
Significance. If the gain estimates are correct, the formulation offers a principled, modular alternative to the iterative CLEAN/self-calibration loop, extending the RML framework to gain calibration with tunable temporal regularization. The closed-form gain update is a useful contribution, and testing on three real ALMA data sets of different morphologies is a strength. However, the current evaluation is self-referential: the same statistics used to select parameters are then reported as validation, and no comparison with standard CLEAN-based self-calibration or with injected gain errors is made. The significance therefore hinges on future validation; as it stands, the claim of 'promising results' is plausible but not yet established.
major comments (3)
- [§3.4, Eq. (16); Tables 2, 4, 6] The selection of lambda_1 and lambda_2 minimizes Eq. (16), which contains penalties hsq(C1* - C1) + hsq(C2* - C2) with C1* = 0.995 and C2* = 0.99, and the selection of mu_1 and mu_2 minimizes Eq. (17), which drives sigma_ph and sigma_amp toward hand-set targets. The paper then reports the achieved C1, C2, sigma_ph, and sigma_amp in Tables 2, 4, and 6 as evidence of image and gain quality. This is circular: these quantities are optimized to satisfy those conditions, so they cannot independently validate the reconstruction. The text should either present C1 and C2 only as enforced constraints, or provide independent validation metrics.
- [§4.1, Figs. 3–5; §5] The empirical evidence for improvement is limited to qualitative statements that the images 'give sharper impressions' and have higher peak intensities than the no-self-calibration case. Because larger allowed gain variance directly permits larger gain fluctuations, the fact that the 'Large variance' runs produce sharper, higher-peak images is a direct consequence of the hand-set targets rather than evidence that the estimated gains are correct. The absence of a comparison with standard CLEAN-based self-calibration or with simulated data with known injected gain errors—both deferred to future work in §5—means the central claim of 'promising results' is not yet supported by the experiments as presented.
- [§3.4, Eq. (17); §5] The target gain variances sigma*_ph and sigma*_amp are set by hand ('We set the target values by hand in this work'), and the paper states that an objective choice from ALMA observations is future work. Since the true gains are unknown, the output of the method depends on these free parameters. The paper needs either a principled, data-driven procedure for setting the targets, or a demonstration that the reconstructions are insensitive to reasonable target choices. The two hand-picked cases shown are not sufficient to establish robustness.
minor comments (4)
- [§4.1] The sentence 'The gains of figures 3b, have larger variances than those of figures 3e' contradicts the figure labels; the Small-variance gains in Fig. 3b should have smaller variances than the Large-variance gains in Fig. 3e.
- [Appendix 1, Eq. (A5)] The symbol y_k in the definition of b_{alpha l} is not defined in the manuscript; it should be introduced explicitly (presumably the model visibility F_k(x) or the calibrated visibility).
- [§3.4] The search over lambda_1, lambda_2, mu_1, and mu_2 is restricted to integer values of their logarithms and limited to 30 Bayesian optimization trials, but the search ranges for Lambda_1, Lambda_2, M_1, and M_2 are not stated; because some selected values lie at the edge of the reported tables (e.g., Lambda_2 = 13 in Table 5), it is unclear whether the optimum is inside the allowed grid.
- [Table 6] For HD 142527 with Small variance, C2 = 0.845, which is far below the C2 >= 0.99 target used in Eq. (16); the text acknowledges this, but it should be discussed explicitly as a failure of the parameter-selection loop to satisfy its own constraint, and the implications for the reconstructed image should be addressed.
Circularity Check
The reported validation is partly circular: λ selection enforces the C1/C2 thresholds that are then presented as quality metrics, and gain scatters in the tables are fitted to hand-set targets by Eq. (17).
-
fitted input called prediction
[Section 3.4, Eq. (16); Section 4, Tables 2/4/6 and §4.1 text]
"For {λ1, λ2}, the Bayesian optimization problem is defined as min_{λ1,λ2} [ L ˜v( ˆx, ˆg) + κ ( hsq(C ∗ 1 − C1) + hsq(C ∗ 2 − C2) ) ] ... and we set κ = 10^6 in § 4. ... The statistics related to the image and the gains are shown in table 2. Almost all the power of the image is concentrated in the covering ellipsoid (C1 > 0.99), and the power of the residual is distributed equally over u–v distance (C2 > 0.99)."
C1 and C2 are not independent diagnostics: they are the exact quantities appearing in the soft constraints of Eq. (16), with a huge penalty κ = 10^6 driving the selected λ1, λ2 toward C1 ≥ 0.995 and C2 ≥ 0.99. Reporting the resulting C1, C2 values in Tables 2, 4, 6 and reading the threshold satisfaction as evidence of a good image-model fit is therefore reporting the objective's own penalty terms. The 'almost all power...' statements in §4 restate the selection constraints rather than validate the reconstruction from the data.
-
fitted input called prediction
[Section 3.4, Eq. (17); Section 4, Tables 2/4/6; Conclusion]
"The parameters µ1 and µ2 are chosen to make (σph, σamp) close to a given target values (σ∗ ph, σ∗ amp). We set the target values by hand in this work. ... Since we do not know the true gains, how to define the variances of gains from ALMA observation information is an open problem."
Eq. (17) explicitly minimizes the relative distance of the estimated gain scatters (σph, σamp) to hand-set targets. Hence the σph and σamp values in Tables 2, 4, 6 (e.g. 4.68 deg vs 16.9 deg for HL Tau) are produced to match the 'Small variance' and 'Large variance' inputs, not inferred from data. The paper's observation that larger targets give sharper images with higher peak intensities is the expected consequence of allowing larger gain modulation, as the paper itself states in §3.4, and cannot validate the gain correction without ground-truth gains or a comparison to CLEAN-based self-calibration. The Conclusion concedes this selection principle is open and the CLEAN comparison is ongoing.
full rationale
The core reformulation in Eq. (14) is genuinely self-contained: it combines a visibility likelihood, image regularization, and gain smoothness regularization into one non-convex problem, and the alternating algorithm is a legitimate optimization proposal with no load-bearing self-citation or imported uniqueness theorem. The circularity is concentrated in the validation pipeline. Eq. (16) selects λ1, λ2 by penalizing C1 < 0.995 and C2 < 0.99, and §4 then presents those same C1, C2 values as evidence of good image-model fit; Eq. (17) forces the estimated gain scatter toward hand-set targets, and Tables 2, 4, 6 report that scatter as a result. The Conclusion's explicit admission that the gain-variance choice is open and that comparison with traditional CLEAN is ongoing reduces the overreach, but it does not remove the fact that the reported diagnostics are, by construction, fitted to the selection criteria. The central method retains independent mathematical content, so the score is 6 (partial circularity in the evidence) rather than 8 or 10.
Assumptions & free parameters
free parameters (9)
- λ1 =
10^Λ1 with Λ1 from -9 to 3 (e.g., -4 for HL Tau, -9 for SDP.81, 2-3 for HD 142527)
- λ2 =
10^Λ2 with Λ2 from 8 to 13 (e.g., 8 for HL Tau, 12 for SDP.81, 13 for HD 142527)
- μ1 =
10^M1 with M1 from 3 to 9 (e.g., 6 and 4 for HL Tau Small/Large, 9 and 6 for HD 142527)
- μ2 =
10^M2 with M2 from -4 to 4 (e.g., 4 for HL Tau Small, 1 for SDP.81 Large, -4 for HD 142527 Small)
- σ*_ph =
5 deg (Small variance) and 15 deg (Large variance)
- σ*_amp =
0.05 (Small variance) and 0.10 (Large variance)
- C*_1 =
0.995
- C*_2 =
0.99
- κ =
1e6
assumptions (4)
- domain assumption Each visibility satisfies ṽ_k g_αl g*_βl = F_k(x) + n_k (Eq. 4), with a single complex gain per antenna per integration.
- domain assumption Gains vary smoothly over time, encoded by the quadratic penalty S (Eq. 13).
- domain assumption Noise is independent Gaussian with known per-visibility variances σ_k^2 (Eq. 8).
- ad hoc to paper The image is non-negative and the covering u-v ellipsoid conditions C1 and C2 are appropriate quality criteria for parameter selection.
Cite this review
Pith. "Pith review of Solving Self-calibration of ALMA Data with an Optimization Method." pith.science (2026). https://pith.science/paper/QLNLQJEA
@misc{pith2026241203183,
author = {Pith},
title = {Pith review of: Solving Self-calibration of ALMA Data with an Optimization Method},
year = {2026},
howpublished = {\url{https://pith.science/paper/QLNLQJEA}},
note = {Machine review of arXiv:2412.03183}
}
read the original abstract
We reformulate the gain correction problem of the radio interferometry as an optimization problem with regularization, which is solved efficiently with an iterative algorithm. Combining this new method with our previously proposed imaging method, PRIISM, the whole process of the self-calibration of radio interferometry is redefined as a single optimization problem with regularization. As a result, the gains are corrected, and an image is estimated. We tested the new approach with ALMA observation data and found it provides promising results.
Figures
Figures from the paper (10 more)
Reference graph
Works this paper leans on
-
[1]
ALMA Partnership , Fomalont, E. B., Vlahakis, C., et al. 2015 a , ApJL, 808, L1
work page 2015
-
[2]
ALMA Partnership , Brogan, C. L., P\' e rez, L. M., et al. 2015 b , ApJL, 808, L3
work page 2015
- [3]
-
[4]
2009, SIAM Journal on Imaging Sciences, 2, 183
Beck, A., & Teboulle, M. 2009, SIAM Journal on Imaging Sciences, 2, 183
work page 2009
- [5]
-
[6]
2022, eht-imaging, doi:10.5281/zenodo.7226661
Chael, A. 2022, eht-imaging, doi:10.5281/zenodo.7226661
- [7]
-
[8]
Clark, B. G. 1980, A&A, 89, 377
1980
Show all 32 references
-
[9]
2019, ApJL, 875, L4
EHT Collaboration . 2019, ApJL, 875, L4
2019
-
[10]
2022, ApJL, 930, L14
---. 2022, ApJL, 930, L14
2022
-
[11]
, Slezak, E
Ellien, A. , Slezak, E. , Martinet, N. , et al. 2021, A&A, 649, A38
2021
-
[12]
2004, SIAM Review, 46, 443
Greengard, L., & Lee, J.-Y. 2004, SIAM Review, 46, 443
2004
-
[13]
2020, scikit-optimize, doi:10.5281/zenodo.4014775
Head, T., Kumar, M., Nahrstaedt, H., Louppe, G., & Shcherbatyi, I. 2020, scikit-optimize, doi:10.5281/zenodo.4014775
2020 doi
-
[14]
H\" o gbom, J. A. 1974, A&A Supplement, 15, 417
1974
-
[15]
2014, PASJ, 66, 95
Honma, M., Akiyama, K., Uemura, M., & Ikeda, S. 2014, PASJ, 66, 95
2014
-
[16]
2016, ApJL, 831, L12
Kataoka, A., Tsukagoshi, T., Momose, M., et al. 2016, ApJL, 831, L12
2016
-
[17]
2018, ApJ, 858, 56
Kuramochi, K., Akiyama, K., Ikeda, S., et al. 2018, ApJ, 858, 56
2018
-
[18]
1993, IEEE trans
Mallat, S., & Zhang, Z. 1993, IEEE trans. Signal Process., 41, 3397
1993
-
[19]
2022, SMILI v0.2.0, doi:10.5281/zenodo.6522933
Moriyama, K., Akiyama, K., Cho, I., et al. 2022, SMILI v0.2.0, doi:10.5281/zenodo.6522933
2022 doi
-
[20]
2020, PRIISM : Python module for Radio Interferometry Imaging with Sparse Modeling, Astrophysics Source Code Library, record ascl:2006.002
Nakazato, T., & Ikeda, S. 2020, PRIISM : Python module for Radio Interferometry Imaging with Sparse Modeling, Astrophysics Source Code Library, record ascl:2006.002
2020
-
[21]
2018, in Proc
Nakazato, T., Ikeda, S., Akiyama, K., et al. 2018, in Proc. of ADASS XXVIII, Maryland, U.S.A., O4.1
2018
-
[22]
2017, MNRAS, 470, 3981–4006
Repetti, A., Birdi, J., Dabbech, A., & Wiaux, Y. 2017, MNRAS, 470, 3981–4006
2017
-
[23]
Richards, A. M. S., Moravec, E., Etoka, S., et al. 2022, ALMA Memo, 620
2022
-
[24]
Schwab, F. R. 1980, in 1980 Intl Optical Computing Conf I, ed. W. T. Rhodes, Vol. 0231, International Society for Optics and Photonics (SPIE), 18 -- 25
1980
-
[25]
Schwab, F. R. 1984, AJ, 89, 1076
1984
-
[26]
2023, PhD thesis, Graduate Institute for Advanced Studies (SOKENDAI)
Takahashi, S. 2023, PhD thesis, Graduate Institute for Advanced Studies (SOKENDAI)
2023
-
[27]
R., James, M
Thompson, A. R., James, M. M., & Swenson Jr., G. W. 2017, Interferometry and Synthesis in Radio Astronomy (Springer International Publishing), doi:10.1007/978-3-319-44431-4
2017 doi
-
[28]
Wiaux, Y., Jacques, L., Puy, G., Scaife, A. M. M., & Vandergheynst, P. 2009, MNRAS, 395, 1733
2009
-
[29]
2021, ApJ, 923, 121
Yamaguchi, M., Tsukagoshi, T., Muto, T., et al. 2021, ApJ, 923, 121
2021
-
[30]
2020, ApJ, 895, 84
Yamaguchi, M., Akiyama, K., Tsukagoshi, T., et al. 2020, ApJ, 895, 84
2020
-
[31]
2024, PASJ, 76, 437
Yamaguchi, M., Muto, T., Tsukagoshi, T., et al. 2024, PASJ, 76, 437
2024
-
[32]
\@bibitem \@bib@author\@prev@author \@set@biblabel \@lbibitem[#1] \@bib@parse#1()\@nil \@set@biblabel \@bib@parse#1(#2)#3\@nil \@bib@author #1 @edef\@bib@year @space#2 \@empty \@set@biblabel#1 \@bib@author\@empty \@latex@warning Author name should be given for reference entry ...
Reviewed August 11, 2026 · model on record in the stance chip above.
Discussion (0). Continue with ORCID to comment.