Searcharxiv⌕ Search

arXiv subjects

S. Gratton

Publications and source records attributed to S. Gratton.

At least 37 records · Page 2Linked to original sources

Complexity of a Class of First-Order Objective-Function-Free Optimization Algorithms

A parametric class of trust-region algorithms for unconstrained nonconvex optimization is considered where the value of the objective function is never computed. The class contains a deterministic version of the first-order Adagrad method typically used for minimization of noisy function, but also allows the use of (possibly approximate) second-order information when available. The rate of convergence of methods in the class is analyzed and is shown to be identical to that known for first-order optimization methods using both function and gradients values, recovering existing results for purely-first order variants and improving the explicit dependence on problem dimension. This rate is shown to be essentially sharp. A new class of methods is also presented, for which a slightly worse and essentially sharp complexity result holds. Limited numerical experiments show that the new methods' performance may be comparable to that of standard steepest descent, despite using significantly less information, and that this performance is relatively insensitive to noise.

math.OC↗

Multilevel Objective-Function-Free Optimization with an Application to Neural Networks Training

A class of multi-level algorithms for unconstrained nonlinear optimization is presented which does not require the evaluation of the objective function. The class contains the momentum-less AdaGrad method as a particular (single-level) instance. The choice of avoiding the evaluation of the objective function is intended to make the algorithms of the class less sensitive to noise, while the multi-level feature aims at reducing their computational cost. The evaluation complexity of these algorithms is analyzed and their behaviour in the presence of noise is then illustrated in the context of training deep neural networks for supervised learning applications.

math.OC↗

Convergence properties of an Objective-Function-Free Optimization regularization algorithm, including an $\mathcal{O}(ε^{-3/2})$ complexity bound

An adaptive regularization algorithm for unconstrained nonconvex optimization is presented in which the objective function is never evaluated, but only derivatives are used. This algorithm belongs to the class of adaptive regularization methods, for which optimal worst-case complexity results are known for the standard framework where the objective function is evaluated. It is shown in this paper that these excellent complexity bounds are also valid for the new algorithm, despite the fact that significantly less information is used. In particular, it is shown that, if derivatives of degree one to $p$ are used, the algorithm will find a $ε_1$-approximate first-order minimizer in at most $O(ε_1^{-(p+1)/p})$ iterations, and an $(ε_1,ε_2)$-approximate second-order minimizer in at most $O(\max[ε^{-(p+1)/p},ε_2^{-(p+1)/(p-1)}])$ iterations. As a special case, the new algorithm using first and second derivatives, when applied to functions with Lipschitz continuous Hessian, will find an iterate $x_k$ at which the gradient's norm is less than $ε_1$ in at most $O(ε_1^{-3/2})$ iterations.

math.OC↗

OFFO minimization algorithms for second-order optimality and their complexity

An Adagrad-inspired class of algorithms for smooth unconstrained optimization is presented in which the objective function is never evaluated and yet the gradient norms decrease at least as fast as $\calO(1/\sqrt{k+1})$ while second-order optimality measures converge to zero at least as fast as $\calO(1/(k+1)^{1/3})$. This latter rate of convergence is shown to be essentially sharp and is identical to that known for more standard algorithms (like trust-region or adaptive-regularization methods) using both function and derivatives' evaluations. A related "divergent stepsize" method is also described, whose essentially sharp rate of convergence is slighly inferior. It is finally discussed how to obtain weaker second-order optimality guarantees at a (much) reduced computional cost.

math.OC↗

Planck 2018 results. VI. Cosmological parameters

We present cosmological parameter results from the final full-mission Planck measurements of the CMB anisotropies. We find good consistency with the standard spatially-flat 6-parameter $Λ$CDM cosmology having a power-law spectrum of adiabatic scalar perturbations (denoted "base $Λ$CDM" in this paper), from polarization, temperature, and lensing, separately and in combination. A combined analysis gives dark matter density $Ω_c h^2 = 0.120\pm 0.001$, baryon density $Ω_b h^2 = 0.0224\pm 0.0001$, scalar spectral index $n_s = 0.965\pm 0.004$, and optical depth $τ= 0.054\pm 0.007$ (in this abstract we quote $68\,\%$ confidence regions on measured parameters and $95\,\%$ on upper limits). The angular acoustic scale is measured to $0.03\,\%$ precision, with $100θ_*=1.0411\pm 0.0003$. These results are only weakly dependent on the cosmological model and remain stable, with somewhat increased errors, in many commonly considered extensions. Assuming the base-$Λ$CDM cosmology, the inferred late-Universe parameters are: Hubble constant $H_0 = (67.4\pm 0.5)$km/s/Mpc; matter density parameter $Ω_m = 0.315\pm 0.007$; and matter fluctuation amplitude $σ_8 = 0.811\pm 0.006$. We find no compelling evidence for extensions to the base-$Λ$CDM model. Combining with BAO we constrain the effective extra relativistic degrees of freedom to be $N_{\rm eff} = 2.99\pm 0.17$, and the neutrino mass is tightly constrained to $\sum m_ν< 0.12$eV. The CMB spectra continue to prefer higher lensing amplitudes than predicted in base -$Λ$CDM at over $2\,σ$, which pulls some parameters that affect the lensing amplitude away from the base-$Λ$CDM model; however, this is not supported by the lensing reconstruction or (in models that also change the background geometry) BAO data. (Abridged)

astro-ph.CO↗

Planck 2018 results. IV. Diffuse component separation

We present full-sky maps of the cosmic microwave background (CMB) and polarized synchrotron and thermal dust emission, derived from the third set of Planck frequency maps. These products have significantly lower contamination from instrumental systematic effects than previous versions. The methodologies used to derive these maps follow those described in earlier papers, adopting four methods (Commander, NILC, SEVEM, and SMICA) to extract the CMB component, as well as three methods (Commander, GNILC, and SMICA) to extract astrophysical components. Our revised CMB temperature maps agree with corresponding products in the Planck 2015 delivery, whereas the polarization maps exhibit significantly lower large-scale power, reflecting the improved data processing described in companion papers; however, the noise properties of the resulting data products are complicated, and the best available end-to-end simulations exhibit relative biases with respect to the data at the few percent level. Using these maps, we are for the first time able to fit the spectral index of thermal dust independently over 3 degree regions. We derive a conservative estimate of the mean spectral index of polarized thermal dust emission of beta_d = 1.55 +/- 0.05, where the uncertainty marginalizes both over all known systematic uncertainties and different estimation techniques. For polarized synchrotron emission, we find a mean spectral index of beta_s = -3.1 +/- 0.1, consistent with previously reported measurements. We note that the current data processing does not allow for construction of unbiased single-bolometer maps, and this limits our ability to extract CO emission and correlated components. The foreground results for intensity derived in this paper therefore do not supersede corresponding Planck 2015 products. For polarization the new results supersede the corresponding 2015 products in all respects.

astro-ph.CO↗

Minimizing convex quadratic with variable precision conjugate gradients

We investigate the method of conjugate gradients, exploiting inaccurate matrix-vector products, for the solution of convex quadratic optimization problems. Theoretical performance bounds are derived, and the necessary quantities occurring in the theoretical bounds estimated, leading to a practical algorithm. Numerical experiments suggest that this approach has significant potential, including in the steadily more important context of multi-precision computations

math.NA↗

Planck 2018 results. V. CMB power spectra and likelihoods

This paper describes the 2018 Planck CMB likelihoods, following a hybrid approach similar to the 2015 one, with different approximations at low and high multipoles, and implementing several methodological and analysis refinements. With more realistic simulations, and better correction and modelling of systematics, we can now make full use of the High Frequency Instrument polarization data. The low-multipole 100x143 GHz EE cross-spectrum constrains the reionization optical-depth parameter $τ$ to better than 15% (in combination with with the other low- and high-$\ell$ likelihoods). We also update the 2015 baseline low-$\ell$ joint TEB likelihood based on the Low Frequency Instrument data, which provides a weaker $τ$ constraint. At high multipoles, a better model of the temperature-to-polarization leakage and corrections for the effective calibrations of the polarization channels (polarization efficiency or PE) allow us to fully use the polarization spectra, improving the constraints on the $Λ$CDM parameters by 20 to 30% compared to TT-only constraints. Tests on the modelling of the polarization demonstrate good consistency, with some residual modelling uncertainties, the accuracy of the PE modelling being the main limitation. Using our various tests, simulations, and comparison between different high-$\ell$ implementations, we estimate the consistency of the results to be better than the 0.5$σ$ level. Minor curiosities already present before (differences between $\ell$<800 and $\ell$>800 parameters or the preference for more smoothing of the $C_\ell$ peaks) are shown to be driven by the TT power spectrum and are not significantly modified by the inclusion of polarization. Overall, the legacy Planck CMB likelihoods provide a robust tool for constraining the cosmological model and represent a reference for future CMB observations. (Abridged)

astro-ph.CO↗

Planck intermediate results. LV. Reliability and thermal properties of high-frequency sources in the Second Planck Catalogue of Compact Sources

We describe an extension of the most recent version of the Planck Catalogue of Compact Sources (PCCS2), produced using a new multi-band Bayesian Extraction and Estimation Package (BeeP). BeeP assumes that the compact sources present in PCCS2 at 857 GHz have a dust-like spectral energy distribution, which leads to emission at both lower and higher frequencies, and adjusts the parameters of the source and its SED to fit the emission observed in Planck's three highest frequency channels at 353, 545, and 857 GHz, as well as the IRIS map at 3000 GHz. In order to reduce confusion regarding diffuse cirrus emission, BeeP's data model includes a description of the background emission surrounding each source, and it adjusts the confidence in the source parameter extraction based on the statistical properties of the spatial distribution of the background emission. BeeP produces the following three new sets of parameters for each source: (a) fits to a modified blackbody (MBB) thermal emission model of the source; (b) SED-independent source flux densities at each frequency considered; and (c) fits to an MBB model of the background in which the source is embedded. BeeP also calculates, for each source, a reliability parameter, which takes into account confusion due to the surrounding cirrus. We define a high-reliability subset (BeeP/base), containing 26 083 sources (54.1 per cent of the total PCCS2 catalogue), the majority of which have no information on reliability in the PCCS2. The results of the BeeP extension of PCCS2, which are made publicly available via the PLA, will enable the study of the thermal properties of well-defined samples of compact Galactic and extra-galactic dusty sources.

astro-ph.GA↗

Planck 2018 results. I. Overview and the cosmological legacy of Planck

The European Space Agency's Planck satellite, which was dedicated to studying the early Universe and its subsequent evolution, was launched on 14 May 2009. It scanned the microwave and submillimetre sky continuously between 12 August 2009 and 23 October 2013, producing deep, high-resolution, all-sky maps in nine frequency bands from 30 to 857GHz. This paper presents the cosmological legacy of Planck, which currently provides our strongest constraints on the parameters of the standard cosmological model and some of the tightest limits available on deviations from that model. The 6-parameter LCDM model continues to provide an excellent fit to the cosmic microwave background data at high and low redshift, describing the cosmological information in over a billion map pixels with just six parameters. With 18 peaks in the temperature and polarization angular power spectra constrained well, Planck measures five of the six parameters to better than 1% (simultaneously), with the best-determined parameter (theta_*) now known to 0.03%. We describe the multi-component sky as seen by Planck, the success of the LCDM model, and the connection to lower-redshift probes of structure formation. We also give a comprehensive summary of the major changes introduced in this 2018 release. The Planck data, alone and in combination with other probes, provide stringent constraints on our models of the early Universe and the large-scale structure within which all astrophysical objects form and evolve. We discuss some lessons learned from the Planck mission, and highlight areas ripe for further experimental advances.

astro-ph.CO↗

Planck 2018 results. X. Constraints on inflation

We report on the implications for cosmic inflation of the 2018 Release of the Planck CMB anisotropy measurements. The results are fully consistent with the two previous Planck cosmological releases, but have smaller uncertainties thanks to improvements in the characterization of polarization at low and high multipoles. Planck temperature, polarization, and lensing data determine the spectral index of scalar perturbations to be $n_\mathrm{s}=0.9649\pm 0.0042$ at 68% CL and show no evidence for a scale dependence of $n_\mathrm{s}.$ Spatial flatness is confirmed at a precision of 0.4% at 95% CL with the combination with BAO data. The Planck 95% CL upper limit on the tensor-to-scalar ratio, $r_{0.002}<0.10$, is further tightened by combining with the BICEP2/Keck Array BK15 data to obtain $r_{0.002}<0.056$. In the framework of single-field inflationary models with Einstein gravity, these results imply that: (a) slow-roll models with a concave potential, $V" (ϕ) < 0,$ are increasingly favoured by the data; and (b) two different methods for reconstructing the inflaton potential find no evidence for dynamics beyond slow roll. Non-parametric reconstructions of the primordial power spectrum consistently confirm a pure power law. A complementary analysis also finds no evidence for theoretically motivated parameterized features in the Planck power spectrum, a result further strengthened for certain oscillatory models by a new combined analysis that includes Planck bispectrum data. The new Planck polarization data provide a stringent test of the adiabaticity of the initial conditions. The polarization data also provide improved constraints on inflationary models that predict a small statistically anisotropic quadrupolar modulation of the primordial fluctuations. However, the polarization data do not confirm physical models for a scale-dependent dipolar modulation.

astro-ph.CO↗

Planck 2018 results. VIII. Gravitational lensing

We present measurements of the cosmic microwave background (CMB) lensing potential using the final $\textit{Planck}$ 2018 temperature and polarization data. We increase the significance of the detection of lensing in the polarization maps from $5\,σ$ to $9\,σ$. Combined with temperature, lensing is detected at $40\,σ$. We present an extensive set of tests of the robustness of the lensing-potential power spectrum, and construct a minimum-variance estimator likelihood over lensing multipoles $8 \le L \le 400$. We find good consistency between lensing constraints and the results from the $\textit{Planck}$ CMB power spectra within the $\rm{ΛCDM}$ model. Combined with baryon density and other weak priors, the lensing analysis alone constrains $σ_8 Ω_{\rm m}^{0.25}=0.589\pm 0.020$ ($1\,σ$ errors). Also combining with baryon acoustic oscillation (BAO) data, we find tight individual parameter constraints, $σ_8=0.811\pm0.019$, $H_0=67.9_{-1.3}^{+1.2}\,\text{km}\,\text{s}^{-1}\,\rm{Mpc}^{-1}$, and $Ω_{\rm m}=0.303^{+0.016}_{-0.018}$. Combining with $\textit{Planck}$ CMB power spectrum data, we measure $σ_8$ to better than $1\,\%$ precision, finding $σ_8=0.811\pm 0.006$. We find consistency with the lensing results from the Dark Energy Survey, and give combined lensing-only parameter constraints that are tighter than joint results using galaxy clustering. Using $\textit{Planck}$ cosmic infrared background (CIB) maps we make a combined estimate of the lensing potential over $60\,\%$ of the sky with considerably more small-scale signal. We demonstrate delensing of the $\textit{Planck}$ power spectra, detecting a maximum removal of $40\,\%$ of the lensing-induced power in all spectra. The improvement in the sharpening of the acoustic peaks by including both CIB and the quadratic lensing reconstruction is detected at high significance (abridged).

astro-ph.CO↗

Planck 2018 results. IX. Constraints on primordial non-Gaussianity

We analyse the Planck full-mission cosmic microwave background (CMB) temperature and E-mode polarization maps to obtain constraints on primordial non-Gaussianity (NG). We compare estimates obtained from separable template-fitting, binned, and modal bispectrum estimators, finding consistent values for the local, equilateral, and orthogonal bispectrum amplitudes. Our combined temperature and polarization analysis produces the following results: f_NL^local = -0.9 +\- 5.1; f_NL^equil = -26 +\- 47; and f_NL^ortho = - 38 +\- 24 (68%CL, statistical). These results include the low-multipole (4 <= l < 40) polarization data, not included in our previous analysis, pass an extensive battery of tests, and are stable with respect to our 2015 measurements. Polarization bispectra display a significant improvement in robustness; they can now be used independently to set NG constraints. We consider a large number of additional cases, e.g. scale-dependent feature and resonance bispectra, isocurvature primordial NG, and parity-breaking models, where we also place tight constraints but do not detect any signal. The non-primordial lensing bispectrum is detected with an improved significance compared to 2015, excluding the null hypothesis at 3.5 sigma. We present model-independent reconstructions and analyses of the CMB bispectrum. Our final constraint on the local trispectrum shape is g_NLl^local = (-5.8 +\-6.5) x 10^4 (68%CL, statistical), while constraints for other trispectra are also determined. We constrain the parameter space of different early-Universe scenarios, including general single-field models of inflation, multi-field and axion field parity-breaking models. Our results provide a high-precision test for structure-formation scenarios, in complete agreement with the basic picture of the LambdaCDM cosmology regarding the statistics of the initial conditions (abridged).

astro-ph.CO↗

A note on solving nonlinear optimization problems in variable precision

This short note considers an efficient variant of the trust-region algorithm with dynamic accuracy proposed Carter (1993) and Conn, Gould and Toint (2000) as a tool for very high-performance computing, an area where it is critical to allow multi-precision computations for keeping the energy dissipation under control. Numerical experiments are presented indicating that the use of the considered method can bring substantial savings in objective function's and gradient's evaluation "energy costs" by efficiently exploiting multi-precision computations.

math.NA↗

Planck 2018 results. XII. Galactic astrophysics using polarized dust emission

We present 353 GHz full-sky maps of the polarization fraction $p$, angle $ψ$, and dispersion of angles $S$ of Galactic dust thermal emission produced from the 2018 release of Planck data. We confirm that the mean and maximum of $p$ decrease with increasing $N_H$. The uncertainty on the maximum polarization fraction, $p_\mathrm{max}=22.0$% at 80 arcmin resolution, is dominated by the uncertainty on the zero level in total intensity. The observed inverse behaviour between $p$ and $S$ is interpreted with models of the polarized sky that include effects from only the topology of the turbulent Galactic magnetic field. Thus, the statistical properties of $p$, $ψ$, and $S$ mostly reflect the structure of the magnetic field. Nevertheless, we search for potential signatures of varying grain alignment and dust properties. First, we analyse the product map $S \times p$, looking for residual trends. While $p$ decreases by a factor of 3--4 between $N_H=10^{20}$ cm$^{-2}$ and $N_H=2\times 10^{22}$ cm$^{-2}$, $S \times p$ decreases by only about 25%, a systematic trend observed in both the diffuse ISM and molecular clouds. Second, we find no systematic trend of $S \times p$ with the dust temperature, even though in the diffuse ISM lines of sight with high $p$ and low $S$ tend to have colder dust. We also compare Planck data with starlight polarization in the visible at high latitudes. The agreement in polarization angles is remarkable. Two polarization emission-to-extinction ratios that characterize dust optical properties depend only weakly on $N_H$ and converge towards the values previously determined for translucent lines of sight. We determine an upper limit for the polarization fraction in extinction of 13%, compatible with the $p_\mathrm{max}$ observed in emission. These results provide strong constraints for models of Galactic dust in diffuse gas.

astro-ph.GA↗

Minimization of nonsmooth nonconvex functions using inexact evaluations and its worst-case complexity

An adaptive regularization algorithm using inexact function and derivatives evaluations is proposed for the solution of composite nonsmooth nonconvex optimization. It is shown that this algorithm needs at most $O(|\log(ε)|\,ε^{-2})$ evaluations of the problem's functions and their derivatives for finding an $ε$-approximate first-order stationary point. This complexity bound therefore generalizes that provided by [Bellavia, Gurioli, Morini and Toint, 2018] for inexact methods for smooth nonconvex problems, and is within a factor $|\log(ε)|$ of the optimal bound known for smooth and nonsmooth nonconvex minimization with exact evaluations. A practically more restrictive variant of the algorithm with worst-case complexity $O(|\log(ε)|+ε^{-2})$ is also presented.

math.OC↗

Planck intermediate results. LIV. The Planck Multi-frequency Catalogue of Non-thermal Sources

This paper presents the Planck Multi-frequency Catalogue of Non-thermal (i.e. synchrotron-dominated) Sources (PCNT) observed between 30 and 857 GHz by the ESA Planck mission. This catalogue was constructed by selecting objects detected in the full mission all-sky temperature maps at 30 and 143 GHz, with a signal-to-noise ratio (S/N)>3 in at least one of the two channels after filtering with a particular Mexican hat wavelet. As a result, 29400 source candidates were selected. Then, a multi-frequency analysis was performed using the Matrix Filters methodology at the position of these objects, and flux densities and errors were calculated for all of them in the nine Planck channels. The present catalogue is the first unbiased, full-sky catalogue of synchrotron-dominated sources published at millimetre and submillimetre wavelengths and constitutes a powerful database for statistical studies of non-thermal extragalactic sources, whose emission is dominated by the central active galactic nucleus. Together with the full multi-frequency catalogue, we also define the Bright Planck Multi-frequency Catalogue of Non-thermal Sources PCNTb, where only those objects with a S/N>4 at both 30 and 143 GHz were selected. In this catalogue 1146 compact sources are detected outside the adopted Planck GAL070 mask; thus, these sources constitute a highly reliable sample of extragalactic radio sources. We also flag the high-significance subsample PCNThs, a subset of 151 sources that are detected with S/N>4 in all nine Planck channels, 75 of which are found outside the Planck mask adopted here. The remaining 76 sources inside the Galactic mask are very likely Galactic objects.

astro-ph.CO↗

Planck 2018 results. II. Low Frequency Instrument data processing

We present a final description of the data-processing pipeline for the Planck, Low Frequency Instrument (LFI), implemented for the 2018 data release. Several improvements have been made with respect to the previous release, especially in the calibration process and in the correction of instrumental features such as the effects of nonlinearity in the response of the analogue-to-digital converters. We provide a brief pedagogical introduction to the complete pipeline, as well as a detailed description of the important changes implemented. Self-consistency of the pipeline is demonstrated using dedicated simulations and null tests. We present the final version of the LFI full sky maps at 30, 44, and 70 GHz, both in temperature and polarization, together with a refined estimate of the Solar dipole and a final assessment of the main LFI instrumental parameters.

astro-ph.CO↗