SearcharxivSearch

arXiv subjects

Gerard Lemson

Publications and source records attributed to Gerard Lemson.

At least 19 recordsLinked to original sources

TornadoNet: Real-Time Building Damage Detection with Ordinal Supervision

We present TornadoNet, a comprehensive benchmark for automated street-level building damage assessment evaluating how modern real-time object detection architectures and ordinal-aware supervision strategies perform under realistic post-disaster conditions. TornadoNet provides the first controlled benchmark demonstrating how architectural design and loss formulation jointly influence multi-level damage detection from street-view imagery, delivering methodological insights and deployable tools for disaster response. Using 3,333 high-resolution geotagged images and 8,890 annotated building instances from the 2021 Midwest tornado outbreak, we systematically compare CNN-based detectors from the YOLO family against transformer-based models (RT-DETR) for multi-level damage detection. Models are trained under standardized protocols using a five-level damage classification framework based on IN-CORE damage states, validated through expert cross-annotation. Baseline experiments reveal complementary architectural strengths. CNN-based YOLO models achieve highest detection accuracy and throughput, with larger variants reaching 46.05% mAP@0.5 at 66-276 FPS on A100 GPUs. Transformer-based RT-DETR models exhibit stronger ordinal consistency, achieving 88.13% Ordinal Top-1 Accuracy and MAOE of 0.65, indicating more reliable severity grading despite lower baseline mAP. To align supervision with the ordered nature of damage severity, we introduce soft ordinal classification targets and evaluate explicit ordinal-distance penalties. RT-DETR trained with calibrated ordinal supervision achieves 44.70% mAP@0.5, a 4.8 percentage-point improvement, with gains in ordinal metrics (91.15% Ordinal Top-1 Accuracy, MAOE = 0.56). These findings establish that ordinal-aware supervision improves damage severity estimation when aligned with detector architecture. Model & Data: https://github.com/crumeike/TornadoNet

cs.CV

Discovery of astrometric accelerations by dark companions in the globular cluster $ω$ Centauri

We present results from the search for astrometric accelerations of stars in $ω$ Centauri using 13 years of regularly-scheduled {\it Hubble Space Telescope} WFC3/UVIS calibration observations in the cluster core. The high-precision astrometry of $\sim$160\,000 sources was searched for significant deviations from linear proper motion. This led to the discovery of four cluster members and one foreground field star with compelling acceleration patterns. We interpret them as the result of the gravitational pull by an invisible companion and determined preliminary Keplerian orbit parameters, including the companion's mass. {For the cluster members} our analysis suggests periods ranging from 8.8 to 19+ years and dark companions in the mass range of $\sim$0.7 to $\sim$1.4$M_\mathrm{sun}$. At least one companion could exceed the upper mass-boundary of white dwarfs and can be classified as a neutron-star candidate.

astro-ph.SR

Indra: a Public Computationally Accessible Suite of Cosmological $N$-body Simulations

Indra is a suite of large-volume cosmological $N$-body simulations with the goal of providing excellent statistics of the large-scale features of the distribution of dark matter. Each of the 384 simulations is computed with the same cosmological parameters and different initial phases, with 1024$^3$ dark matter particles in a box of length 1 Gpc/h, 64 snapshots of particle data and halo catalogs, and 505 time steps of the Fourier modes of the density field, amounting to almost a petabyte of data. All of the Indra data are immediately available for analysis via the SciServer science platform, which provides interactive and batch computing modes, personal data storage, and other hosted data sets such as the Millennium simulations and many astronomical surveys. We present the Indra simulations, describe the data products and how to access them, and measure ensemble averages, variances, and covariances of the matter power spectrum, the matter correlation function, and the halo mass function to demonstrate the types of computations that Indra enables. We hope that Indra will be both a resource for large-scale structure research and a demonstration of how to make very large datasets public and computationally accessible.

astro-ph.CO

SciServer: a Science Platform for Astronomy and Beyond

We present SciServer, a science platform built and supported by the Institute for Data Intensive Engineering and Science at the Johns Hopkins University. SciServer builds upon and extends the SkyServer system of server-side tools that introduced the astronomical community to SQL (Structured Query Language) and has been serving the Sloan Digital Sky Survey catalog data to the public. SciServer uses a Docker/VM based architecture to provide interactive and batch mode server-side analysis with scripting languages like Python and R in various environments including Jupyter (notebooks), RStudio and command-line in addition to traditional SQL-based data analysis. Users have access to private file storage as well as personal SQL database space. A flexible resource access control system allows users to share their resources with collaborators, a feature that has also been very useful in classroom environments. All these services, wrapped in a layer of REST APIs, constitute a scalable collaborative data-driven science platform that is attractive to science disciplines beyond astronomy.

astro-ph.IM

Six Dimensional Streaming Algorithm for Cluster Finding in N-Body Simulations

Cosmological N-body simulations are crucial for understanding how the Universe evolves. Studying large-scale distributions of matter in these simulations and comparing them to observations usually involves detecting dense clusters of particles called "halos,'' which are gravitationally bound and expected to form galaxies. However, traditional cluster finders are computationally expensive and use massive amounts of memory. Recent work by Liu et al (Liu et al. (2015)) showed the connection between cluster detection and memory-efficient streaming algorithms and presented a halo finder based on heavy hitter algorithm. Later, Ivkin et al. (Ivkin et al. (2018)) improved the scalability of suggested streaming halo finder with efficient GPU implementation. Both works map particles' positions onto a discrete grid, and therefore lose the rest of the information, such as their velocities. Therefore, two halos travelling through each other are indistinguishable in positional space, while the velocity distribution of those halos can help to identify this process which is worth further studying. In this project we analyze data from the Millennium Simulation Project (Springel et al. (2005)) to motivate the inclusion of the velocity into streaming method we introduce. We then demonstrate a use of suggested method, which allows one to find the same halos as before, while also detecting those which were indistinguishable in prior methods.

astro-ph.GA

Scalable Streaming Tools for Analyzing $N$-body Simulations: Finding Halos and Investigating Excursion Sets in One Pass

Cosmological $N$-body simulations play a vital role in studying models for the evolution of the Universe. To compare to observations and make a scientific inference, statistic analysis on large simulation datasets, e.g., finding halos, obtaining multi-point correlation functions, is crucial. However, traditional in-memory methods for these tasks do not scale to the datasets that are forbiddingly large in modern simulations. Our prior paper proposes memory-efficient streaming algorithms that can find the largest halos in a simulation with up to $10^9$ particles on a small server or desktop. However, this approach fails when directly scaling to larger datasets. This paper presents a robust streaming tool that leverages state-of-the-art techniques on GPU boosting, sampling, and parallel I/O, to significantly improve performance and scalability. Our rigorous analysis of the sketch parameters improves the previous results from finding the centers of the $10^3$ largest halos to $\sim 10^4-10^5$, and reveals the trade-offs between memory, running time and number of halos. Our experiments show that our tool can scale to datasets with up to $\sim 10^{12}$ particles while using less than an hour of running time on a single GPU Nvidia GTX 1080.

astro-ph.IM

Probabilistic Cross-Identification of Galaxies with Realistic Clustering

Probabilistic cross-identification has been successfully applied to a number of problems in astronomy from matching simple point sources to associating stars with unknown proper motions and even radio observations with realistic morphology. Here we study the Bayes factor for clustered objects and focus in particular on galaxies to assess the effect of typical angular correlations. Numerical calculations provide the modified relationship, which (as expected) suppresses the evidence for the associations at the shortest separations where the 2-point auto-correlation function is large. Ultimately this means that the matching probability drops at somewhat shorter scales than in previous models.

astro-ph.GA

Galaxy formation in the Planck cosmology - IV. Mass and environmental quenching, conformity and clustering

We study the quenching of star formation as a function of redshift, environment and stellar mass in the galaxy formation simulations of Henriques et al. (2015), which implement an updated version of the Munich semi-analytic model (L-GALAXIES) on the two Millennium Simulations after scaling to a Planck cosmology. In this model massive galaxies are quenched by AGN feedback depending on both black hole and hot gas mass, and hence indirectly on stellar mass. In addition, satellite galaxies of any mass can be quenched by ram-pressure or tidal stripping of gas and through the suppression of gaseous infall. This combination of processes produces quenching efficiencies which depend on stellar mass, host halo mass, environment density, distance to group centre and group central galaxy properties in ways which agree qualitatively with observation. Some discrepancies remain in dense regions and close to group centres, where quenching still seems too efficient. In addition, although the mean stellar age of massive galaxies agrees with observation, the assumed AGN feedback model allows too much ongoing star formation at late times. The fact that both AGN feedback and environmental effects are stronger in higher density environments leads to a correlation between the quenching of central and satellite galaxies which roughly reproduces observed conformity trends inside haloes.

astro-ph.GA

Galaxy formation in the Planck cosmology II. Star formation histories and post-processing magnitude reconstruction

We adapt the L-Galaxies semi-analytic model to follow the star-formation histories (SFH) of galaxies -- by which we mean a record of the formation time and metallicities of the stars that are present in each galaxy at a given time. We use these to construct stellar spectra in post-processing, which offers large efficiency savings and allows user-defined spectral bands and dust models to be applied to data stored in the Millennium data repository. We contrast model SFHs from the Millennium Simulation with observed ones from the VESPA algorithm as applied to the SDSS-7 catalogue. The overall agreement is good, with both simulated and SDSS galaxies showing a steeper SFH with increased stellar mass. The SFHs of blue and red galaxies, however, show poor agreement between data and simulations, which may indicate that the termination of star formation is too abrupt in the models. The mean star-formation rate (SFR) of model galaxies is well-defined and is accurately modelled by a double power law at all redshifts: SFR proportional to $1/(x^{-1.39}+x^{1.33})$, where $x=(t_a-t)/3.0\,$Gyr, $t$ is the age of the stars and $t_a$ is the loopback time to the onset of galaxy formation; above a redshift of unity, this is well approximated by a gamma function: SFR proportional to $x^{1.5}e^{-x}$, where $x=(t_a-t)/2.0\,$Gyr. Individual galaxies, however, show a wide dispersion about this mean. When split by mass, the SFR peaks earlier for high-mass galaxies than for lower-mass ones, and we interpret this downsizing as a mass-dependence in the evolution of the quenched fraction: the SFHs of star-forming galaxies show only a weak mass dependence.

astro-ph.GA

The EAGLE simulations of galaxy formation: public release of halo and galaxy catalogues

We present the public data release of halo and galaxy catalogues extracted from the EAGLE suite of cosmological hydrodynamical simulations of galaxy formation. These simulations were performed with an enhanced version of the GADGET code that includes a modified hydrodynamics solver, time-step limiter and subgrid treatments of baryonic physics, such as stellar mass loss, element-by-element radiative cooling, star formation and feedback from star formation and black hole accretion. The simulation suite includes runs performed in volumes ranging from 25 to 100 comoving megaparsecs per side, with numerical resolution chosen to marginally resolve the Jeans mass of the gas at the star formation threshold. The free parameters of the subgrid models for feedback are calibrated to the redshift z=0 galaxy stellar mass function, galaxy sizes and black hole mass - stellar mass relation. The simulations have been shown to match a wide range of observations for present-day and higher-redshift galaxies. The raw particle data have been used to link galaxies across redshifts by creating merger trees. The indexing of the tree produces a simple way to connect a galaxy at one redshift to its progenitors at higher redshift and to identify its descendants at lower redshift. In this paper we present a relational database which we are making available for general use. A large number of properties of haloes and galaxies and their merger trees are stored in the database, including stellar masses, star formation rates, metallicities, photometric measurements and mock gri images. Complex queries can be created to explore the evolution of more than 10^5 galaxies, examples of which are provided in appendix. (abridged)

astro-ph.GA

Galaxy formation in the Planck Cosmology - I. Matching the observed evolution of star formation rates, colours and stellar masses

We have updated the Munich galaxy formation model to the Planck first-year cosmology, while modifying the treatment of baryonic processes to reproduce recent data on the abundance and passive fractions of galaxies from z= 3 down to z=0. Matching these more extensive and more precise observational results requires us to delay the reincorporation of wind ejecta, to lower the surface density threshold for turning cold gas into stars, to eliminate ram-pressure stripping in haloes less massive than ~10^14 Msun, and to modify our model for radio mode feedback. These changes cure the most obvious failings of our previous models, namely the overly early formation of low-mass galaxies and the overly large fraction of them that are passive at late times. The new model is calibrated to reproduce the observed evolution both of the stellar mass function and of the distribution of star formation rate at each stellar mass. Massive galaxies (M>10^11 [Msun]) assemble most of their mass before z=1 and are predominantly old and passive at z=0, while lower mass galaxies assemble later and, for M<10^9.5 (Msun), are still predominantly blue and star forming at z=0. This phenomenological but physically based model allows the observations to be interpreted in terms of the efficiency of the various processes that control the formation and evolution of galaxies as a function of their stellar mass, gas content, environment and time.

astro-ph.GA

IVOA Recommendation: Simulation Data Model

In this document and the accompanying documents we describe a data model (Simulation Data Model) describing numerical computer simulations of astrophysical systems. The primary goal of this standard is to support discovery of simulations by describing those aspects of them that scientists might wish to query on, i.e. it is a model for meta-data describing simulations. This document does not propose a protocol for using this model. IVOA protocols are being developed and are supposed to use the model, either in its original form or in a form derived from the model proposed here, but more suited to the particular protocol. The SimDM has been developed in the IVOA Theory Interest Group with assistance of representatives of relevant working groups, in particular DM and Semantics.

astro-ph.IM

IVOA Recommendation: IVOA Photometry Data Model

The Photometry Data Model (PhotDM) standard describes photometry filters, photometric systems, magnitude systems, zero points and its interrelation with the other IVOA data models through a simple data model. Particular attention is given necessarily to optical photometry where specifications of magnitude systems and photometric zero points are required to convert photometric measurements into physical flux density units.

astro-ph.IM

On the Spin Bias of Satellite Galaxies in the Local Group-like Environment

We utilize the Millennium-II simulation databases to study the spin bias of dark subhalos in the Local Group-like systems which have two prominent satellites with comparable masses. Selecting the group-size halos with total mass similar to that of the Local Group (LG) from the friends-of-friends halo catalog and locating their subhalos from the substructure catalog, we determine the most massive (main) and second to the most massive (submain) ones among the subhalos hosted by each selected halo. When the dimensionless spin parameter (lambda) of each subhalo is derived from its specific angular momentum and circular velocity at virial radius, a signal of correlation is detected between the spin parameters of the subhalos and the main-to-submain mass ratios of their host halos at z=0: The higher main-to-submain mass ratio a host halo has, the higher mean spin parameter its subhalos have. It is also found that the correlations exist even for the subhalo progenitors at z=0.5 and z=1. Our interpretation of this result is that the subhalo spin bias is not a transient effect but an intrinsic property of a LG-like system with higher main-to- submain mass ratio, caused by stronger anisotropic stress in the region. A cosmological implication of our result is also discussed.

astro-ph.CO

Galaxy formation in WMAP1 and WMAP7 cosmologies

Using the technique of Angulo & White (2010) we scale the Millennium and Millennium-II simulations of structure growth in a LCDM universe from the cosmological parameters with which they were carried out (based on first-year results from the Wilkinson Microwave Anisotropy Probe, WMAP1) to parameters consistent with the seven-year WMAP data (WMAP7). We implement semi-analytic galaxy formation modelling on both simulations in both cosmologies to investigate how the formation, evolution and clustering of galaxies are predicted to vary with cosmological parameters. The increased matter density Omega_m and decreased linear fluctuation amplitude sigma8 in WMAP7 have compensating effects, so that the abundance and clustering of dark halos are predicted to be very similar to those in WMAP1 for z <= 3. As a result, local galaxy properties can be reproduced equally well in the two cosmologies by slightly altering galaxy formation parameters. The evolution of the galaxy populations is then also similar. In WMAP7, structure forms slightly later. This shifts the peak in cosmic star formation rate to lower redshift, resulting in slightly bluer galaxies at z=0. Nevertheless, the model still predicts more passive low-mass galaxies than are observed. For rp< 1Mpc, the z=0 clustering of low-mass galaxies is weaker for WMAP7 than for WMAP1 and closer to that observed, but the two cosmologies give very similar results for more massive galaxies and on large scales. At z>1 galaxies are predicted to be more strongly clustered for WMAP7. Differences in galaxy properties, including, clustering, in these two cosmologies are rather small up to redshift 3. Given that there are still considerable residual uncertainties in galaxy formation models, it is very difficult to distinguish WMAP1 from WMAP7 through observations of galaxy properties or their evolution.

astro-ph.CO

Confronting theoretical models with the observed evolution of the galaxy population out to z=4

[abridged] We construct lightcones for the semi-analytic galaxy formation simulation of Guo et al. (2011) and make mock catalogues for comparison with deep high-redshift surveys. Photometric properties are calculated with two different stellar population synthesis codes (Bruzual & Charlot 2003; Maraston 2005) in order to study sensitivity to this aspect of the modelling. The catalogues are publicly available and include photometry for a large number of observed bands from 4000°A to 6μm, as well as rest-frame photometry and intrinsic properties of the galaxies. Guo et al. (2011) tuned their model to fit the low-redshift galaxy population but noted that at z > 1 it overpredicts the abundance of galaxies below the "knee" of the stellar mass function. Here we extend the comparison to deep galaxy counts in the B, i, J, K and IRAC 3.6μm, 4.5μm and 5.8μm bands, to the redshift distributions of K and 5.8μm selected galaxies, and to the evolution of rest-frame luminosity functions in the B and K bands. The B, i and J counts are well reproduced, but at longer wavelengths the overabundant high-redshift galaxies produce excess faint counts. The predicted redshift distributions for K and 5.8μm selected samples highlight the effect of emission from thermally pulsing AGB stars. The full treatment of Maraston (2005) predicts three times as many z~2 galaxies in faint 5.8μm selected samples as the model of Bruzual & Charlot (2003), whereas the two models give similar predictions for K-band selected samples. Although luminosity functions are adequately reproduced out to z~3 in rest-frame B, the same is true at rest-frame K only if TP-AGB emission is included, and then only at high luminosity. Fainter than L* the two synthesis models agree but overpredict the number of galaxies, another reflection of the overabundance of ~10^10M\odot model galaxies at z > 1.

astro-ph.CO

Simulations of the galaxy population constrained by observations from z=3 to the present day: implications for galactic winds and the fate of their ejecta

We apply Monte Carlo Markov Chain (MCMC) methods to large-scale simulations of galaxy formation in a LambdaCDM cosmology in order to explore how star formation and feedback are constrained by the observed luminosity and stellar mass functions of galaxies. We build models jointly on the Millennium and Millennium-II simulations, applying fast sampling techniques which allow observed galaxy abundances over the ranges 7<log(M*/Msun)<12 and z=0 to z=3 to be used simultaneously as constraints in the MCMC analysis. When z=0 constraints alone are imposed, we reproduce the results of previous modelling by Guo et al. (2012), but no single set of parameters can reproduce observed galaxy abundances at all redshifts simultaneously, reflecting the fact that low-mass galaxies form too early and thus are overabundant at high redshift in this model. The data require the efficiency with which galactic wind ejecta are reaccreted to vary with redshift and halo mass quite differently than previously assumed, but in a similar way as in some recent hydrodynamic simulations of galaxy formation. We propose a specific model in which reincorporation timescales vary inversely with halo mass and are independent of redshift. This produces an evolving galaxy population which fits observed abundances as a function of stellar mass, B- and K-band luminosity at all redshifts simultaneously. It also produces a significant improvement in two other areas where previous models were deficient. It leads to present day dwarf galaxy populations which are younger, bluer, more strongly star-forming and more weakly clustered on small scales than before, although the passive fraction of faint dwarfs remains too high.

astro-ph.CO

Observing simulated galaxy clusters with PHOX: a novel X-ray photon simulator

We present a novel, virtual X-ray observatory designed to obtain synthetic observations from hydro-numerical simulations, named PHOX. In particular, we provide a description of the code constituting the photon simulator and of the new approach implemented. We apply PHOX to simulated galaxy clusters in order to demonstrate its capabilities. In fact, X-ray observations of clusters of galaxies continue to provide us with an increasingly detailed picture of their structure and of the underlying physical phenomena governing the gaseous component, which dominates their baryonic content. Therefore, it is fundamental to find the most direct and faithful way to compare such observational data with hydrodynamical simulations of cluster-like objects, which can currently include various complex physical processes. Here, we present and analyse synthetic Suzaku observations of two cluster-size haloes obtained by processing with PHOX the hydrodynamical simulation of the large-scale, filament-like region in which they reside. Taking advantage of the simulated data, we test the results inferred from the X-ray analysis of the mock observations against the underlying, known solution. Remarkably, we are able to recover the theoretical temperature distribution of the two haloes by means of the multi-temperature fitting of the synthetic spectra. Moreover, the shapes of the reconstructed distributions allow us to trace the different thermal structure that distinguishes the dynamical state of the two haloes.

astro-ph.CO