SearcharxivSearch

arXiv subjects

Andrew Kirby

Publications and source records attributed to Andrew Kirby.

11 recordsLinked to original sources

Validation and extension of an analytic momentum availability model for the two-scale momentum theory of wind farm flows

A key parameter in the two-scale momentum theory of wind farm flows is the momentum availability, which quantifies the supply of momentum to a wind farm from various different momentum transport mechanisms (advection, pressure gradient, Coriolis, turbulence and unsteadiness). In this study, the contribution of each of these mechanisms to the momentum availability is evaluated directly from large-eddy simulation (LES) data in order to validate an analytic momentum availability model (Kirby, Dunstan, & Nishino, J. Fluid Mech., vol. 976, 2023, A24). Application of the model to six wind farm cases, three with different atmospheric boundary-layer (ABL) heights and three with different turbine layouts, shows that the full model performs well across all cases, but that its linearized version increasingly overpredicts the momentum availability for increasing ABL heights. It is found that the overprediction is related to the ABL Rossby number, and based on this observation, we propose an extension of the original linear model, which improves its accuracy for the considered cases and makes it more generally applicable, in particular to cases with tall ABL heights or strong Coriolis forcing.

physics.flu-dyn

Retuve: Automated Multi-Modality Analysis of Hip Dysplasia with Open Source AI

Developmental dysplasia of the hip (DDH) poses significant diagnostic challenges, hindering timely intervention. Current screening methodologies lack standardization, and AI-driven studies suffer from reproducibility issues due to limited data and code availability. To address these limitations, we introduce Retuve, an open-source framework for multi-modality DDH analysis, encompassing both ultrasound (US) and X-ray imaging. Retuve provides a complete and reproducible workflow, offering open datasets comprising expert-annotated US and X-ray images, pre-trained models with training code and weights, and a user-friendly Python Application Programming Interface (API). The framework integrates segmentation and landmark detection models, enabling automated measurement of key diagnostic parameters such as the alpha angle and acetabular index. By adhering to open-source principles, Retuve promotes transparency, collaboration, and accessibility in DDH research. This initiative has the potential to democratize DDH screening, facilitate early diagnosis, and ultimately improve patient outcomes by enabling widespread screening and early intervention. The GitHub repository/code can be found here: https://github.com/radoss-org/retuve

eess.IV

An analytical model of momentum availability for predicting large wind farm power

Turbine-wake and farm-atmosphere interactions influence wind farm power production. For large offshore farms, the farm-atmosphere interaction is usually the more significant effect. This study proposes an analytical model of the `momentum availability factor' to predict the impact of farm-atmosphere interactions. It models the effects of net advection, pressure gradient forcing and turbulent entrainment, using steady quasi-1D flow assumptions. Turbulent entrainment is modelled by assuming self-similar vertical shear stress profiles. We used the model with the `two-scale momentum theory' to predict the power of large finite-sized farms. The model compared well with existing results of large-eddy simulations (LES) of finite wind farms in conventionally neutral boundary layers. The model captured most of the effects of atmospheric boundary layer (ABL) height on farm performance by considering the undisturbed vertical shear stress profile of the ABL as an input. In particular, the model predicted the power of staggered wind farms with a typical error of 5% or less. The developed model provides a novel way of instantly predicting the power of large wind farms, including the farm blockage effects. A further simplification of the model to analytically predict the 'wind extractability factor' is also presented. This study provides a novel framework for modelling farm-atmosphere interactions. Future studies can use the framework to better model large wind farms.

physics.flu-dyn

Validated respiratory drug deposition predictions from 2D and 3D medical images with statistical shape models and convolutional neural networks

For the one billion sufferers of respiratory disease, managing their disease with inhalers crucially influences their quality of life. Generic treatment plans could be improved with the aid of computational models that account for patient-specific features such as breathing pattern, lung pathology and morphology. Therefore, we aim to develop and validate an automated computational framework for patient-specific deposition modelling. To that end, an image processing approach is proposed that could produce 3D patient respiratory geometries from 2D chest X-rays and 3D CT images. We evaluated the airway and lung morphology produced by our image processing framework, and assessed deposition compared to in vivo data. The 2D-to-3D image processing reproduces airway diameter to 9% median error compared to ground truth segmentations, but is sensitive to outliers of up to 33% due to lung outline noise. Predicted regional deposition gave 5% median error compared to in vivo measurements. The proposed framework is capable of providing patient-specific deposition measurements for varying treatments, to determine which treatment would best satisfy the needs imposed by each patient (such as disease and lung/airway morphology). Integration of patient-specific modelling into clinical practice as an additional decision-making tool could optimise treatment plans and lower the burden of respiratory diseases.

physics.med-ph

Data-driven modelling of turbine wake interactions and flow resistance in large wind farms

Turbine wake and local blockage effects are known to alter wind farm power production in two different ways: (1) by changing the wind speed locally in front of each turbine; and (2) by changing the overall flow resistance in the farm and thus the so-called farm blockage effect. To better predict these effects with low computational costs, we develop data-driven emulators of the `local' or `internal' turbine thrust coefficient $C_T^*$ as a function of turbine layout. We train the model using a multi-fidelity Gaussian Process (GP) regression with a combination of low (engineering wake model) and high-fidelity (Large-Eddy Simulations) simulations of farms with different layouts and wind directions. A large set of low-fidelity data speeds up the learning process and the high-fidelity data ensures a high accuracy. The trained multi-fidelity GP model is shown to give more accurate predictions of $C_T^*$ compared to a standard (single-fidelity) GP regression applied only to a limited set of high-fidelity data. We also use the multi-fidelity GP model of $C_T^*$ with the two-scale momentum theory (Nishino \& Dunstan 2020, J. Fluid Mech. 894, A2) to demonstrate that the model can be used to give fast and accurate predictions of large wind farm performance under various mesoscale atmospheric conditions. This new approach could be beneficial for improving annual energy production (AEP) calculations and farm optimisation in the future.

physics.flu-dyn

Two-scale interaction of wake and blockage effects in large wind farms

Turbine wake and farm blockage effects may significantly impact the power produced by large wind farms. In this study, we perform Large-Eddy Simulations (LES) of 50 infinitely large offshore wind farms with different turbine layouts and wind directions. The LES results are combined with the two-scale momentum theory (Nishino & Dunstan 2020, J. Fluid Mech. 894, A2) to investigate the aerodynamic performance of large but finite-sized farms as well. The power of infinitely large farms is found to be a strong function of the array density, whereas the power of large finite-sized farms depends on both the array density and turbine layout. An analytical model derived from the two-scale momentum theory predicts the impact of array density very well for all 50 farms investigated and can therefore be used as an upper limit to farm performance. We also propose a new method to quantify turbine-scale losses (due to turbine-wake interactions) and farm-scale losses (due to the reduction of farm-average wind speed). They both depend on the strength of atmospheric response to the farm, and our results suggest that, for large offshore wind farms, the farm-scale losses are typically more than twice as large as the turbine-scale losses. This is found to be due to a two-scale interaction between turbine wake and farm induction effects, explaining why the impact of turbine layout on farm power varies with the strength of atmospheric response.

physics.flu-dyn

Accuracy and Performance Comparison of Video Action Recognition Approaches

Over the past few years, there has been significant interest in video action recognition systems and models. However, direct comparison of accuracy and computational performance results remain clouded by differing training environments, hardware specifications, hyperparameters, pipelines, and inference methods. This article provides a direct comparison between fourteen off-the-shelf and state-of-the-art models by ensuring consistency in these training characteristics in order to provide readers with a meaningful comparison across different types of video action recognition algorithms. Accuracy of the models is evaluated using standard Top-1 and Top-5 accuracy metrics in addition to a proposed new accuracy metric. Additionally, we compare computational performance of distributed training from two to sixty-four GPUs on a state-of-the-art HPC system.

cs.CV

Best of Both Worlds: High Performance Interactive and Batch Launching

Rapid launch of thousands of jobs is essential for effective interactive supercomputing, big data analysis, and AI algorithm development. Achieving thousands of launches per second has required hardware to be available to receive these jobs. This paper presents a novel preemptive approach to implement spot jobs on MIT SuperCloud systems allowing the resources to be fully utilized for both long running batch jobs while still providing fast launch for interactive jobs. The new approach separates the job preemption and scheduling operations and can achieve 100 times faster performance in the scheduling of a job with preemption when compared to using the standard scheduler-provided automatic preemption-based capability. The results demonstrate that the new approach can schedule interactive jobs preemptively at a performance comparable to when the required computing resources are idle and available. The spot job capability can be deployed without disrupting the interactive user experience while increasing the overall system utilization.

cs.DC

Fast Mapping onto Census Blocks

Pandemic measures such as social distancing and contact tracing can be enhanced by rapidly integrating dynamic location data and demographic data. Projecting billions of longitude and latitude locations onto hundreds of thousands of highly irregular demographic census block polygons is computationally challenging in both research and deployment contexts. This paper describes two approaches labeled "simple" and "fast". The simple approach can be implemented in any scripting language (Matlab/Octave, Python, Julia, R) and is easily integrated and customized to a variety of research goals. This simple approach uses a novel combination of hierarchy, sparse bounding boxes, polygon crossing-number, vectorization, and parallel processing to achieve 100,000,000+ projections per second on 100 servers. The simple approach is compact, does not increase data storage requirements, and is applicable to any country or region. The fast approach exploits the thread, vector, and memory optimizations that are possible using a low-level language (C++) and achieves similar performance on a single server. This paper details these approaches with the goal of enabling the broader community to quickly integrate location and demographic data.

cs.DC

Multi-Temporal Analysis and Scaling Relations of 100,000,000,000 Network Packets

Our society has never been more dependent on computer networks. Effective utilization of networks requires a detailed understanding of the normal background behaviors of network traffic. Large-scale measurements of networks are computationally challenging. Building on prior work in interactive supercomputing and GraphBLAS hypersparse hierarchical traffic matrices, we have developed an efficient method for computing a wide variety of streaming network quantities on diverse time scales. Applying these methods to 100,000,000,000 anonymized source-destination pairs collected at a network gateway reveals many previously unobserved scaling relationships. These observations provide new insights into normal network background traffic that could be used for anomaly detection, AI feature engineering, and testing theoretical models of streaming networks.

cs.NI

TapirXLA: Embedding Fork-Join Parallelism into the XLA Compiler in TensorFlow Using Tapir

This work introduces TapirXLA, a replacement for TensorFlow's XLA compiler that embeds recursive fork-join parallelism into XLA's low-level representation of code. Machine-learning applications rely on efficient parallel processing to achieve performance, and they employ a variety of technologies to improve performance, including compiler technology. But compilers in machine-learning frameworks lack a deep understanding of parallelism, causing them to lose performance by missing optimizations on parallel computation. This work studies how Tapir, a compiler intermediate representation (IR) that embeds parallelism into a mainstream compiler IR, can be incorporated into a compiler for machine learning to remedy this problem. TapirXLA modifies the XLA compiler in TensorFlow to employ the Tapir/LLVM compiler to optimize low-level parallel computation. TapirXLA encodes the parallelism within high-level TensorFlow operations using Tapir's representation of fork-join parallelism. TapirXLA also exposes to the compiler implementations of linear-algebra library routines whose parallel operations are encoded using Tapir's representation. We compared the performance of TensorFlow using TapirXLA against TensorFlow using an unmodified XLA compiler. On four neural-network benchmarks, TapirXLA speeds up the parallel running time of the network by a geometric-mean multiplicative factor of 30% to 100%, across four CPU architectures.

cs.PF