SearcharxivSearch

arXiv subjects

Chris Burr

Publications and source records attributed to Chris Burr.

4 recordsLinked to original sources

Data Preservation in High Energy Physics: Global Report 2026

This document summarizes the contributions to the 5th DPHEP workshop March 5-6, 2026, CERN, and reflects the advancements since 2024, as well as future milestones and tendencies. Impressive progress in HEP data preservation is observed. Legacy data revival was showcased through successful reanalysis of archived data using contemporary methods, demonstrating the long-term scientific value of preservation. Sustainability challenges were noted, emphasizing the need for long-term funding and institutional support to maintain data preservation infrastructure, particularly for legacy experiments transitioning to archival modes. Innovative transverse projects display constant progress towards common technologies for a robust and transferrable DP. In particular, there is a clear shift toward automation, with increasing use of AI and machine learning for data curation, metadata extraction, and workflow optimization. Open science momentum is growing, with wider adoption of FAIR principles and open data policies, and experiments committing to public releases.

hep-ex

A framework for assuring the accuracy and fidelity of an AI-enabled Digital Twin of en route UK airspace

Digital Twins combine simulation, operational data and Artificial Intelligence (AI), and have the potential to bring significant benefits across the aviation industry. Project Bluebird, an industry-academic collaboration, has developed a probabilistic Digital Twin of en route UK airspace as an environment for training and testing AI Air Traffic Control (ATC) agents. There is a developing regulatory landscape for this kind of novel technology. Regulatory requirements are expected to be application specific, and may need to be tailored to each specific use case. We draw on emerging guidance for both Digital Twin development and the use of Artificial Intelligence/Machine Learning (AI/ML) in Air Traffic Management (ATM) to present an assurance framework. This framework defines actionable goals and the evidence required to demonstrate that a Digital Twin accurately represents its physical counterpart and also provides sufficient functionality across target use cases. It provides a structured approach for researchers to assess, understand and document the strengths and limitations of the Digital Twin, whilst also identifying areas where fidelity could be improved. Furthermore, it serves as a foundation for engagement with stakeholders and regulators, supporting discussions around the regulatory needs for future applications, and contributing to the emerging guidance through a concrete, working example of a Digital Twin. The framework leverages a methodology known as Trustworthy and Ethical Assurance (TEA) to develop an assurance case. An assurance case is a nested set of structured arguments that provides justified evidence for how a top-level goal has been realised. In this paper we provide an overview of each structured argument and a number of deep dives which elaborate in more detail upon particular arguments, including the required evidence, assumptions and justifications.

cs.AI

The LHCb Sprucing and Analysis Productions

The LHCb detector underwent a comprehensive upgrade in preparation for the third data-taking run of the Large Hadron Collider (LHC), known as LHCb Upgrade I. The increased data rate of Run 3 not only posed data collection (Online) challenges but also significant Offline data processing and analysis ones. The offline processing and analysis model was consequently upgraded to handle the factor 30 increase in data volume and the associated demands of ever-growing analyst-level datasets, led by the LHCb Data Processing and Analysis (DPA) project. This paper documents the LHCb "Sprucing" - the centralised offline processing, selections and streaming of data - and "Analysis Productions" - the centralised and highly automated declarative nTuple production system. The DaVinci application used by analysis productions for tupling spruced data is described as well as the apd and lbconda tools for data retrieval and analysis environment configuration. These tools allow for greatly improved analyst workflows and analysis preservation. Finally, the approach to data processing and analysis in the High-Luminosity Large Hadron Collider (HL-LHC) era - LHCb Upgrade II - is discussed.

hep-ex

The Scikit HEP Project -- overview and prospects

Scikit-HEP is a community-driven and community-oriented project with the goal of providing an ecosystem for particle physics data analysis in Python. Scikit-HEP is a toolset of approximately twenty packages and a few "affiliated" packages. It expands the typical Python data analysis tools for particle physicists. Each package focuses on a particular topic, and interacts with other packages in the toolset, where appropriate. Most of the packages are easy to install in many environments; much work has been done this year to provide binary "wheels" on PyPI and conda-forge packages. The Scikit-HEP project has been gaining interest and momentum, by building a user and developer community engaging collaboration across experiments. Some of the packages are being used by other communities, including the astroparticle physics community. An overview of the overall project and toolset will be presented, as well as a vision for development and sustainability.

physics.comp-ph