Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “HEP”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 235 records · Page 13

dCache: The Storage System of Choice for Data-Intensive Applications

The ever-increasing volumes of data produced by modern scientific facilities like EuXFEL and LHC put significant stress on data management infrastructure operated by laboratories and research centers. The challenges to be addressed span the entire data life cycle, from ingest and efficient data analysis to long-term preservation, typically involving large tape libraries. dCache, a storage system developed in collaboration between the Deutsches Elektronen-Synchrotron (DESY), Fermi National Accelerator Laboratory, and Nordic e-Infrastructure Collaboration (NeIC), is designed to manage a large number of disk servers and to facilitate transparent data migration to and from archival storage. Its multifaceted approach offers a unified method to support a variety of scientific use cases with the same storage infrastructure, including high-throughput data ingest, data sharing over wide area networks, efficient access from HPC clusters, and long-term data preservation on tertiary storage. Initially developed for high energy physics (HEP) experiments, dCache is now used by various scientific communities, including astrophysics, biomedical research, and life sciences, each having specific requirements. This paper presents architecture, deployment strategies, performance and scalability enhancements, and recent advancements in dCache addressing the needs of scientific communities. Finally, we touch on the development and release process, ensuring the software’s high quality.

DCache↗

Feasibility of using crystal channeling for the beam loss mitigation in slow extraction at 8GeV

In the accelerator applications of slow extraction for High Energy Physics (HEP), one of the prime challenges is the mitigation of beam losses, which is becoming increasingly critical with the continuous rise in beam power. A significant breakthrough was achieved earlier at CERN through the successful deployment of proton beam channeling in a crystal at 400GeV, effectively diverting the beam away from extraction septa and offering new avenues for improving slow extraction efficiency. However, a crucial question remains whether this approach is still effective at lower and medium proton beam energies. In this paper, we present the promising results of the computer simulation studies of the septum shadowing at 8GeV for the Mu2e project slow extraction at Fermilab. In conclusion, across the wide range of beam parameters, the beam loss reduction is shown to range between 40% and factor of three.

43 PARTICLE ACCELERATORS↗

Performance of a triple-GEM detector with capacitive-sharing 3-coordinate (X–Y–U)-strip anode readout

The concept of capacitive-sharing readout, described in detail in a previous study, offers the possibility for the development of high-performance three-coordinates (X--Y--U)-strip readout for Micro Pattern Gaseous Detectors (MPGDs) using simple standard PCB fabrication techniques. Capacitive-sharing (X--Y--U)-strip readout allows simultaneous measurement of the Cartesian coordinates x and y of the position of the particles together with a third coordinate u along the diagonal axis in a single readout PCB. This provides a powerful tool to address multiple-hit ambiguity and enable pattern recognition capabilities in moderate particle flux environment of collider or fixed target experiments in high energy physics HEP) and nuclear physics (NP). We present in this paper the performance of a 10 cm × 10 cm triple-GEM detector with capacitive-sharing (X--Y--U)-strip anode readout. Spatial resolutions of the order of $\sigma_{x}^{res}$ = 71.6 $\pm$ 0.8 $\mu$m for X-strips, $\sigma_{y}^{res}$ = 56.2 $\pm$ 0.9 $\mu$m for Y-strips and $\sigma_{u}^{res}$ = 75.2 $\pm$ 0.9 $\mu$m for U-strips have been obtained at a beam test at Thomas Jefferson National Accelerator Facility (Jefferson Lab). Modifications of the readout design of future prototypes to improve the spatial resolution and challenges in scaling to large-area MPGDs are discussed.

(X-Y-U) strip↗

Modeling of surface-state induced inter-electrode isolation of n -on- p devices in mixed-field and γ -irradiation environments

Position sensitive n-on-p silicon sensors will be utilized in the tracker and in the High Granularity Calorimeter (HGCAL) of the Compact Muon Solenoid (CMS) experiment at High Luminosity Large Hadron Collider (HL-LHC). The detrimental effect of the radiation-induced accumulation of positive net oxide charge on position resolution in n-on-p sensors has typically been countered by the application of isolation implants like p-stop or p-spray between n + -electrodes. In addition to the positively charged layer inside the oxide and close to the Si/SiO 2 -interface, surface damage introduced by ionizing radiation in SiO 2 -passivated silicon particle detectors includes the accumulation of trapped-oxide-charge and interface traps. A previous study of either n/γ (mixed field)- or γ-irradiated Metal-Oxide-Semiconductor (MOS) capacitors showed evidence of substantially higher introduction rates of acceptor- and donor-type deep interface traps (N it,acc/don ) in mixed-field environment. Here, in this work, an inter-pad and -strip resistance (or resistivity (ρ int )) simulation study of n-on-p sensors with and without p-stop isolation implants was conducted for both irradiation types. Higher levels of ρint showed correlation to higher densities of deep N it,acc/don , with the inter-pad isolation performance of the mixed-field irradiated sensors becoming independent of the presence of p-stop implant between the n + - electrodes up to about 100 kGy. The low introduction rates of deep N it,acc/don in γ-irradiated sensors resulted in high sensitivity of ρ int to the presence and peak doping of p-stop above the lowest dose of about 7 kGy in the study. As a consequence of the advantageous influence of radiation-induced accumulation of deep N it on the inter-electrode isolation, position sensitive n-on-p sensors without isolation implants may be considered for future HEP-experiments where the radiation is largely due to hadrons.

Akchurin, N. [Texas Tech Univ., Lubbock, TX (Unite↗

Determining the density of the sun with neutrinos

The discovery of solar neutrinos confirmed that the inner workings of the Sun generally match our theoretical understanding of the fusion process. Solar neutrinos have also played a role in discovering that neutrinos have mass and that they oscillate. We combine the latest solar neutrino data along with other oscillation data from reactors to determine the Sun's density profile. We derive constraints given the current data and show the anticipated improvements with more reactor neutrino data from JUNO constraining the true oscillation parameters and more solar neutrino data from DUNE which should provide a crucial measurement of hep neutrinos.

79 ASTRONOMY AND ASTROPHYSICS↗

Highly efficient photon detection systems for noble liquid detectors based on perovskite quantum dots

Abstract Wavelength shifting photon detection systems (PDS) are the critical functioning components in noble liquid detectors used for high energy physics (HEP) experiments and dark matter search. The vacuum ultraviolet (VUV) scintillation light emitted by these Liquid argon (LAr) and liquid Xenon (LXe) detectors are shifted to higher wavelengths resulting in its efficient detection using the state-of-the-art photodetectors such as silicon photomultipliers (SiPM). The currently used organic wavelength shifting materials [such as 1,1,4,4 Tetraphenyl Butadiene (TPB)] have several disadvantages and are unreliable for longterm use. In this study, we demonstrate the application of the inorganic perovskite cesium lead bromide (CsPbBr 3 ) quantum dots (QDs) as highly efficient wavelength shifters. The absolute photoluminescence quantum yield of the PDS fabricated using these QDs exceeds 70%. CsPbBr 3 -based PDS demonstrated an enhancement in the SiPM signal enhancement by up to 3 times when compared to a 3 µm-thick TPB-based PDS. The emission spectrum from the QDs was optimized to match the highest quantum efficiency region of the SiPMs. In addition, we have demonstrated the deposition of the QD-based wavelength shifting material on a large area PDS substrate using low capital cost and widely scalable solution-based techniques providing a pathway appropriate for meter-scale PDS fabrication and widespread use for other wavelength shifting applications.

47 OTHER INSTRUMENTATION↗

Heterogeneous data-processing optimization with CLARA’s adaptive workflow orchestrator

The hardware landscape used in HEP and NP is changing from homogeneous multi-core systems towards heterogeneous systems with many different computing units, each with their own characteristics. To achieve maximum performance with data processing, the main challenge is to place the right computing on the right hardware. In this paper, we discuss CLAS12 charge particle tracking workflow orchestration that allows us to utilize both CPU and GPU to improve the performance. The tracking application algorithm was decomposed into micro-services that are deployed on CPU and GPU processing units, where the best features of both are intelligently combined to achieve maximum performance. In this heterogeneous environment, CLARA aims to match the requirements of each micro-service to the strength of a CPU or a GPU architecture. A predefined execution of a micro-service on a CPU or a GPU may not be the most optimal solution due to the streaming data-quantum size and the data-quantum transfer latency between CPU and GPU. So, the CLARA workflow orchestrator is designed to dynamically assign micro-service execution to a CPU or a GPU, based on the online benchmark results analyzed for a period of real-time data-processing.

Gyurjyan, Vardan↗

dCache: Inter-disciplinary storage system

The dCache project provides open-source software deployed internationally to satisfy ever more demanding storage requirements. Its multifaceted approach provides an integrated way of supporting different use-cases with the same storage, from high throughput data ingest, data sharing over wide area networks, efficient access from HPC clusters and long term data persistence on a tertiary storage. Though it was originally developed for the HEP experiments, today it is used by various scientific communities, including astrophysics, biomed, life science, which have their specific requirements. In this paper we describe some of the new requirements as well as demonstrate how dCache developers are addressing them.

Mkrtchyan, Tigran↗

Grid-based minimization at scale: Feldman-Cousins corrections for light sterile neutrino search

High Energy Physics (HEP) experiments generally employ sophisticated statistical methods to present results in searches of new physics. In the problem of searching for sterile neutrinos, likelihood ratio tests are applied to short-baseline neutrino oscillation experiments to construct confidence intervals for the parameters of interest. The test statistics of the form Δχ2 is often used to form the confidence intervals, however, this approach can lead to statistical inaccuracies due to the small signal rate in the region-of-interest. In this paper, we present a computational model for the computationally expensive Feldman-Cousins corrections to construct a statistically accurate confidence interval for neutrino oscillation analysis. The program performs a grid-based minimization over oscillation parameters and is written in C++. Our algorithms make use of vectorization through Eigen3, yielding a single-core speed-up of 350 compared to the original implementation, and achieve MPI data parallelism by employing DIY. We demonstrate the strong scaling of the application at High-Performance Computing (HPC) sites. We utilize HDF5 along with HighFive to write the results of the calculation to file.

Wospakrik, Marianette↗

Performance of CUDA Unified Memory in CMS Heterogeneous Pixel Reconstruction

The management of separate memory spaces of CPUs and GPUs brings an additional burden to the development of software for GPUs. To help with this, CUDA unified memory provides a single address space that can be accessed from both CPU and GPU. The automatic data transfer mechanism is based on page faults generated by the memory accesses. This mechanism has a performance cost, that can be with explicit memory prefetch requests. Various hints on the inteded usage of the memory regions can also be given to further improve the performance. The overall effect of unified memory compared to an explicit memory management can depend heavily on the application. In this paper we evaluate the performance impact of CUDA unified memory using the heterogeneous pixel reconstruction code from the CMS experiment as a realistic use case of a GPU-targeting HEP reconstruction software. We also compare the programming model using CUDA unified memory to the explicit management of separate CPU and GPU memory spaces.

Kortelainen, Matti J.↗

An Error Analysis Toolkit for Binned Counting Experiments

We introduce the MINERvA Analysis Toolkit (MAT), a utility for centralizing the handling of systematic uncertainties in HEP analyses. The fundamental utilities of the toolkit are the MnvHnD, a powerful histogram container class, and the systematic Universe classes, which provide a modular implementation of the many universe error analysis approach. These products can be used stand-alone or as part of a complete error analysis prescription. They support the propagation of systematic uncertainty through all stages of analysis, and provide flexibility for an arbitrary level of user customization. This extensible solution to error analysis enables the standardization of systematic uncertainty definitions across an experiment and a transparent user interface to lower the barrier to entry for new analyzers.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Integration of Rucio Metadata in Belle II

Rucio is a Data Management software that has become a de-facto standard in the HEP community and beyond. It allows the management of large volumes of data over their full lifecycle. The Belle II experiment located at KEK (Japan) recently moved to Rucio to manage its data over the coming decade (O(10) PB/year). In addition to its Data Management functionalities, Rucio also provides support for storing generic metadata. Rucio metadata already provides accurate accounting of the data stored all over the sites serving Belle II. Annotating files with generic metadata opens up possibilities for finer-grained metadata query support. We will first introduce some of the new developments aimed at providing good performance that were done to cover Belle II use-cases like bulk insert methods, metadata inheritance, etc. We will then describe the various tests performed to validate Rucio generic metadata at Belle II scale (O(100M) files), detailing the import and performance tests that were made.

97 MATHEMATICS AND COMPUTING↗

BigPanDA monitoring system evolution in the ATLAS Experiment

Monitoring services play a crucial role in the day-to-day operation of distributed computing systems. The ATLAS Experiment at LHC uses the Production and Distributed Analysis workload management system (PanDA WMS), which allows a million computational jobs to run daily at over 170 computing centers of the WLCG and opportunistic resources, utilizing 600k cores simultaneously on average. The BigPanDA monitor is an essential part of the monitoring infrastructure for the ATLAS Experiment that provides a wide range of views, from top-level summaries to a single computational job and its logs. Over the past few years of the PanDA WMS advancement in the ATLAS Experiment, several new components were developed, such as Harvester, iDDS, Data Carousel, and Global Shares. Due to its modular architecture, the BigPanDA monitor naturally grew into a platform where the relevant data from all PanDA WMS components and accompanying services are accumulated and displayed in the form of interactive charts and tables. Moreover the system has been adopted by other experiments beyond HEP. In this paper we describe the evolution of the BigPanDA monitor system, the development of new modules, and the integration process into other experiments.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

ROOT’s RNTuple I/O Subsystem: The Path to Production

The RNTuple I/O subsystem is ROOT’s future event data file format and access API. It is driven by the expected data volume increase at upcoming HEP experiments, e.g. at the HL-LHC, and recent opportunities in the storage hardware and software landscape such as NVMe drives and distributed object stores. RNTuple is a redesign of the TTree binary format and API and has shown to deliver substantially faster data throughput and better data compression both compared to TTree and to industry standard formats. In order to let HENP computing workflows benefit from RNTuple’s superior performance, however, the I/O stack needs to connect efficiently to the rest of the ecosystem, from grid storage to (distributed) analysis frameworks to (multithreaded) experiment frameworks for reconstruction and ntuple derivation. With the RNTuple binary format soon arriving at its first production release, we present RNTuple’s feature set, integration efforts, and its performance impact on the time-to-solution. We show the latest performance figures of RDataFrame analysis code of realistic complexity, comparing RNTuple and TTree as data sources. We discuss RNTuple’s approach to functionality critical to the HENP I/O (such as multithreaded writes, fast data merging, schema evolution) and we provide an outlook on the road to its use in production.

Blomer, Jakob↗

Evaluating Performance Portability with the CMS Heterogeneous Pixel Reconstruction code

In the past years the landscape of tools for expressing parallel algorithms in a portable way across various compute accelerators has continued to evolve significantly. There are many technologies on the market that provide portability between CPU, GPUs from several vendors, and in some cases even FPGAs. These technologies include C++ libraries such as Alpaka and Kokkos, compiler directives such as OpenMP, the SYCL open specification that can be implemented as a library or in a compiler, and standard C++ where the compiler is solely responsible for the offloading. Given this developing landscape, users have to choose the technology that best fits their applications and constraints. For example, in the CMS experiment the experience so far in heterogeneous reconstruction algorithms suggests that the full application contains a large number of relatively short computational kernels and memory transfer operations. In this work we use a stand-alone version of the CMS heterogeneous pixel reconstruction code as a realistic use case of HEP reconstruction software that is capable of leveraging GPUs effectively. We summarize the experience of porting this code base from CUDA to Alpaka, Kokkos, SYCL, std::par, and OpenMP offloading. We compare the event processing throughput achieved by each version on NVIDIA and AMD GPUs as well as on a CPU, and compare those to what a native version of the code achieves on each platform.

Andriotis, Nikolaos↗

Performance of Heterogeneous Algorithm Scheduling in CMSSW

The CMS experiment started to utilize Graphics Processing Units (GPU) to accelerate the online reconstruction and event selection running on its High Level Trigger (HLT) farm in the 2022 data taking period. The projections of the HLT farm to the High-Luminosity LHC foresee a significant use of compute accelerators in the LHC Run 4 and onwards in order to keep the cost, size, and power budget of the farm under control. This direction of leveraging compute accelerators has synergies with the increasing use of HPC resources in HEP computing, as HPC machines are employing more and more compute accelerators that are predominantly GPUs today. In this work we review the features developed for the CMS data processing framework, CMSSW, to support the effective utilization of both compute accelerators and many-core CPUs within a highly concurrent task-based framework. We measure the impact of various design choices for the scheduling of heterogeneous algorithms on the event processing throughput, using the Run-3 HLT application as a realistic use case.

Bocci, Andrea↗

dCache project status and update

The dCache project delivers an open-source, massively scalable, distributed storage system deployed internationally to satisfy today’s scientists’ ever-demanding storage requirements. Its multifaceted approach supports different use cases with the same storage, from high throughput data ingest, data sharing over wide area networks, efficient access from HPC clusters, and longterm data persistence on tertiary storage. Even though dCache was initially developed for HEP experiments, today, it is used by various scientific communities, including astrophysics, biomed, and life science, each with their specific requirements. To match the needs of these new communities and keep up with the scaling demands of existing experiments, dCache is permanently evolving. With this contribution, we would like to highlight the recent developments in dCache regarding integration with CERN Tape Archive (CTA), advanced metadata handling, token-based authorization support, bulk API for QoS transitions, REST API to control interaction with the tape system, and future development directions.

Mkrtchyan, Tigran [DESY]↗

Using the ATLAS experiment software on heterogeneous resources

With the large dataset expected from 2030 onwards by the HL-LHC at CERN, the ATLAS experiment is reaching the limits of the current data processing model in terms of traditional CPU resources based on x86_64 architectures and an extensive program for software upgrades towards the HL-LHC has been set up. The ARM CPU architecture is becoming a competitive and energy efficient alternative. Accelerators like GPUs are available in any recent HPC. In the past years ATLAS has successfully ported its full data processing and simulation software framework Athena to ARM and has invested significant effort in porting parts of the reconstruction and simulation algorithms to GPUs. We report on the successful usage of the ATLAS experiment offline and online software framework Athena on ARM and GPUs through the PanDA workflow management system at various WLCG sites. Furthermore we report on performance optimizations of the builds for ARM CPUs and the GPU integration efforts. We will discuss performance comparisons of different ARM and x86_64 architectures on WLCG resources and Cloud compute providers like GCP and AWS using ATLAS productions workflows as used in the Hep-Score23 benchmark suite.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗