Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Performance Trace”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

Extraction of the muon signals recorded with the surface detector of the Pierre Auger Observatory using recurrent neural networks

The Pierre Auger Observatory, at present the largest cosmic-ray observatory ever built, is instrumented with a ground array of 1600 water-Cherenkov detectors, known as the Surface Detector (SD). The SD samples the secondary particle content (mostly photons, electrons, positrons and muons) of extensive air showers initiated by cosmic rays with energies ranging from 1017eV up to more than 1020eV. Measuring the independent contribution of the muon component to the total registered signal is crucial to enhance the capability of the Observatory to estimate the mass of the cosmic rays on an event-by-event basis. However, with the current design of the SD, it is difficult to straightforwardly separate the contributions of muons to the SD time traces from those of photons, electrons and positrons. In this paper, we present a method aimed at extracting the muon component of the time traces registered with each individual detector of the SD using Recurrent Neural Networks. We derive the performances of the method by training the neural network on simulations, in which the muon and the electromagnetic components of the traces are known. We conclude this work showing the performance of this method on experimental data of the Pierre Auger Observatory. We find that our predictions agree with the parameterizations obtained by the AGASA collaboration to describe the lateral distributions of the electromagnetic and muonic components of extensive air showers.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

A Roadmap for Edge Computing Enabled Automated Multidimensional Transmission Electron Microscopy

The advent of modern, high-speed electron detectors has made the collection of multidimensional hyperspectral transmission electron microscopy datasets, such as 4D-STEM, a routine. However, many microscopists find such experiments daunting since analysis, collection, long-term storage, and networking of such datasets remain challenging. Some common issues are their large and unwieldy size that often are several gigabytes, non-standardized data analysis routines, and a lack of clarity about the computing and network resources needed to utilize the electron microscope. The existing computing and networking bottlenecks introduce significant penalties in each step of these experiments, and thus, real-time analysis-driven automated experimentation for multidimensional TEM is challenging. One solution is to integrate microscopy with edge computing, where moderately powerful computational hardware performs the preliminary analysis before handing off the heavier computation to high-performance computing (HPC) systems. In this work, we trace the roots of computation in modern electron microscopy, demonstrate deep learning experiments running on an edge system, and discuss the networking requirements for tying together microscopes, edge computers, and HPC systems.

47 OTHER INSTRUMENTATION↗

Exceptional Electrical Detection of Trace NO 2 via Mixed Metal MOF-on-MOF Film-Based Sensors

The tunability of metal–organic frameworks (MOFs) makes them exceptional materials for the development of highly selective, low-power sensors for toxic gas detection. Herein, we demonstrate enhanced detection of NO 2 gas by a MOF-based electrical impedance sensor made using a unique mixed metal MOF-on-MOF synthesis. For this work, a combined experimental and computational study was performed using the exemplar Ni x Mg 1–x -MOF-74 to understand the fundamental structure–property relationships behind metal mixing and MOF film synthesis methods on sensor performance. Density functional theory results indicated that the presence of Ni in Mg-MOF-74 increased framework stability and increased the electron density of states at lower energies near the HOMO, as well as enhanced the NO 2 –Mg adsorption interaction. Impedance data of the Ni x Mg 1–x -MOF-74 films with larger Ni contents showed greater impedance change after exposure to 1 ppm of NO 2 gas. Furthermore, when synthesized through either a drop-cast or direct solvothermal film growth approach, the monometallic Ni-based sensors had the best performance. However, the mixed metal Ni x Mg 1–x -MOF-74 sensors synthesized through a MOF-on-MOF approach resulted in the highest impedance change, outperforming all monometallic Ni-based sensors. In particular, the mixed metal Ni-on-Mg-MOF-74 film was the best-performing sensor with an impedance change of 309 upon trace NO 2 exposure. Change in impedance response after NO 2 exposure was improved by 52% compared to the best monometallic Ni-on-Ni-MOF-74 sensor. Structural analysis of the Ni-on-Mg film showed that the first Mg-MOF-74 layer acts as a structural template controlling the structural features of the final film after metal exchange with Ni. This led to improved film quality, evidenced by the greater crystallinity and larger MOF grain sizes, and resulted in enhanced sensor performance which was not achievable through other metal mixing methods. Altogether, this study identifies structure–property relationships and synthetic templating methods that inform MOF-based sensor design, allowing for improved detection of toxic compounds.

36 MATERIALS SCIENCE↗

Exploratory Studies of Mechanical Properties, Residual Stress, and Grain Evolution at the Local Regions near Pores in Additively Manufactured Metals

Directed energy deposition (DED) is gaining widespread acceptance in various industrial applications since its unique manufacturing features allow the DED to print metallic parts with very complex geometries. However, DED inevitably generates a lot of internal pores which can limit the widespread applications of the DED technique. The current studies on DED porosity are mostly focused on analyzing pores’ bulk-scale influences on mechanical properties and performances. Since DED pores have a micro-scale existence, with dimensions ranging from a few microns to several hundred microns, it is fundamental to explore the pores’ influences on the micro-scale, including local mechanical properties, residual stress, and grains near pores. However, this important research direction has been neglected. The objective of this work is to fill the above gap in DED porosity research and acquire a fundamental understanding of the role of porosity on a microscopic scale. The authors used nanoindentation approaches to investigate internal pores’ effects on mechanical properties and residual stress in local regions surrounding the pores. In addition, the grains near pores were observed through EBSD, and simulated with the Kinetic Monte Carlo model. The research findings can be provided for DED researchers and industrial practitioners as technical guidance. Most importantly, the research results can work as a good reference for tracing the source of bulk-scale mechanical performances and properties of DED parts with internal pores.

36 MATERIALS SCIENCE↗

Thermal Analysis of a Solid Particle Light-Trapping Planar Cavity Receiver Using Computational Fluid Dynamics

Concentrated solar power (CSP) is one of the most effective ways of harnessing solar power to create efficient, durable, and resilient energy systems. This study entails thermal modeling and analysis of a novel central tower receiver configuration. This receiver uses solid particles as the heat transfer fluid (HTF), a promising option for third-generation CSP systems. The configuration considered here is the light-trapping planar cavity receiver (LTPCR) introduced by the National Renewable Energy Laboratory. While heat transfer studies of various LTPCR subsystems have been done, system-level thermal analysis of the LTPCR receiver has not been attempted. This study also presents important sensitivity analyses of the operating parameters of the CSP system, which can help guide the design of future central tower receivers. This study employs Ansys Fluent as a computational fluid dynamics (CFD) tool to model fluid dynamics and heat transfer in the receiver, intending to quantify its thermal performance. The model seamlessly integrates Monte Carlo ray tracing data, which generates absorbed solar flux profiles from the heliostat field design, with the heat transfer characteristics of the fluidized particle bed. This unified model is designed to accurately predict the thermal behavior of the LTPCR. Analysis of preliminary results reveals that the primary loss mechanisms are radiative and natural convective losses, in that order. Based on observations from a baseline case, several strategies are suggested and numerically tested. These solutions include selective cooling of high-temperature regions and manipulation of particle bed parameters. Selective cooling of high-temperature regions reduced the peak temperature by 151 degrees C and decreased thermal losses by 0.9%. Improving the particle-wall heat transfer coefficient (P-W HTC) of the particle bed decreased the thermal losses by 1.7% and decreased the peak temperatures by 57 degrees C. Decreasing the particle inlet temperature (PIT) also reduced thermal losses by 3.5% and decreased peak temperatures by 29 degrees C. Compounding these strategies improved the thermal losses of the receiver from 13.5% in the baseline case to 7.5%. Additionally, the study explores the variation in thermal performance across different locations of the receiver, where a variation of thermal losses from 12.9% to 17.3% is found. This allows a comprehensive evaluation of potential improvements in efficiency and temperature management.

computational fluid dynamics↗

Analysis of dislocation configurations in SiC crystals through X-ray topography aided by ray tracing simulations

Silicon carbide as a wide bandgap semiconductor is of great research interest for its widespread deployment in a range of electronic and optoelectronic devices, particularly in power electronics. However, defects in silicon carbide crystals are still major concerns that is hampering the development of high-performance devices. X-ray topography, particularly using the synchrotron beam has been instrumental in characterizing and analyzing defect configurations in silicon carbide crystals to optimize crystal growth as well as understand the effect of defects on device performance. Here, in recent years, the use of ray-tracing simulation technique based on the orientation contrast mechanism to simulate contrast of defects observed on actual X-ray topographs has proven to be an effective approach to investigate the nature of crystallographic defects in various semiconductors. This review discusses the principle of ray-tracing simulation and its application and modifications to incorporate the effects of surface relaxation and photoelectric absorption to better simulate different dislocations observed in 4H–SiC as well as 6H–SiC crystals of various orientations. The adaptation to weak beam topography and plane wave topography is also discussed. The application of ray-tracing simulation in dislocation characterization of silicon carbide of different polytypes is systematically reviewed including different types of dislocations observed in both off-axis wafers and axial-sliced samples through synchrotron X-ray topography under various beam conditions, recording geometries and reflections. The result of ray-tracing simulation is further utilized in other studies including the investigation of effective penetration depth of all types of dislocations lying on the basal plane on grazing-incidence X-ray topography.

Engineering↗

Hybrid Solar System (Final Scientific/Technical Report)

GTI Energy (GTI) teamed with the University of California at Merced (UCM) to scaleup the hybrid solar system (HSS) technology for demonstrating its performance at the US Gypsum (USG) plant in Plaster City, California. The technology integrates two-stage concentrating solar collector with matching particle thermal transport and storage (TSS) system to deliver cost-effective, and on-demand distributed high temperature industrial process heat up to 600°C with solar thermal, in this case to a gypsum kettle, to reduce its fuel use and carbon footprint. Current solar technologies, which reach these temperatures, are not distributable (towers) or cost-effective (dish). The research team developed a conceptual system design for host site retrofit, including preliminary heat balance, process flow diagram, particle to process heat exchanger and equipment placements at the site. Subsequently, parallel efforts were carried out at UCM to design, build and test a 12 m long commercial scale prototype concentrating thermal-only collector system and at GTI to design, build and test a matching 650°C capable particle TTS system. The nominal 50 kWth collector consists of a parabolic trough and three 4 m long two-stage receivers in series. Prior to on-sun testing, a 4 m long receiver was fabricated and successfully tested at 650 °C in a laboratory setting for 100 hrs of continuous operation showing less than 15% radiation loss. A 7 m wide x 17 m long parabolic trough was then installed at UCM for on-sun testing of the 12 m long receiver, and concurrently several 4 m long receivers were built. The optics of the parabolic trough were calibrated, and on-sun test were carried out on 12 m long receivers. During tests, the intense solar radiation (53x) caused the absorber tubes in the receivers to bend, reducing the overall optical efficiency. To address the bending issue, a self-consistent algorithm that includes ray tracing, thermal and deformation models was developed to perform thermal stress analysis on absorbers for parabolic solar collectors. Results obtained with this algorithm showed a dramatic rise in deformation as absorber tube length increases. A combined efficiency parameter that includes the occluded area for the mounts was developed to obtain an optimized tube length obtained. Based on the results, a length of 2.7 m for the absorber + 0.2 m for the coupler was chosen to minimize any bending and optimize optical efficiency while maintaining ease of mounting. The associated particle TTS system was designed, built and successfully tested at GTI. It includes storage, receiving and lock hoppers and piping that simulates the transfer of captured solar energy to an actual industrial furnace. Tests over 77 charge-discharge cycles demonstrated <2% particle degradation, with no problematic particle accumulations and no flow interruptions. The piping pressure drop was about 5 psi. The team also worked with Stanley Consultants (Stanley) to prepare conceptual and preliminary engineering packages to facilitate follow-on development and commercialization efforts. These include process and instrumentation diagram’s (P&ID’s), general arrangements, electrical one-line, project definitions document, equipment data sheets, schedule, and construction cost estimate for 2 MWth system. Updated HSS technology commercialization and customer engagement plans and detailed costs and evaluated market trade-offs and manufacturing.

03 NATURAL GAS↗

Assessment of the CTF subchannel code for modeling a large-break loss-of-coolant accident reflood transient

With increased industry interest in extending reactor operating cycles, the Nuclear Energy Advanced Modeling and Simulation (NEAMS) program has been investigating the behavior of high-burnup fuel during design basis accidents such as the large-break loss-of-coolant accident (LBLOCA) with consideration for risk of fuel fragmentation, relocation, and dispersal (FFRD). As part of that activity, the NEAMS subchannel thermal/ hydraulics (T/H) code, CTF, is being used for modeling of LBLOCA and to determine the impact of subchannel resolution on results. Although CTF includes a wide range of models for LBLOCA conditions, the code has not been used for this application while maintained at Oak Ridge National Laboratory (ORNL) until now. Therefore, here, in this work, a preliminary assessment of several of these models was performed using openly available reflood experimental data from the Flooding Experiments in Blocked Arrays (FEBA) tests. One coarse mesh and one fine mesh model were set up in CTF for high and low flooding rate tests performed in the unblocked FEBA facility. A coarse TRACE model was set up to be as consistent as possible with the coarse CTF model to allow for code-to-code benchmarking. The assessment shows a tendency of the codes to over-predict peak cladding temperature (PCT) near the top of the bundle and to quench early. Advanced spacer grid models were shown to improve upper bundle predictions in CTF. The resolved CTF model over-predicted PCT by a larger degree in the center channels in the low-flooding rate test, and it is believed that the radiative heat transfer model, which was not used in this study, may be needed to correct this over-prediction. Finally, this work demonstrates the importance of the droplet model in determining quench time and vapor temperature and PCT prediction, which necessitates a more in-depth validation of these models in the future.

21 SPECIFIC NUCLEAR REACTORS AND ASSOCIATED PLANTS↗

First high-power helicon results from DIII-D

Abstract More than 0.6 MW of rf power at 476 MHz has been coupled to DIII-D plasmas by launching helicon (whistler) waves with a traveling-wave antenna (comb-line) in the fast-wave polarization (Van Compernolle et al 2021 Nucl. Fusion 61 116034) which resulted in the observation of electron heating of the core plasma with single-pass absorption based on ray-tracing in L-mode discharges. The coupling performance of the 1.5 m wide 30-element comb-line traveling-wave antenna has been consistent with expectations based on the 2015–2016 experiments on DIII-D with a low-power 12-element prototype (Pinsker et al 2018 Nucl. Fusion 58 106007). The conditioning process that was necessary to carry out high-power experiments is discussed; rf-specific impurities have not been observed. Parametric decay instabilities have been observed and are being investigated as a potential edge absorption mechanism (Porkolab et al 2023 AIP Conf. Proc. 2984 070004).

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Indoor Wireless Localization of Uncooperative Sources Using a Ray Tracing Model

Indoor geolocation of radio frequency (RF) trans- mitters is challenging due to site-specific multipath effects, especially when the sources are uncooperative. A source being uncooperative means that a-priori the transmit time, location, and power are unknown. It is proposed to use a site-specific ray tracing model to generate realistic indoor RF responses, perform geolocation using multiple methods, and compare their performance. This work will look at two types of geolocation algorithms: (1) time-difference of arrival (TDOA) and (2) re- ceived signal strength Indicator (RSSI) fingerprinting or pattern- matching-based methods. An indoor space will be simulated with a grid of fixed receivers and a grid of transmit locations using Remcom’s Wireless InSite. Using the output of the Wireless InSite simulation the response from any given transmit location can be generated and used to evaluate geolocation performance.

indoor source localization, Wireless InSite, re- c↗

A Scalable Gaussian Process Approach to Shear Mapping with MuyGPs

Analysis of cosmic shear is an integral part of understanding structure growth across cosmic time, which in turn provides us with information about the nature of dark energy. Conventional methods generate shear maps from which we can infer the matter distribution in the universe. Current methods (e.g., Kaiser–Squires inversion) for generating these maps, however, are tricky to implement and can introduce bias. Recent alternatives construct a spatial process prior for the lensing potential, which allows for inference of the convergence and shear parameters given lensing shear measurements. Realizing these spatial processes, however, scales cubically in the number of observations—an unacceptable expense as near-term surveys expect billions of correlated measurements. Therefore, we present a linearly scaling shear map construction alternative using a scalable Gaussian process prior called MuyGPs. MuyGPs avoids cubic scaling by conditioning interpolation on only nearest neighbors and fits hyperparameters using batched leave-one-out cross-validation. This work is the first step toward a full, scalable mass mapping method. We work in a simplified regime where we validate our method by interpolating and analyzing maps given noisy point-estimate data from all three shear fields, taken from a suite of N -body ray-tracing simulations. We also show that we can perform these operations at the scale of billions of galaxies on high-performance computing platforms.

79 ASTRONOMY AND ASTROPHYSICS↗

NLR HPC Kestrel Jobs Data

Overview: Anonymized job-level records from the Kestrel HPC system at the National Laboratory of the Rockies (NLR). Each record represents a Slurm batch job with scheduling metadata, resource requests, utilization, energy estimates, and efficiency metrics. Sensitive fields (user, account, job name, submit line, working directory, submit script, and job type) are replaced with 7-character cryptographic hashes. System & Timeframe: Kestrel is located at the NLR campus. Standard compute nodes have 104 cores and 256 GB RAM; bigmem nodes have 2,000 GB. GPU nodes (gpu-h100 partition) use NVIDIA H100 GPUs. Data covers jobs submitted August 2023 through December 2025. Funding provided by the U.S. Department of Energy, EERE. Files: esif.hpc.kestrel.job-anon.zip — Anonymized job records (Hive-partitioned Parquet) datacard.md — Full dataset documentation ~11 million rows, 50 variables. Readable with PyArrow, pandas, DuckDB, Apache Spark, or any Parquet-compatible tool. Data Collection: Jobs collected via sacct with timezone-aware export (SLURM_TIME_FORMAT="%Y-%m-%dT%H:%M:%S%z"), loaded into PostgreSQL. Calculated columns updated via database triggers and batch functions. All timestamps use timestamptz and correctly handle DST transitions. Preprocessing: Anonymization of name, user, account, submit_line, work_dir, submit_script, and job_type via 7-char hex hashes Derived columns: queue_wait, cpu_eff, max/min/avg_mem_eff, energy estimates Simplified job state mapping (e.g., "CANCELLED by 132357" → "CANCELLED") Boolean flags: python_job, reframe_job Temporal decomposition: year, month, day, day_of_week, hour, minute from submit_time Shared node tracking: shared_job_count, nodes_shared, jobs_shared Key Variables: Scheduling: job_id, partition, state_simple, submit_time, start_time, end_time, queue_wait Resources: nodes_req/used, processors_req/used, memory_req, wallclock_req/used, gpus_requested Efficiency: cpu_eff, max/min/avg_mem_eff Energy: cpu_energy_tdp_estimated_max/used_watt_hours, consumed_energy_raw_joules, consumed_energy_raw_watt_hours Sharing: shared_job_count, nodes_shared, jobs_shared Partitions: short, standard, debug, gpu-h100 Job States: CANCELLED, COMPLETED, FAILED, PENDING, RUNNING QoS Levels: normal, high Important Notes: Timestamps include timezone offsets; DST transitions are handled correctly, though adding intervals across DST boundaries requires offset adjustment shared_job_count reflects physical node co-residency, not use of the shared partition Job step records and raw Slurm JSONB fields are excluded Do not attempt to re-identify individuals from hashed fields

97 MATHEMATICS AND COMPUTING↗

NLR HPC Eagle Jobs Data and Additional Energy Metrics

Overview: Anonymized job-level records from the Eagle high-performance computing (HPC) system at the National Laboratory of the Rockies (NLR). Each record represents a Slurm batch job with scheduling metadata, resource requests, resource utilization, CPU/GPU energy consumption, and efficiency metrics. Sensitive fields (user, account, job name) are replaced with cryptographic hashes. System & Timeframe: Eagle was a 2,000-node, 8-petaflop system operated at NLR from 2019–2024. Data covers the full operational lifetime of the system. Slurm data was processed nightly; timestamps are in Mountain Time. Funding provided by the U.S. Department of Energy, EERE. Files: esif.hpc.eagle.job-anon.zip — Core anonymized job records (Hive-partitioned Parquet) esif.hpc.eagle.job-anon-energy-metrics.zip — Same records with additional iLO and Ganglia energy metrics datacard.md — Full dataset documentation ~13.8 million rows, 62 variables. Readable with PyArrow, pandas, DuckDB, Apache Spark, or any Parquet-compatible tool. Data Collection: Jobs collected via sacct through a pipeline: Eagle Jobs API → Redpanda → StreamSets → HPCMON API → PostgreSQL. Node-level power from iLO (HP Integrated Lights-Out); GPU power from Ganglia monitoring, joined to jobs via node lists and time ranges. Preprocessing: Anonymization of name, user, and account fields via cryptographic hashing Derived columns: queue_wait, cpu_eff, max_mem_eff Simplified job state mapping (e.g., "CANCELLED BY 12345" → "CANCELLED") QoS accounting rules (buy-in, standby, or Slurm QoS value) CPU energy estimated from TDP (200W, Intel Xeon Gold 6154, 18 cores) Timezone-aware columns (_tz) sourced from LEX accounting database to correctly handle DST transitions Key Variables: Scheduling: job_id, partition, state_simple, submit_time_tz, start_time_tz, end_time_tz, queue_waitResources: nodes_req/used, processors_req/used, memory_req, wallclock_req/used, gpus_requested Efficiency: cpu_eff, max_mem_eff Energy: cpu_energy_tdp_estimated_max/used_watt_hours, node_energy_total_watt_hours (iLO), gpu0/1_energy_total_watt_hours (Ganglia) Partitions: bigmem, bigmem-8600, bigscratch, csc, dav, ddn, debug, gpu, haswell, long, mono, short, standard Job States: CANCELLED, COMPLETED, FAILED, NODE_FAIL, OUT_OF_MEMORY, PENDING, RUNNING, TIMEOUT QoS Levels: Unknown, normal, buy-in, debug, penalty, high, standby Important Notes: Non-_tz timestamp columns may be off by one hour across DST boundaries; use _tz columns for time difference calculations Energy fields are null for jobs without monitoring coverage Job step records and raw Slurm JSONB fields are excluded from this extract Do not attempt to re-identify individuals from hashed fields

97 MATHEMATICS AND COMPUTING↗

Developing a Nuclear Quality Assurance Compliant Design Methodology for Neutronic Analysis of Xe-100 Design

The primary objective of this work is to develop a design methodology compliant with nuclear quality assurance standards for the Xe-100 neutronic design verification studies. To achieve this, a Monte Carlo model of the Xe-100 reactor was constructed using the exclusion principle, transformation technique, and universe-based level specification following Idaho National Laboratory (INL) NQA level-1 compliant standards and an NQA-1 compliant version of MCNP6. The model encompasses the entire reactor core structures, including the upper plenum, core region, and lower plenum sections, along with all sub-components. The active core section was represented using the spectral regions, each comprising a particular fuel composition and temperature averaged over the considered zone, calculated by X-energy using Very Superior Old Programs (VSOP). Additionally, a component-wise temperature map was implemented into the model, not only for the core region but also for the structural components. Temperature-dependent cross-section libraries, along with thermal scattering law libraries, generated using INL NQA-1 compliant version of NJOY21, were utilized for each isotope in the burnt fuel and the structural materials. Furthermore, the volume of each modeled component was estimated using a stochastic approach with the ray tracing method in MCNP and criticality calculations were performed.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

A new highly enriched 233 U reference material for improved simultaneous determination of uranium amount and isotope amount ratios in trace level samples

A highly-enriched 233 U reference material (>0.99987 n( 233 U)/n(U)) has been prepared and characterized for use as an isotope dilution mass spectrometry spike. An ion exchange separation was performed on 1 g of high purity 233 U to further reduce trace amounts of contaminant Pu in the material. The purified 233 U was then prepared as a master solution which was analyzed for molality of uranium by modified Davies and Gray titration. A portion of the master solution was quantitatively diluted and dispensed for reference material units. Selected units were analyzed for verification of uranium amount and to characterize uranium isotope amount ratios by multi-collector inductively couple plasma mass spectrometry. Furthermore, modelling of spike-corrected isotopic data show that the new spike will enable simultaneous measurements of uranium amount and isotope amount ratios with resulting uncertainties that are substantially less sensitive to over spiking than widely used 233 U certified reference materials.

233U↗

Suppressing ion migration in metal halide perovskite via interstitial doping with a trace amount of multivalent cations

Cations with suitable sizes to occupy an interstitial site of perovskite crystals have been widely used to inhibit ion migration and promote the performance and stability of perovskite optoelectronics. However, such interstitial doping inevitably leads to lattice microstrain that impairs the long-range ordering and stability of the crystals, causing a sacrificial trade-off. Here, we unravel the evident influence of the valence states of the interstitial cations on their efficacy to suppress the ion migration. Incorporation of a trivalent neodymium cation (Nd 3+ ) effectively mitigates the ion migration in the perovskite lattice with a reduced dosage (0.08%) compared to a widely used monovalent cation dopant (Na + , 0.45%). As a result, the photovoltaic performances and operational stability of the prototypical perovskite solar cells are enhanced with a trace amount of Nd 3+ doping while minimizing the sacrificial trade-off.

36 MATERIALS SCIENCE↗

Towards a new generation long trace profiler LTP-2020: optical design of pencil beam interferometry sensor

Improvements in the quality of synchrotron beamline x-ray optics required for next-generation light sources (e.g. the ALS Upgrade project) drive the need to improve the performance of the metrology instrumentation used to measure these components. The Long Trace Profiler (LTP) that is in use at many synchrotron metrology laboratories around the world has some known issues that affect the accuracy of its measurements. The main error source is optical path difference (OPD) phase error introduced into the probe beam by inhomogeneities in the glass components used in the optical head. We have developed a new optical head design, LTP-2020, that replaces the cube polarizing beamsplitter (PBS) with a thin wedge plate polarizing beamsplitter (WPBS) and replaces the cemented doublet lens with an aspheric singlet. Both of these components significantly reduce the glass volume traversed by the laser probe beam. Careful attention to ghost ray interference produced by back reflection from optical surfaces is necessary to minimize distortion in the primary image that translates into systematic error in the slope angle measurement. We make extensive use of a commercial raytracing program to model the back reflections and adjust component parameters as necessary to minimize distortion. Deliberate misalignment of components is necessary to make the system perform correctly. Stringent requirements are placed on the 45◦ incidence coatings on the WPBS and on the normal incidence coatings on the lens and camera window elements. We encourage our colleagues who wish to upgrade their current LTP systems to join us in the procurement of these custom optical components.

Takacs, Peter Z.↗

An Integrated Framework for Memory-Centric Analysis: From Trace Collection to Co-Design

The memory wall phenomenon—where advances in processor performance significantly outpace those in memory subsystems—poses a fundamental challenge for contemporary computing systems. In memory-bound applications, memory subsystem behavior dominates performance, yet existing analysis approaches present significant limitations: detailed microarchitectural simulators require days to weeks to simulate modest workloads; hardware performance counters provide only aggregate statistics that obscure temporal and spatial access patterns; and scaled simulation approaches face challenges in capturing certain behaviors that emerge at larger scales. These limitations reflect a processor-centric design philosophy increasingly misaligned with memory-bound workloads where detailed understanding of memory access patterns, cache hierarchy interactions, and contention is critical for effective optimization. This paper presents an integrated framework for memory-centric analysis that enables effective hardware-software co-design. We describe practical trace collection techniques, including hardware-assisted processor tracing with minimal overhead and portable software-based instrumentation with statistical sampling. We present multi-perspective analysis methods that examine memory behavior from temporal, sequential, spatial, and relational viewpoints, revealing distinct optimization opportunities invisible in aggregate metrics. We detail an architectural modeling framework that uses sampled traces with temporal interpolation and confidence-based filtering to evaluate cache and memory configurations. Evaluation on representative benchmarks demonstrates that this framework achieves practical accuracy (L2 cache errors of 2.64\%, confidence-filtered L3 errors of 9.92\%, bandwidth errors of 7.33\%) while providing substantial speedup (26.8×) over cycle-accurate simulation, enabling rapid design space exploration. We demonstrate how this integrated framework enables systematic identification of both hardware optimizations (memory controller tuning, bank partitioning, NUMA configuration) and software optimizations (data layout restructuring, prefetching strategies, memory-aware scheduling). Through this comprehensive treatment of the memory-centric analysis pipeline—from trace collection through architectural modeling to co-design application—we provide researchers and practitioners with practical techniques for addressing memory bottlenecks in contemporary computing systems.

Gajaria, Dhruv Mayur↗