Search arXivSearch

arXiv subjects

Jan Strube

Publications and source records attributed to Jan Strube.

At least 19 recordsLinked to original sources

The Imaging Time-of-Propagation Detector at Belle II

We report on the construction, operation, and performance of the Time-of-Propagation detector with imaging used for the Belle II experiment running at the Super-KEKB $e^+e^-$ collider. This detector is located in the central barrel region and uses Cherenkov light to provide particle identification among hadrons. The Cherenkov light is radiated in highly polished bars of synthetic fused silica (quartz) and transported to the ends of the bars via total internal reflection. One bar end is instrumented with finely segmented micro-channel-plate photomultiplier tubes to record the light, while the other end has a mirror attached to reflect the photons back to the instrumented end. Both the propagation times and hit positions of the Cherenkov photons are measured; these depend on the Cherenkov angle and together provide good discrimination among charged pions, kaons, and protons with momenta up to around 4 GeV/$c$. To date, the detector has been used to record and analyze almost 600 fb$^{-1}$ of Belle II data.

hep-ex

SuperSAM: Crafting a SAM Supernetwork via Structured Pruning and Unstructured Parameter Prioritization

Neural Architecture Search (NAS) is a powerful approach of automating the design of efficient neural architectures. In contrast to traditional NAS methods, recently proposed one-shot NAS methods prove to be more efficient in performing NAS. One-shot NAS works by generating a singular weight-sharing supernetwork that acts as a search space (container) of subnetworks. Despite its achievements, designing the one-shot search space remains a major challenge. In this work we propose a search space design strategy for Vision Transformer (ViT)-based architectures. In particular, we convert the Segment Anything Model (SAM) into a weight-sharing supernetwork called SuperSAM. Our approach involves automating the search space design via layer-wise structured pruning and parameter prioritization. While the structured pruning applies probabilistic removal of certain transformer layers, parameter prioritization performs weight reordering and slicing of MLP-blocks in the remaining layers. We train supernetworks on several datasets using the sandwich rule. For deployment, we enhance subnetwork discovery by utilizing a program autotuner to identify efficient subnetworks within the search space. The resulting subnetworks are 30-70% smaller in size compared to the original pre-trained SAM ViT-B, yet outperform the pretrained model. Our work introduces a new and effective method for ViT NAS search-space design.

cs.CV

AI-Enabled Operations at Fermi Complex: Multivariate Time Series Prediction for Outage Prediction and Diagnosis

The Main Control Room of the Fermilab accelerator complex continuously gathers extensive time-series data from thousands of sensors monitoring the beam. However, unplanned events such as trips or voltage fluctuations often result in beam outages, causing operational downtime. This downtime not only consumes operator effort in diagnosing and addressing the issue but also leads to unnecessary energy consumption by idle machines awaiting beam restoration. The current threshold-based alarm system is reactive and faces challenges including frequent false alarms and inconsistent outage-cause labeling. To address these limitations, we propose an AI-enabled framework that leverages predictive analytics and automated labeling. Using data from $2,703$ Linac devices and $80$ operator-labeled outages, we evaluate state-of-the-art deep learning architectures, including recurrent, attention-based, and linear models, for beam outage prediction. Additionally, we assess a Random Forest-based labeling system for providing consistent, confidence-scored outage annotations. Our findings highlight the strengths and weaknesses of these architectures for beam outage prediction and identify critical gaps that must be addressed to fully harness AI for transitioning downtime handling from reactive to predictive, ultimately reducing downtime and improving decision-making in accelerator management.

cs.LG

Final Report for CHESS: Cloud, High-Performance Computing, and Edge for Science and Security

Automating the theory-experiment cycle requires effective distributed workflows that utilize a computing continuum spanning lab instruments, edge sensors, computing resources at multiple facilities, data sets distributed across multiple information sources, and potentially cloud. Unfortunately, the obvious methods for constructing continuum platforms, orchestrating workflow tasks, and curating datasets over time fail to achieve scientific requirements for performance, energy, security, and reliability. Furthermore, achieving the best use of continuum resources depends upon the efficient composition and execution of workflow tasks, i.e., combinations of numerical solvers, data analytics, and machine learning. Pacific Northwest National Laboratory's LDRD "Cloud, High-Performance Computing (HPC), and Edge for Science and Security" (CHESS) has developed a set of interrelated capabilities for enabling distributed scientific workflows and curating datasets. This report describes the results and successes of CHESS from the perspective of open science.

cs.DC

WeQA: A Benchmark for Retrieval Augmented Generation in Wind Energy Domain

Wind energy project assessments present significant challenges for decision-makers, who must navigate and synthesize hundreds of pages of environmental and scientific documentation. These documents often span different regions and project scales, covering multiple domains of expertise. This process traditionally demands immense time and specialized knowledge from decision-makers. The advent of Large Language Models (LLM) and Retrieval Augmented Generation (RAG) approaches offer a transformative solution, enabling rapid, accurate cross-document information retrieval and synthesis. As the landscape of Natural Language Processing (NLP) and text generation continues to evolve, benchmarking becomes essential to evaluate and compare the performance of different RAG-based LLMs. In this paper, we present a comprehensive framework to generate a domain relevant RAG benchmark. Our framework is based on automatic question-answer generation with Human (domain experts)-AI (LLM) teaming. As a case study, we demonstrate the framework by introducing WeQA, a first-of-its-kind benchmark on the wind energy domain which comprises of multiple scientific documents/reports related to environmental aspects of wind energy projects. Our framework systematically evaluates RAG performance using diverse metrics and multiple question types with varying complexity level, providing a foundation for rigorous assessment of RAG-based systems in complex scientific domains and enabling researchers to identify areas for improvement in domain-specific applications.

cs.CL

SAM-I-Am: Semantic Boosting for Zero-shot Atomic-Scale Electron Micrograph Segmentation

Image segmentation is a critical enabler for tasks ranging from medical diagnostics to autonomous driving. However, the correct segmentation semantics - where are boundaries located? what segments are logically similar? - change depending on the domain, such that state-of-the-art foundation models can generate meaningless and incorrect results. Moreover, in certain domains, fine-tuning and retraining techniques are infeasible: obtaining labels is costly and time-consuming; domain images (micrographs) can be exponentially diverse; and data sharing (for third-party retraining) is restricted. To enable rapid adaptation of the best segmentation technology, we propose the concept of semantic boosting: given a zero-shot foundation model, guide its segmentation and adjust results to match domain expectations. We apply semantic boosting to the Segment Anything Model (SAM) to obtain microstructure segmentation for transmission electron microscopy. Our booster, SAM-I-Am, extracts geometric and textural features of various intermediate masks to perform mask removal and mask merging operations. We demonstrate a zero-shot performance increase of (absolute) +21.35%, +12.6%, +5.27% in mean IoU, and a -9.91%, -18.42%, -4.06% drop in mean false positive masks across images of three difficulty classes over vanilla SAM (ViT-L).

cond-mat.mtrl-sci

Evaluating Physically Motivated Loss Functions for Photometric Redshift Estimation

Physical constraints have been suggested to make neural network models more generalizable, act scientifically plausible, and be more data-efficient over unconstrained baselines. In this report, we present preliminary work on evaluating the effects of adding soft physical constraints to computer vision neural networks trained to estimate the conditional density of redshift on input galaxy images for the Sloan Digital Sky Survey. We introduce physically motivated soft constraint terms that are not implemented with differential or integral operators. We frame this work as a simple ablation study where the effect of including soft physical constraints is compared to an unconstrained baseline. We compare networks using standard point estimate metrics for photometric redshift estimation, as well as metrics to evaluate how faithful our conditional density estimate represents the probability over the ensemble of our test dataset. We find no evidence that the implemented soft physical constraints are more effective regularizers than augmentation.

astro-ph.IM

The International Linear Collider: Report to Snowmass 2021

The International Linear Collider (ILC) is on the table now as a new global energy-frontier accelerator laboratory taking data in the 2030s. The ILC addresses key questions for our current understanding of particle physics. It is based on a proven accelerator technology. Its experiments will challenge the Standard Model of particle physics and will provide a new window to look beyond it. This document brings the story of the ILC up to date, emphasizing its strong physics motivation, its readiness for construction, and the opportunity it presents to the US and the global particle physics community.

physics.acc-ph

Strategy for Understanding the Higgs Physics: The Cool Copper Collider

A program to build a lepton-collider Higgs factory, to precisely measure the couplings of the Higgs boson to other particles, followed by a higher energy run to establish the Higgs self-coupling and expand the new physics reach, is widely recognized as a primary focus of modern particle physics. We propose a strategy that focuses on a new technology and preliminary estimates suggest that can lead to a compact, affordable machine. New technology investigations will provide much needed enthusiasm for our field, resulting in trained workforce. This cost-effective, compact design, with technologies useful for a broad range of other accelerator applications, could be realized as a project in the US. Its technology innovations, both in the accelerator and the detector, will offer unique and exciting opportunities to young scientists. Moreover, cost effective compact designs, broadly applicable to other fields of research, are more likely to obtain financial support from our funding agencies.

hep-ex

Strange quark as a probe for new physics in the Higgs sector

This paper describes a novel algorithm for tagging jets originating from the hadronisation of strange quarks (strange-tagging) with the future International Large Detector (ILD) at the International Linear Collider (ILC). It also presents the first application of such a strange-tagger to a Higgs to strange ($h \rightarrow s\bar{s}$) analysis with the $P(e^-,e^+) = (-80\%,+30\%)$ polarisation scenario, corresponding to 900 fb$^{-1}$ of the initial proposed 2000 fb$^{-1}$ of data which will be collected by ILD during its first 10 years of data taking at $\sqrt{s} = 250$ GeV. Upper limits on the Standard Model Higgs-strange coupling strength modifier, $\kappa_s$, are derived at the 95% confidence level to be 7.14. The paper includes as well a preliminary study of a Ring Imaging Cherenkov (RICH) system capable of discriminating between kaons and pions at high momenta (up to 25 GeV), and thus enhancing strange-tagging performance at future Higgs factory detectors.

hep-ex

Neural Ordinary Differential Equations for Nonlinear System Identification

Neural ordinary differential equations (NODE) have been recently proposed as a promising approach for nonlinear system identification tasks. In this work, we systematically compare their predictive performance with current state-of-the-art nonlinear and classical linear methods. In particular, we present a quantitative study comparing NODE's performance against neural state-space models and classical linear system identification methods. We evaluate the inference speed and prediction performance of each method on open-loop errors across eight different dynamical systems. The experiments show that NODEs can consistently improve the prediction accuracy by an order of magnitude compared to benchmark methods. Besides improved accuracy, we also observed that NODEs are less sensitive to hyperparameters compared to neural state-space models. On the other hand, these performance gains come with a slight increase of computation at the inference time.

cs.LG

Accelerated Computation of a High Dimensional Kolmogorov-Smirnov Distance

Statistical testing is widespread and critical for a variety of scientific disciplines. The advent of machine learning and the increase of computing power has increased the interest in the analysis and statistical testing of multidimensional data. We extend the powerful Kolmogorov-Smirnov two sample test to a high dimensional form in a similar manner to Fasano (Fasano, 1987). We call our result the d-dimensional Kolmogorov-Smirnov test (ddKS) and provide three novel contributions therewith: we develop an analytical equation for the significance of a given ddKS score, we provide an algorithm for computation of ddKS on modern computing hardware that is of constant time complexity for small sample sizes and dimensions, and we provide two approximate calculations of ddKS: one that reduces the time complexity to linear at larger sample sizes, and another that reduces the time complexity to linear with increasing dimension. We perform power analysis of ddKS and its approximations on a corpus of datasets and compare to other common high dimensional two sample tests and distances: Hotelling's T^2 test and Kullback-Leibler divergence. Our ddKS test performs well for all datasets, dimensions, and sizes tested, whereas the other tests and distances fail to reject the null hypothesis on at least one dataset. We therefore conclude that ddKS is a powerful multidimensional two sample test for general use, and can be calculated in a fast and efficient manner using our parallel or approximate methods. Open source implementations of all methods described in this work are located at https://github.com/pnnl/ddks.

stat.CO

ILC Study Questions for Snowmass 2021

To aid contributions to the Snowmass 2021 US Community Study on physics at the International Linear Collider and other proposed $e^+e^-$ colliders, we present a list of study questions that could be the basis of useful Snowmass projects. We accompany this with links to references and resources on $e^+e^-$ physics, and a description of a new software framework that we are preparing for $e^+e^-$ studies at Snowmass.

hep-ph

Scaling the training of particle classification on simulated MicroBooNE events to multiple GPUs

Measurements in Liquid Argon Time Projection Chamber (LArTPC) neutrino detectors, such as the MicroBooNE detector at Fermilab, feature large, high fidelity event images. Deep learning techniques have been extremely successful in classification tasks of photographs, but their application to LArTPC event images is challenging, due to the large size of the events. Events in these detectors are typically two orders of magnitude larger than images found in classical challenges, like recognition of handwritten digits contained in the MNIST database or object recognition in the ImageNet database. Ideally, training would occur on many instances of the entire event data, instead of many instances of cropped regions of interest from the event data. However, such efforts lead to extremely long training cycles, which slow down the exploration of new network architectures and hyperparameter scans to improve the classification performance. We present studies of scaling a LArTPC classification problem on multiple architectures, spanning multiple nodes. The studies are carried out on simulated events in the MicroBooNE detector. We emphasize that it is beyond the scope of this study to optimize networks or extract the physics from any results here. Institutional computing at Pacific Northwest National Laboratory and the SummitDev machine at Oak Ridge National Laboratory's Leadership Computing Facility have been used. To our knowledge, this is the first use of state-of-the-art Convolutional Neural Networks for particle physics and their attendant compute techniques onto the DOE Leadership Class Facilities. We expect benefits to accrue particularly to the Deep Underground Neutrino Experiment (DUNE) LArTPC program, the flagship US High Energy Physics (HEP) program for the coming decades.

physics.comp-ph

Performance of Julia for High Energy Physics Analyses

We argue that the Julia programming language is a compelling alternative to implementations in Python and C++ for common data analysis workflows in high energy physics. We compare the speed of implementations of different workflows in Julia with those in Python and C++. Our studies show that the Julia implementations are competitive for tasks that are dominated by computational load rather than data access. For work that is dominated by data access, we demonstrate an application with concurrent file reading and parallel data processing.

physics.comp-ph

The International Linear Collider: A Global Project

The International Linear Collider (ILC) is now under consideration as the next global project in particle physics. In this report, we review of all aspects of the ILC program: the physics motivation, the accelerator design, the run plan, the proposed detectors, the experimental measurements on the Higgs boson, the top quark, the couplings of the W and Z bosons, and searches for new particles. We review the important role that polarized beams play in the ILC program. The first stage of the ILC is planned to be a Higgs factory at 250 GeV in the centre of mass. Energy upgrades can naturally be implemented based on the concept of a linear collider. We discuss in detail the ILC program of Higgs boson measurements and the expected precision in the determination of Higgs couplings. We compare the ILC capabilities to those of the HL-LHC and to those of other proposed e+e- Higgs factories. We emphasize throughout that the readiness of the accelerator and the estimates of ILC performance are based on detailed simulations backed by extensive RandD and, for the accelerator technology, operational experience.

hep-ex

The International Linear Collider. A Global Project

A large, world-wide community of physicists is working to realise an exceptional physics program of energy-frontier, electron-positron collisions with the International Linear Collider (ILC). This program will begin with a central focus on high-precision and model-independent measurements of the Higgs boson couplings. This method of searching for new physics beyond the Standard Model is orthogonal to and complements the LHC physics program. The ILC at 250 GeV will also search for direct new physics in exotic Higgs decays and in pair-production of weakly interacting particles. Polarised electron and positron beams add unique opportunities to the physics reach. The ILC can be upgraded to higher energy, enabling precision studies of the top quark and measurement of the top Yukawa coupling and the Higgs self-coupling. The key accelerator technology, superconducting radio-frequency cavities, has matured. Optimised collider and detector designs, and associated physics analyses, were presented in the ILC Technical Design Report, signed by 2400 scientists. There is a strong interest in Japan to host this international effort. A detailed review of the many aspects of the project is nearing a conclusion in Japan. Now the Japanese government is preparing for a decision on the next phase of international negotiations, that could lead to a project start within a few years. The potential timeline of the ILC project includes an initial phase of about 4 years to obtain international agreements, complete engineering design and prepare construction, and form the requisite international collaboration, followed by a construction phase of 9 years.

hep-ex

Precision Higgs Measurements at the 250 GeV ILC

The plan for the International Linear Collider is now being prepared as a staged design, with the first stage at 250 GeV and later stages achieving the full project specifications with 4 ab-1 at 500 GeV. This talk will present the capabilities for precision Higgs boson measurements at 250 GeV and their relation to the full ILC program. It will show that the 250 GeV stage of ILC will already provide many compelling results in Higgs physics, with new measurements not available at LHC, model-independent determinations of key parameters, and tests for and possible discrimination of a variety of scenarios for new physics.

hep-ex