Search arXiv⌕ Search

arXiv · 2506.15585

Engineering Supercomputing Platforms for Biomolecular Applications

Abstract

A range of computational biology software (GROMACS, AMBER, NAMD, LAMMPS, OpenMM, Psi4 and RELION) was benchmarked on a representative selection of HPC hardware, including AMD EPYC 7742 CPU nodes, NVIDIA V100 and AMD MI250X GPU nodes, and an NVIDIA GH200 testbed. The raw performance, power efficiency and data storage requirements of the software was evaluated for each HPC facility, along with qualitative factors such as the user experience and software environment. It was found that the diversity of methods used within computational biology means that there is no single HPC hardware that can optimally run every type of HPC job, and that diverse hardware is the only way to properly support all methods. New hardware, such as AMD GPUs and Nvidia AI chips, are mostly compatible with existing methods, but are also more labour-intensive to support. GPUs offer the most efficient way to run most computational biology tasks, though some tasks still require CPUs. A fast HPC node running molecular dynamics can produce around 10GB of data per day, however, most facilities and research institutions lack short-term and long-term means to store this data. Finally, as the HPC landscape has become more complex, deploying software and keeping HPC systems online has become more difficult. This situation could be improved through hiring/training in DevOps practices, expanding the consortium model to provide greater support to HPC system administrators, and implementing build frameworks/containerisation/virtualisation tools to allow users to configure their own software environment, rather than relying on centralised software installations.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Robert Welch, Charles Laughton, Oliver Henrich, Tom Burnley, Daniel Cole, Alan Real, Sarah Harris, James Gebbie-Rayet. 2025-10-13. Engineering Supercomputing Platforms for Biomolecular Applications. https://arxiv.org/abs/2506.15585

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Rapid prediction of organisation in engineered corneal, glial and fibroblast tissues using machine learning and biophysical models

We present a machine learning approach for predicting the organisation of corneal, glial and fibroblast cells in 3D cultures used for tissue engineering. Our machine-learning-based method uses a generative adversarial network architecture called pix2pix, which we train using results from biophysical contractile network dipole orientation (CONDOR) simulations. In the following, we refer to the machine learning method as the RAPTOR (RApid Prediction of Tissue ORganisation) approach. A training data set containing a range of CONDOR simulations is created, covering a range of underlying model parameters. Predictions of the trained neural network are compared with cultured glial, corneal, and fibroblast tissues, with good agreements for both CONDOR and RAPTOR approaches. An approach is developed to determine CONDOR model parameters for specific tissues using both RAPTOR and CONDOR fits to tissue properties. RAPTOR outputs a variety of tissue properties, including cell densities, cell alignments and tension. RAPTOR yields predictions of tissue properties within fractions of a second. This speed makes it valuable for the design of tethered moulds for tissue growth.

physics.bio-ph↗

How do incorrect ligands help detect a correct ligand?

Intrigued by the response of T cell receptors to the presence of a few agonist ligands, we propose a minimal model that can achieve similar performance. The model consists of a small cluster of immobile receptors that bind reversibly to two types (correct/incorrect) of ligands in the environment, with slightly weaker binding strength for the incorrect one. It features binding-state coupling between nearest-neighbor receptors, and receptors in the bound/free states are activated/deactivated by specific enzymes, with rates that allow kinetic proofreading. It is found that, for a range of binding-state coupling strength, incorrect ligands alone cannot activate the receptors, but the binding of merely one correct ligand to a receptor is sufficient to promote the activation of other receptors via induced binding to incorrect ligands. Both response time and signal amplification increase as the receptor binding-state coupling strength increases until it reaches an optimal range to achieve the most rapid and sensitive response. These results suggest a possible mechanism for a speedy and specific response of receptors to very few correct ligands in biological and artificial systems at the subcellular scale.

physics.bio-ph↗

Spatially Resolved Nucleated Polymerization: A Free-Boundary Model of Protein Aggregation in Concentrated Solutions

Kinetic models of protein aggregation describe populations by size, not by spatial organization or morphology. We extend Lumry-Eyring nucleated polymerization to a model in which the monomer is a density field and each aggregate is a region bounded by a level set. Growth is a flux condition on the available sites of a surface. Condensation is a reaction between the bonding sites of two surfaces in contact, at a rate set by the bond rate and the contact geometry. The availability of those sites is a field on the interface, and its equilibrium value follows from Wertheim's perturbation theory. The collision efficiency and the Fuchs stability ratio are therefore computed, not fitted. In a well-mixed limit the model's spatial averages satisfy the rate equations term by term; the monomer fraction agrees to eight parts in ten thousand, a difference that arises from equating aggregate size with volume. The condensation kernel's exponent is $0.5806\pm0.0013$ against the $0.600\pm0.010$ fitted to a monoclonal antibody. The computed stability ratio reproduces thirteen of fourteen published conditions at twelve $k_BT$, but only with the bond rate at the top of its range. In a many-body box, aggregates merge at $2.2$ to $3.8$ times the two-body rate.

physics.bio-ph↗