Search arXiv⌕ Search

arXiv · 2609.39546

FAIR-Compliant Architecture for Heterogeneous Astronomical Data: KazVO Framework

Abstract

This paper presents the architecture of the Kazakhstani National Virtual Observatory (KazVO) - an International Virtual Observatory Alliance (IVOA) compliant node that unifies the heterogeneous observational datasets of the Fesenkov Astrophysical Institute (FAI). The architectural foundation of the system is built upon a partitioned Data Lake, a German Astrophysical Virtual Observatory (GAVO) Data Center Helper Suite (DaCHS) publishing backend, and a PostgreSQL database that maps heterogeneous metadata to the unified IVOA ObsCore standard. We follow Open Science by introducing metadata-only model featuring based on a four-class data embargo mechanism. Machine-to-machine services deployed via IVOA protocols are listed in the global Registry of Registers, enabling analysis of Kazakhstani observational assets within external clients such as Tool for OPerations on Catalogues And Tables (TOPCAT), Aladin, and PyVO. As a result, KazVO framework provides a scalable platform driven by a developed end-to-end pipeline that bridges two fundamentally distinct data types - digitized historical glass-plate heritage (1950-1997) and live operational photometric and spectroscopic digital streams from telescopes at the Assy-Turgen and Tien-Shan Observatories - opening FAI's combined data datasets to the global scientific and time-domain astrophysics community.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Ildana Izmailova, Denis Yurin, Yerlan Aimuratov, Maxim Makukov, Aleksander Serebryanskiy, Saule Shomshekova, Vitaliy Kim, Chingis Omarov, Adel Umirbayeva, Laura Aktay, Dana Kuvatova, Anton Gluchshenko, Gulnara Suliyeva, Maxim Krugov, Nadezhda Vaidman, Daulet Anarbek, Inna Reva, Gauhar Aimanova, Rashit Valiullin, Raushan Kokumbayeva. 2026-09-30. FAIR-Compliant Architecture for Heterogeneous Astronomical Data: KazVO Framework. https://arxiv.org/abs/2609.39546

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

LeoNet: A Machine Learning Method for Binary Pulsar Classification

Binary pulsars provide valuable laboratories for testing theories of gravity, but orbital Doppler shifts complicate their detection. Fourier-domain acceleration and jerk searches address this challenge via matched filtering, but at substantial computational cost. We present LeoNet, a convolutional neural network that uses ten learnable filters to extract features of signals affected by Doppler shifts. The resulting ten-channel feature map provides a compact, lower-dimensional alternative to an explicitly sampled acceleration-jerk response grid and is analysed by a convolutional classifier to identify candidate signals. For simulated observations lasting 500 s, LeoNet achieves a mean relative reduction in false negative rate of 55.5% across five sampling intervals compared with the evaluated PRESTO acceleration-search configuration. TensorRT-optimised LeoNet processes each 500 s observation in 3.44-4.37 ms in FP32 on an NVIDIA H100 PCIe GPU across eight sampling intervals, including preprocessing, inference, and postprocessing. At a sampling interval of 128 microseconds, its mean processing time is 3.54 ms, compared with 1.767 s for PRESTO FDAS on an AMD EPYC 9825 CPU with search-frequency limits of 96-1000 Hz, corresponding to an approximately 499-fold speedup in the measured processing time. These results suggest that LeoNet has the potential to improve detection performance, while its millisecond-scale processing time supports its use as a candidate-identification stage in real-time binary pulsar search pipelines.

astro-ph.IM↗

Comparing Optimized Systematic Error Correction Methods on Selected TESS Light Curves

The correction of systematic errors in TESS light curves is crucial for all astrophysical analyses employing these observations. Here we present a data analysis and a software package, SysCoCoPy, to investigate and directly compare the performance of the Presearch Data Conditioning (PDC) correcting method from the Science Processing Operations Center (SPOC) pipeline and three correctors developed by the community. We incorporate these three correctors based on their implementations in Lightkurve and are particularly interested in the ability of the four correctors to remove scattered light contamination from the Earth and the Moon, which is a key systematic for TESS. We implemented these correctors in SysCoCoPy with a framework that allows an automatic optimization of their parameters to scale the analysis towards increasingly larger samples. SysCoCoPy provides qualitative and quantitative products for the comparison of individual cases as well as statistical results for selected samples. We currently find that while our automatic parameter optimization provides a significant number of successful scattered-light corrections for two of the correctors with a design that favors this purpose, an statistical analysis of our largest sample indicates that PDC is presently more robust, with a larger overall success level of the metrics used.

astro-ph.IM↗

hyprfine: simulating the 21-cm signal from the Dark Ages through to the Epoch of Reionization on a GPU

hyprfine is an analytic simulation of the sky-averaged 21-cm signal from $z=1100 - 6$ written using JAX and Python for native GPU capabilities. It models the average temperature of the 21-cm signal over cosmic time as a function of the $Λ$-CDM cosmology parameters and the astrophysics of the first stars and galaxies. As far as we are aware, the code is the first analytic GPU native simulation of the 21-cm signal. It runs in a fraction of a second, parallelises efficiently across a GPU and is differentiable through the Dark Ages ($z \geq 35$).

astro-ph.IM↗