Skip to content
Open access

Unsupervised sifting of XMM-Newton EPIC observations using Variational Autoencoders

Jul 2026 · RAS Techniques and Instruments · Vol 5 · 0 citations

TL;DR

A novel unsupervised framework based on Variational Autoencoders (VAEs) to automatically sift through the XSA, which provides a scalable screening tool for multi-band light curves generated by the EPIC-pn Pipeline Processing System (PPS).

Abstract

Finding specific events of interest within the XMM-Newton Science Archive (XSA) is a significant challenge due to the sheer volume of data, which contains over 300,000 source light curves. While the mission provides standardized multi-band products, the vast majority of these time series remain unexplored, and it is often unknown a priori whether they contain intrinsic variability, transients, or flares. Traditional screening is further complicated by non-stationary instrumental backgrounds that often overlap with physical signals. This paper presents a novel unsupervised framework based on Variational Autoencoders (VAEs) to automatically sift through the XSA. Unlike previous methods that rely on hand-crafted statistical features or multi-epoch historical data, our approach operates directly on (pipeline-processed) contemporaneous multi-band light curves. By learning patterns of variability shared across the five contemporaneous energy-band light curves, the VAE maps high-dimensional observations into a low-dimensional latent space. This allows for the discovery of complex variability patterns and the identification of anomalies based on their intrinsic morphological structure. Ultimately, the proposed framework provides a scalable screening tool for multi-band light curves generated by the EPIC-pn Pipeline Processing System (PPS), ranking candidate anomalies using model-based scores, significantly reducing the manual burden of exploratory searches, and prioritizing potentially interesting sources for follow-up analysis. The validity of the method has been assessed both qualitatively and quantitatively.

Read PDF

Similar papers

Review Open access Sep 2026

Flow-variational auto-encoder for the data augmentation of the Horizon-AGN dual-AGN dataset

A variational auto-encoder (VAE) with a normalising flow is constructed, developing a specific training loss function for the VAE in both the direct and Fourier spaces, which has enabled the construction of the small-scale and thin features associated with the gravitational interaction between merging galaxies like tid...

C. Mahé, H. Bretonnière, E. Slezak · 0 citations
Preprint Aug 2026

A Machine Learning Based Search for Lunar Anomalies

This investigation further gauged the Beta-Variational Autoencoder's ability to locate anomalous surface features, successfully recovering two places of interest and numerous landed technological assets at a statistically significant rate.

Cameron Kelahan, D. Angerhausen, Adam Lesnikowski et al. · 0 citations
Review Open access Sep 2026

Archival Transients Detected with the MeerLICHT Telescope. I. Supernovae

Wide-field optical transient surveys collectively discover ∼50 new transients per night, making rapid, multi-filter and/or spectroscopic follow-up of all but the most exceptional events unfeasible. Consequently, the (early) light-curve properties of most supernovae (SNe), particularly sub-luminous events, are often poo...

O. Mogawana, P. Groot, N. Blagorodnova et al. · 0 citations
Review Open access Aug 2026

The promise of self-supervised and active learning for Strong Lens discovery: Astronomaly applied to KiDS

Strong gravitational lenses (SGLs) are rare systems whose discovery currently relies primarily on supervised machine learning methods trained on large simulated datasets. We present the first application of Astronomaly:PROTEGE to SGL discovery, demonstrating that a human-in-the-loop active learning framework can effi...

M. Grespan, Aprajita Verma, Michelle Lochner et al. · 0 citations
Review Sep 2026

Learning JWST. I. A Foundation Model for New Population Discoveries and Morphology-Aware Photometric Redshift Measurements in the JADES Survey

We present FM-JADES-v1, a self-supervised foundation model for James Webb Space Telescope ({\em JWST}) deep-field science, trained with 482,444 objects from the {\em JWST} Advanced Deep Extragalactic Survey (JADES) Data Release 5 using multi-band imaging and the photometric catalog. The shared embedding space is traine...

Jia-Ni Ding, M. Yue, Yong-Da Zhu et al. · 1 citation

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.