Speaker
Description
The HSF Conditions Database (CDB) is a community-driven solution for managing conditions data - non-event data required for event processing - which present common challenges across HENP and astro-particle experiments. In the three years of production operation for sPHENIX at BNL, where the HSF CDB supports over 70,000 concurrent jobs on a farm running 132,000 logical cores, it has evolved into a robust and scalable service shaped by continuous real-world feedback. It is also being adopted by the Belle II experiment at KEK, which runs up to about 36,000 concurrent jobs across a distributed grid of about 38 computing sites worldwide. We will present the insights from the migration, including the challenges of data migration, as well as the developments on both the server and client software sides required to adopt and integrate the new system.
These experiences highlight practical considerations for experiment-wide deployment and provide guidance for future adopters of the HSF CDB. Recent developments further enhance the system, including caching techniques, experiment-specific authentication plugins, database replication for high availability, and a deployment framework based on Helm charts and OpenShift. An IntelligentLogging pipeline, developed through the HSF GSoC programme, provides central log aggregation, storage, and monitoring, and uses DeepLog-based anomaly detection. Beyond sPHENIX and Belle II, DUNE is preparing for adoption after initial prototype deployments demonstrated sufficient performance and Einstein Telescope is integrating it into an Open Science solution, demonstrating the system’s versatility across diverse experimental environments.