2003/06/13 by Tofigh Azemoon, Azemoon, Tofigh, Adil Hasan +5
Computer Science · #Databases (cs.DB) #Distributed #E.5 #FOS: Computer and information sciences #H.2.7 #J.2 #Parallel #and Cluster Computing (cs.DC) #cs.DB #cs.DC
paper · pdf · doi:10.48550/arxiv.cs/0306061
Presented for Computing in High Energy Physics, San Diego, March 2003
arxiv created 2003/06/13 · arxiv updated 2009/11/30
To date, the BaBar experiment has stored over 0.7PB of data in an Objectivity/DB database. Approximately half this data-set comprises simulated data of which more than 70% has been produced at more than 20 collaborating institutes outside of SLAC. The operational aspects of managing such a large data set and providing access to the physicists in a timely manner is a challenging and complex problem. We describe the operational aspects of managing such a large distributed data-set as well as importing and exporting data from geographically spread BaBar collaborators. We also describe problems common to dealing with such large datasets.