Batch is back: CasJobs, serving multi-TB data on the Web

التفاصيل البيبلوغرافية
العنوان: Batch is back: CasJobs, serving multi-TB data on the Web
المؤلفون: Wil O'Mullane, Aniruddha R. Thakar, Alexander S. Szalay, Maria Nieto-Santisteban, Nolan Li, Jim Gray
المصدر: ICWS
بيانات النشر: IEEE, 2005.
سنة النشر: 2005
مصطلحات موضوعية: Database, Interface (Java), SOAP, computer.internet_protocol, Computer science, Joins, computer.software_genre, World Wide Web, Upload, Transfer (computing), Batch processing, Web service, computer, Server-side
الوصف: The Sloan Digital Sky Survey (SDSS) science database describes over 230 million objects and is over 1.6 TB in size. The SDSS Catalog Archive Server (CAS) provides several levels of query interface to the SDSS data via the SkyServer website. Most queries execute in seconds or minutes. However, some queries can take hours or days, either because they require non-index scans of the largest tables, or because they request very large result sets, or because they represent very complex aggregations of the data. These "monster queries" not only take a long time, they also affect response times for everyone else - one or more of them can clog the entire system. To ameliorate this problem, we developed a multiserver multiqueue batch job submission, execution, and tracking system for the CAS called CasJobs. The transfer of very large result sets from queries over the network is another serious problem. Statistics suggested that much of this data transfer is unnecessary; users would prefer to store results locally in order to allow further joins and filtering. To allow local analysis, a system was developed that gives users their own personal databases (MyDB) at the server side. Users may transfer data to their MyDB, and then perform further analysis before extracting it to their own machine. MyDB tables also provide a convenient way to share results of queries with collaborators without downloading them. CasJobs is built using SOAP XML Web services and has been in operation since May 2004.
URL الوصول: https://explore.openaire.eu/search/publication?articleId=doi_________::d53b09c6599288a7f77a9d345143dd47
https://doi.org/10.1109/icws.2005.29
رقم الأكسشن: edsair.doi...........d53b09c6599288a7f77a9d345143dd47
قاعدة البيانات: OpenAIRE