NSF PAR Search | NSF Public Access Repository

Note: When clicking on a Digital Object Identifier (DOI) number, you will be taken to an external site maintained by the publisher. Some full text articles may not yet be available without a charge during the embargo (administrative interval).
What is a DOI Number?

Some links on this page may take you to non-federal websites. Their policies may differ from this site.

Computing at the HL-LHC and beyond

https://doi.org/10.22323/1.478.0333

Gardner, Robert William (December 2024, Sissa Medialab)

The High-Luminosity Large Hadron Collider (HL-LHC) is set to introduce unprecedented data volumes and computational demands, necessitating significant enhancements in the current LHC computing infrastructure. We summarize efforts by the experiments to integrate high-performance computing clusters and public cloud resources into their processing frameworks. We also examine the adoption of cloud technologies for implementation of advanced service infrastructure which are finding applications in Tier 2 centers and prototyping of future analysis facilities. We highlight the crucial role of scalable networking capabilities and challenge exercises to prepare for the expected increased data throughput.
more » « less
Full Text Available
Microalgae, soil and plants: A critical review of microalgae as renewable resources for agriculture

https://doi.org/10.1016/j.algal.2021.102200

Alvarez, Adriana L.; Weyers, Sharon L.; Goemann, Hannah M.; Peyton, Brent M.; Gardner, Robert D. (April 2021, Algal Research)
null (Ed.)
Full Text Available
Towards a NoOps Model for WLCG

https://doi.org/https://doi.org/10.1051/epjconf/202024507024

Gardner, Robert; Bryant, Lincoln; Stephen, Judith; Vukotic, Ilija; Weaver, Christopher; Wu, Wenjing (November 2020, 24th International Conference on Computing in High Energy and Nuclear Physics (CHEP 2019))
null (Ed.)
One of the most costly factors in providing a global computing infrastructure such as the WLCG is the human effort in deployment, integration, and operation of the distributed services supporting collaborative computing, data sharing and delivery, and analysis of extreme scale datasets. Furthermore, the time required to roll out global software updates, introduce new service components, or prototype novel systems requiring coordinated deployments across multiple facilities is often increased by communication latencies, staff availability, and in many cases expertise required for operations of bespoke services. While the WLCG (and distributed systems implemented throughout HEP) is a global service platform, it lacks the capability and flexibility of a modern platform-as-a-service including continuous integration/continuous delivery (CI/CD) methods, development-operations capabilities (DevOps, where developers assume a more direct role in the actual production infrastructure), and automation. Most importantly, tooling which reduces required training, bespoke service expertise, and the operational effort throughout the infrastructure, most notably at the resource endpoints (sites), is entirely absent in the current model. In this paper, we explore ideas and questions around potential NoOps models in this context: what is realistic given organizational policies and constraints? How should operational responsibility be organized across teams and facilities? What are the technical gaps? What are the social and cybersecurity challenges? Conversely what advantages does a NoOps model deliver for innovation and for accelerating the pace of delivery of new services needed for the HL-LHC era? We will describe initial work along these lines in the context of providing a data delivery network supporting IRIS-HEP DOMA R&D.
more » « less
Full Text Available
The Scalable Systems Laboratory: a Platform for Software Innovation for HEP

https://doi.org/10.1051/epjconf/202024505019

Gardner, Robert; Bryant, Lincoln; Neubauer, Mark; Wuerthwein, Frank; Stephen, Judith; Chien, Andrew (January 2020, EPJ Web of Conferences)
Doglioni, C.; Kim, D.; Stewart, G.A.; Silvestris, L.; Jackson, P.; Kamleh, W. (Ed.)
The Scalable Systems Laboratory (SSL), part of the IRIS-HEP Software Institute, provides Institute participants and HEP software developers generally with a means to transition their R&D from conceptual toys to testbeds to production-scale prototypes. The SSL enables tooling, infrastructure, and services supporting innovation of novel analysis and data architectures, development of software elements and tool-chains, reproducible functional and scalability testing of service components, and foundational systems R&D for accelerated services developed by the Institute. The SSL is constructed with a core team having expertise in scale testing and deployment of services across a wide range of cyberinfrastructure. The core team embeds and partners with other areas in the Institute, and with LHC and other HEP development and operations teams as appropriate, to define investigations and required service deployment patterns. We describe the approach and experiences with early application deployments, including analysis platforms and intelligent data delivery systems.
more » « less
Full Text Available
Distributed Computing Software and Data Access Patterns in OSG Midscale Collaborations

https://doi.org/10.1051/epjconf/202024503005

Paschos, Pascal; Riedel, Benedikt; Rynge, Mats; Bryant, Lincoln; Stephen, Judith; Gardner, Robert; Fajardo, Edgar; Hicks, John; Wuerthwein, Frank; Clark, James (January 2020, EPJ Web of Conferences)
Doglioni, C.; Kim, D.; Stewart, G.A.; Silvestris, L.; Jackson, P.; Kamleh, W. (Ed.)
In this paper we showcase the support in Open Science Grid (OSG) of Midscale collaborations, the region of computing and storage scale where multi-institutional researchers collaborate to execute their science workflows on the grid without having dedicated technical support teams of their own. Collaboration Services enables such collaborations to take advantage of the distributed resources of the Open Science Grid by facilitating access to submission hosts, the deployment of their applications and supporting their data management requirements. Distributed computing software adopted from large scale collaborations, such as CVMFS, Rucio, xCache lower the barrier of intermediate scale research to integrate with existing infrastructure.
more » « less
Full Text Available
WLCG Networks: Update on Monitoring and Analytics

https://doi.org/10.1051/epjconf/202024507053

Babik, Marian; McKee, Shawn; Andrade, Pedro; Bockelman, Brian Paul; Gardner, Robert; Fajardo Hernandez, Edgar Mauricio; Martelli, Edoardo; Vukotic, Ilija; Weitzel, Derek; Zvada, Marian (January 2020, EPJ Web of Conferences)
Doglioni, C.; Kim, D.; Stewart, G.A.; Silvestris, L.; Jackson, P.; Kamleh, W. (Ed.)
WLCG relies on the network as a critical part of its infrastructure and therefore needs to guarantee effective network usage and prompt detection and resolution of any network issues including connection failures, congestion and traffic routing. The OSG Networking Area, in partnership with WLCG, is focused on being the primary source of networking information for its partners and constituents. It was established to ensure sites and experiments can better understand and fix networking issues, while providing an analytics platform that aggregates network monitoring data with higher level workload and data transfer services. This has been facilitated by the global network of the perfSONAR instances that have been commissioned and are operated in collaboration with WLCG Network Throughput Working Group. An additional important update is the inclusion of the newly funded NSF project SAND (Service Analytics and Network Diagnosis) which is focusing on network analytics. This paper describes the current state of the network measurement and analytics platform and summarises the activities taken by the working group and our collaborators. This includes the progress being made in providing higher level analytics, alerting and alarming from the rich set of network metrics we are gathering.
more » « less
Full Text Available
Buzzard: Georgia Tech’s Foray into the Open Science Grid

https://doi.org/10.1145/3491418.3535135

Belgin, Mehmet; Sarajlic, Semir; Lara, Ruben; Cadonati, Laura; Otte, A. Nepomuk; Taboada, Ignacio J.; Beyer, Gregory L; Bonner, Norman B; Brandon, Michael; Buffington, Pam; et al (July 2022, PEARC '22: Practice and Experience in Advanced Research Computing)

Full Text Available
Managing Privilege and Access on Federated Edge Platforms

https://doi.org/10.1145/3332186.3332234

Breen, Joe; Bryant, Lincoln; Chen, Jiahui; Ford, Emerson; Gardner, Robert W.; Glupker, Gage; Griffith, Skyler; Kulbertis, Ben; McKee, Shawn; Pierce, Rose; et al (January 2019, Proceedings of the Practice and Experience in Advanced Research Computing on Rise of the Machines (learning))

Full Text Available
Building the SLATE Platform

https://doi.org/10.1145/3219104.3219144

Breen, Joe; McKee, Shawn; Riedel, Benedikt; Stidd, Jason; Truong, Luan; Vukotic, Ilija; Bryant, Lincoln; Carcassi, Gabriele; Chen, Jiahui; Gardner, Robert W.; et al (July 2018, Proceedings of the Practice and Experience on Advanced Research Computing)

We describe progress on building the SLATE (Services Layer at the Edge) platform. The high level goal of SLATE is to facilitate creation of multi-institutional science computing systems by augmenting the canonical Science DMZ pattern with a generic, "programmable", secure and trusted underlayment platform. This platform permits hosting of advanced container-centric services needed for higher-level capabilities such as data transfer nodes, software and data caches, workflow services and science gateway components. SLATE uses best-of-breed data center virtualization and containerization components, and where available, software defined networking, to enable distributed automation of deployment and service lifecycle management tasks by domain experts. As such it will simplify creation of scalable platforms that connect research teams, institutions and resources to accelerate science while reducing operational costs and development cycle times.
more » « less
Full Text Available
Developing Edge Services for Federated Infrastructure Using MiniSLATE

https://doi.org/10.1145/3332186.3332236

Breen, Joe; Bryant, Lincoln; Chen, Jiahui; Ford, Emerson; Gardner, Robert W.; Glupker, Gage; Griffith, Skyler; Kulbertis, Ben; McKee, Shawn; Pierce, Rose; et al (January 2019, Proceedings of the Practice and Experience in Advanced Research Computing on Rise of the Machines (Learning))

Full Text Available

« Prev Next »

Search for: All records