ConferenceProceedings of the 19th Usenix Symposium on Networked Systems Design and Implementation Nsdi 2022 · January 1, 2022
A cloud provider today provides its network resources to its tenants as a black box, such that cloud tenants have little knowledge of the underlying network characteristics. Meanwhile, data-intensive applications have increasingly migrated to the cloud, an ...
Cite
ConferenceProceedings International Conference on Computer Communications and Networks ICCCN · July 1, 2021
This paper presents the rationale and design of the trust plane for ImPACT, a federated platform for managed sharing of restricted data. Key elements of the architecture include Web-based notaries for credential establishment based on declarative templates ...
Full textCite
ConferenceProceedings 2021 IEEE 35th International Parallel and Distributed Processing Symposium IPDPS 2021 · May 1, 2021
This work seeks to advance the state of the art in HPC I/O performance analysis and interpretation. In particular, we demonstrate effective techniques to: (1) model output performance in the presence of I/O interference from production loads; (2) build fea ...
Full textCite
ConferenceProceedings IEEE International Conference on Cluster Computing Iccc · January 1, 2021
This paper introduces WIRE that manages resources for the DAG-based workflows on IaaS clouds. WIRE predicts and plans resources over the MAPE (Monitor-AnalyzePlan-Execute) loops to: 1) Estimate task performance with online data, 2) Conduct simulations to p ...
Full textCite
ConferenceIEEE INFOCOM 2020 IEEE Conference on Computer Communications Workshops INFOCOM Wkshps 2020 · July 1, 2020
Research testbed fabrics have potential to support long-lived, evolving, interdomain experiments, including opt-in application traffic across multiple campuses and edge sites. We propose abstractions and security infrastructure to facilitate multi-domain n ...
Full textCite
ConferenceProceedings of Pdsw 2019 IEEE ACM 4th International Parallel Data Systems Workshop Held in Conjunction with Sc 2019 the International Conference for High Performance Computing Networking Storage and Analysis · November 1, 2019
In high-performance computing (HPC), I/O performance prediction offers the potential to improve the efficiency of scientific computing. In particular, accurate prediction can make runtime estimates more precise, guide users toward optimal checkpoint strate ...
Full textCite
Conference2018 International Scientific and Technical Conference Modern Computer Network Technologies Monetec 2018 Proceedings · December 10, 2018
In this paper we present a vision of an environment composed of multiple independent cloud providers of various sizes, interconnected by programmable networks in which tenants may acquire resources from the providers and interconnect them together to serve ...
Full textCite
ConferenceINFOCOM 2018 IEEE Conference on Computer Communications Workshops · July 6, 2018
A key dimension of reproducibility in testbeds is stable performance that scales in regular and predictable ways in accordance with declarative specifications for virtual resources. We contend that reproducibility is crucial for elastic performance control ...
Full textCite
Conference2017 IEEE Conference on Computer Communications Workshops INFOCOM Wkshps 2017 · November 20, 2017
The GENI network testbed was designed to enable experimentation with network protocols by offering the capability to construct virtual networks at the link layer (L2). GENI users build virtual networks in their GENI slices that span resources on multiple G ...
Full textCite
ConferenceHpdc 2017 Proceedings of the 26th International Symposium on High Performance Parallel and Distributed Computing · June 26, 2017
In this paper, we develop a predictive model useful for output performance prediction of supercomputer file systems under production load. Our target environment is Titan-the 3rd fastest supercomputer in the world-and its Lustre-based multi-stage write pat ...
Full textCite
ConferenceLecture Notes of the Institute for Computer Sciences Social Informatics and Telecommunications Engineering Lnicst · January 1, 2017
This paper describes advanced capabilities that were deployed recently in the ExoGENI testbed to offer increased flexibility in provisioning, modifying, and recovering the topologies and the configuration settings of the virtual systems, or slices, in whic ...
Full textCite
ConferenceLecture Notes in Computer Science Including Subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics · January 1, 2017
This paper reports our observations from a top-tier supercomputer Titan and its Lustre parallel file stores under production load. In summary, we find that supercomputer file systems are highly variable across the machine at fine time scales. This variabil ...
Full textCite
Conference9th Usenix Workshop on Hot Topics in Cloud Computing Hotcloud 2017 Co Located with Usenix Atc 2017 · January 1, 2017
One way to establish trust in a service is to know what code it is running. However, verified code identity is currently not possible for programs launched on a cloud by another party. We propose an approach to integrate support for code attestation—authen ...
Cite
ConferenceProceedings of the 7th ACM Symposium on Cloud Computing Socc 2016 · October 5, 2016
Cloud providers are in a position to greatly improve the trust clients have in network services: IaaS platforms can isolate services so they cannot leak data, and can help verify that they are securely deployed. We describe a new system called CQSTR that a ...
Full textCite
ConferenceProceedings 2015 IEEE ACM 8th International Conference on Utility and Cloud Computing Ucc 2015 · January 1, 2015
Recent advances in cloud technologies and on-demand network circuits have created an unprecedented opportunity to enable complex data-intensive scientific applications to run on dynamic, networked cloud infrastructure. However, there is a lack of tools for ...
Full textCite
ConferenceLecture Notes of the Institute for Computer Sciences Social Informatics and Telecommunications Engineering Lnicst · January 1, 2015
Researchers and educators in computer science and other domains are increasingly turning to distributed test beds that offer access to a variety of resources, including networking, computation, storage, sensing, and actuation. The provisioning of resources ...
Full textCite
ConferenceWorkshop on Hot Topics in Cloud Computing Hotcloud 2009 · January 1, 2009
This paper addresses “reflective” control for applications that use server resources from a shared cloud infrastructure opportunistically. In this approach, an external reflective controller launches application functions based on knowledge of what resourc ...
Cite
ConferenceProceedings of the 2008 Usenix Annual Technical Conference Usenix 2008 · January 1, 2008
A common approach to benchmarking a server is to measure its behavior under load from a workload generator. Often a set of such experiments is required—perhaps with different server configurations or workload parameters—to obtain a statistically sound resu ...
Cite
ConferenceFast 2007 5th Usenix Conference on File and Storage Technologies · January 1, 2007
This paper presents the design, implementation, and evaluation of CATS, a network storage service with strong accountability properties. A CATS server annotates read and write responses with evidence of correct execution, and offers audit and challenge int ...
Cite
ConferenceProceedings of the 2nd International Workshop on Hot Topics in Autonomic Computing Hotac 2007 Held in Conjunction with Icac 2007 · January 1, 2007
This paper introduces Automat, a testbed architecture and prototype for research in autonomic services and hosting centers. Automat is an interactive web-based laboratory in which users allocate resources from an ondemand server cluster to experiment with ...
Cite
ConferenceUsenix 2006 Annual Technical Conference · January 1, 2006
This paper presents the design and implementation of Shirako, a system for on-demand leasing of shared networked resources. Shirako is a prototype of a service-oriented architecture for resource providers and consumers to negotiate access to resources over ...
Cite
ConferenceVLDB 2006 Proceedings of the 32nd International Conference on Very Large Data Bases · January 1, 2006
We present the NIMO system that automatically learns cost models for predicting the execution time of computational-science applications running on large-scale networked utilities such as computational grids. Accurate cost models are important for selectin ...
Cite
ConferenceUsenix 2005 Annual Technical Conference · January 1, 2005
Trends towards consolidation and higher-density computing configurations make the problem of heat management one of the critical challenges in emerging data centers. Conventional approaches to addressing this problem have focused at the facilities level to ...
Cite
ConferenceOsdi 2004 6th Symposium on Operating Systems Design and Implementation · January 1, 2004
This paper studies the use of statistical induction techniques as a basis for automated performance diagnosis and performance management. The goal of the work is to develop and evaluate tools for offline and online analysis of system metrics gathered from ...
Cite
ConferenceProceedings of the 3rd Usenix Conference on File and Storage Technologies Fast 2004 · January 1, 2004
Whole-file transfer is a basic primitive for Internet content dissemination. Content servers are increasingly limited by disk arm movement given the rapid growth in disk density, disk transfer rates, server network bandwidth, and content size. Individual f ...
Cite
ConferenceProceedings of the 3rd Usenix Conference on File and Storage Technologies Fast 2004 · January 1, 2004
Losing information when a storage device or data center fails can bring a company to its knees—or put it out of business altogether. Such catastrophic outcomes can readily be prevented with today’s storage technology, albeit with some difficulty: the desig ...
Cite
ConferenceProceedings of the ACM SIGCOMM Workshop on Network I O Convergence Experience Lessons Implications Niceli 2003 · August 25, 2003
Periodic order-of-magnitude jumps in Ethernet bandwidth regularly reawaken interest in TCP/IP transport protocol offload. This time the jump to 10-Gigabit Ethernet coincides with the emergence of new network storage protocols (iSCSI and DAFS), and vendors ...
Full textCite
ConferenceProceedings IEEE Computer Society S Annual International Symposium on Modeling Analysis and Simulation of Computer and Telecommunications Systems Mascots · January 1, 2003
Scalability is the primary challenge to studying large complex network systems with network emulation. This paper studies topology partitioning, assigning disjoint pieces of the network topology across processors, as a technique to increase emulation capac ...
Full textCite
Conference4th Usenix Symposium on Internet Technologies and Systems Usits 2003 · January 1, 2003
Internet service utilities host multiple server applications on a shared server cluster. A key challenge for these systems is to provision shared resources on demand to meet service quality targets at least cost. This paper presents a new approach to utili ...
Cite
ConferenceProceedings of the IEEE International Symposium on High Performance Distributed Computing · January 1, 2003
This paper presents new mechanisms for dynamic resource management in a cluster manager called Cluster-on-Demand (COD). COD allocates servers from a common pool to multiple virtual clusters (vclusters), with independently configured software environments, ...
Full textCite
Conference4th Usenix Symposium on Internet Technologies and Systems Usits 2003 · January 1, 2003
Anypoint is a new model for one-to-many communication with ensemble sites—aggregations of end nodes that appear to the external Internet as a unified site. Policies for routing Anypoint traffic are defined by application-layer plugins residing in extensibl ...
Cite
ConferenceOperating Systems Review ACM · December 31, 2002
This paper presents ModelNet, a scalable Internet emulation environment that enables researchers to deploy unmodified software prototypes in a configurable Internet-like environment and subject them to faults and varying network conditions. Edge nodes runn ...
Full textCite
Conference2002 IEEE Open Architectures and Network Programming Proceedings Openarch 2002 · January 1, 2002
Today, an increasing number of important network services, such as content distribution, replicated services, and storage systems, are deploying overlays across multiple Internet sites to deliver better performance, reliability and adaptability. Currently ...
Full textCite
ConferenceLecture Notes in Computer Science · January 1, 2002
The key principles behind current peer-to-peer research include fully distributing service functionality among all nodes participating in the system and routing individual requests based on a small amount of locally maintained state. The goals extend much ...
Full textCite
ConferenceProceedings of the 2002 Usenix Annual Technical Conference · January 1, 2002
The Direct Access File System (DAFS) is an emerging industrial standard for network-attached storage. DAFS takes advantage of new user-level network interface standards. This enables a user-level file system structure in which client-side functionality for ...
Cite
ConferenceProceedings of the IEEE International Symposium on High Performance Distributed Computing · January 1, 2002
One approach to high-performance processing of massive data sets is to incorporate computation into storage systems. Previous work has shown that this active storage model is effective for a variety of problems. This paper explores opportunities to use act ...
Full textCite
ConferenceProceedings of the 2001 Usenix Annual Technical Conference · January 1, 2001
Large-scale network services such as data delivery often incorporate new functions by interposing intermediaries on the network. Examples of forwarding intermediaries include firewalls, content routers, protocol converters, caching proxies, and multicast s ...
Cite
ConferenceProceedings 2nd IEEE Workshop on Internet Applications Wiapp 2001 · January 1, 2001
Server switches distribute incoming request traffic across the nodes of Internet server clusters and Web proxy cache arrays. These switches are a standard building block for large-scale Internet services, with many commercial products on the market. As Int ...
Full textCite
ConferenceProceedings of the 4th Conference on Symposium on Operating System Design and Implementation Osdi 2000 · October 22, 2000
This paper explores interposed request routing in Slice, a new storage system architecture for high-speed networks incorporating network-attached block storage. Slice interposes a request switching filter | called a /iproxy | along each client's network pa ...
Cite
Conference4th Symposium on Operating System Design and Implementation Osdi 2000 · January 1, 2000
This paper explores interposed request routing in Slice, a new storage system architecture for high-speed networks incorporating networkattached block storage. Slice interposes a request switching filter — called a ßproxy — along each client's network path ...
Cite
ConferenceUsenix 1998 Annual Technical Conference · January 1, 1998
Recent advances in I/O bus structures (e.g., PCI), highspeed networks, and fast, cheap disks have significantly expanded the I/O capacity of desktop-class systems. This paper describes a messaging system designed to deliver the potential of these advances ...
Cite
ConferenceUsenix 1998 Annual Technical Conference · January 1, 1998
While the availability of platform-independent code on the Internet is increasing, third-party code rarely exhibits all of the features desired by end users. Unfortunately, developers cannot foresee and provide for all possible extensions. In this paper, w ...
Cite
ConferenceIEEE International Symposium on High Performance Distributed Computing Proceedings · January 1, 1997
New network technology continues to improve both the latency and bandwidth of communication in computer clusters. The fastest high-speed networks approach or exceed the I/O bus bandwidths of 'gigabit-ready' hosts. These advances introduce new consideration ...
Cite
ConferenceProceedings of the Annual Hawaii International Conference on System Sciences · January 1, 1996
This paper describes object-based runtime support for eficient access to protected objects, i.e., objects belonging to server programs that export protected services to untrusted clients. Modern operating systems use hardware-based protection domains to pr ...
Full textCite
ConferenceProceedings of the 1st Usenix Conference on Operating Systems Design and Implementation Osdi 1994 · November 14, 1994
We propose a technique for maintaining coherency of a transactional distributed shared memory, used by applications accessing a shared persistent store. Our goal is to improve support for fine-grained distributed data sharing in collaborative design applic ...
Cite
ConferenceProceedings of IEEE 4th Workshop on Workstation Operating Systems Wwos 1993 · January 1, 1993
We previously described Opal, an OS environment that has a single virtual address space common to all protection domains, rather than the usual private virtual address space per protection domain (e.g., a Unix process). All threads on an Opal node see the ...
Full textCite
ConferenceConference on Object Oriented Programming Systems Languages and Applications · December 1, 1992
Object-oriented models are a popular basis for supporting uniform sharing of data and services in operating systems, distributed programming systems, and database systems. We term systems that use objects for these purposes object sharing systems. Operatin ...
Cite
ConferenceProceedings of the 5th ACM Sigops European Workshop Models and Paradigms for Distributed Systems Structuring Ew 1992 · September 21, 1992
The recent, appearance of architectures with Hat 64-bit virtual addressing opens an opportunity to reconsider the way our operating systems use virtual address spaces. We are building an operating system called Opal for these wide-address architectures. Th ...
Full textCite
ConferenceInternational Conference on Architectural Support for Programming Languages and Operating Systems ASPLOS · September 1, 1992
Recent microprocessor announcements show a trend toward wide-address computers: architectures that support 64 bits of virtual address space. Such architectures facilitate fundamentally new operating system organizations that promote efficient data sharing ...
Cite
ConferenceProceedings 2nd International Workshop on Object Orientation in Operating Systems Iwooos 1992 · January 1, 1992
An alternative to surrogates is to use ordinary virtual addresses for inter-object referencing. Usually (but not always) this involves mapping distributed or persistent data into specified parts of the application's address space relying on page faults to ...
Full textCite
Conference3rd Workshop on Workstation Operating Systems Wwos 1992 · January 1, 1992
The recent appearance of architectures with flat 64-bit virtual addressing opens an opportunity to reconsider the way in which operating systems use virtual address spaces. An operating system called Opal is being built for these wide-address architectures ...
Full textCite
ConferenceProceedings of the ACM SIGPLAN Symposium on Principles and Practice of Parallel Programming PPOPP · April 1, 1991
Idle workstations in a network represent a significant computing potential. In particular, their processing power can be used by parallel-distributed programs that treat the network as a loosely-coupled multiprocessor. But the set of machines free to parti ...
Full textCite