Scooped, Again

Total Page:16

File Type:pdf, Size:1020Kb

Scooped, Again Scooped, Again The Harvard community has made this article openly available. Please share how this access benefits you. Your story matters Citation Ledlie, Jonathan, Jeff Shneidman, Margo Seltzer, and John Huth. 2003. Scooped, again. Peer-to-peer systems II: Second international workshop, IPTPS 2003, Berkeley, California, February 21-22, 2003, ed. IPTPS 2003, Frans Kaashoek, and Ion Stoica. Berlin: Springer Verlang. Previously published in Lecture Notes in Computer Science 2735: 129-138. Published Version http://dx.doi.org/10.1007/b11823 Citable link http://nrs.harvard.edu/urn-3:HUL.InstRepos:2799042 Terms of Use This article was downloaded from Harvard University’s DASH repository, and is made available under the terms and conditions applicable to Other Posted Material, as set forth at http:// nrs.harvard.edu/urn-3:HUL.InstRepos:dash.current.terms-of- use#LAA Scooped, Again Jonathan Ledlie, Jeff Shneidman, Margo Seltzer, John Huth Harvard University jonathan,jeffsh,margo ¡ @eecs.harvard.edu [email protected] Abstract p2p Grid Users (scientists) users Users (disparate The Peer-to-Peer (p2p) and Grid infrastructure commu- groups ) ©§ nities are tackling an overlapping set of problems. In ad- App. ¢¤£¦¥§¥§¨writers place demand demand dressing these problems, p2p solutions are usually moti- demand (e.g., on vated by elegance or research interest. Grid researchers, App. writers OGSA) writers} under pressure from thousands of scientists with real file research and Infrastructure writers Infrastructure writers sharing and computational needs, are pooling their solu- development tions from a wide range of sources in an attempt to meet user demand. Driven by this need to solve large scientific Figure 1: In serving their well-defined user base, the problems quickly, the Grid is being constructed with the Grid community has needed to draw from both its ances- tools at hand: FTP or RPC for data transfer, centraliza- try of supercomputing and from Systems research (in- tion for scheduling and authentication, and an assump- cluding p2p and distributed computing). P2p has essen- tially invented its user base through its technology. tion of correct, obediant nodes. If history is any guide, the World Wide Web depicts viscerally that systems that his own work and the work of other physicists at address user needs can have enormous staying power and CERN led him to develop HTML, HTTP, and a sim- affect future research. The Grid infrastructure is a great ple browser [4]. customer waiting for future p2p products. By no means should we make them our only customers, but we should While HTTP and HTML are simple, they have at least put them on the list. If p2p research does not exhibited serious network and language-based defi- at least address the Grid, it may eventually be sidelined ciencies as the Web has grown [21]. These problems by defacto distributed algorithms that are less elegant but have been examined and patched to some extent, but were used to solve Grid problems. In essense, we’ll have this simple inelegant solution remains at the core of been scooped, again. the Web. A parallel situation exists today with the p2p and 1 Introduction Grid communities. A large group of users (scien- Butler Lampson, in his SOSP99 Invited Talk, tists) are pushing for immediately useable tools to stated that the greatest failure of the systems re- pool large sets of resources. The Grid is currently search community over the past ten years was that building these tools. The Systems community is in “we did not invent the Web” [39]. The systems danger of falling victim to the contemporary version research community laid the groundwork, but did of Lampson’s admonishment if we do not partici- not follow through. The same situation exists to- pate in this process. day with the overlapping efforts of the Grid and p2p The Grid is an active area of research and devel- communities. The former is building and using a opment. The number of academic grids has jumped global resource sharing system, while the latter is six-fold in the last year [43]. As Figure 1 depicts, repeating their mistake of the past by focusing on the “customer base” of scientists, who are often also elegant solutions without regard to a vast potential application writers, drive Grid developers to pro- user community. duce tangible solutions, even when the solutions In 1989, Tim Berner-Lee’s need to communicate are not ideal, from a computer systems perspec- tive. Most current p2p users are people sharing files; 2.3 Manifestations sometimes these people are pursuing noble causes Grids historially arose out of a need to perform like anonymous document distribution, but more of- massive computation. A manifestation of shared ten they are simply trying to circumvent copyright, computation, Condor accepts compiled jobs, sched- and rarely are they interacting with the systems’ de- ules and runs them on remote idle machines [42]. velopers. P2p does not have the driving force that Exemplifying its focus on computation, it issues interactive users provide and therefore has focused RPCs back to the job originator’s machine for data. on solutions that are interesting primarily from a re- An example of computational middleware is Globus search perspective. [23]; it “meta-schedules” jobs among Grids like The difficulty with this parallel development is Condor and ships host data between them using not that it is wasteful, but that, without the p2p com- GridFTP, an FTP wrapper. Manifestations of Data munity’s input, the Grid will most likely be built in Grids include the European Data Grid project [20]. a way incompatible and non-inclusive of many of The Grid community is currently authoring a p2p’s strong points: search and storage scalability, Web Services-oriented API called Open Grid Ser- decentralization, anonymity and pseudonymity, and vices Architecture (OGSA) [24, 25]. The next gen- denial of service prevention. eration of Globus is intended to be a reference im- In the rest of this paper, we first discuss the char- plentation of this API. ters of the two different communities. We introduce three fallacies that may have kept the communities 3 P2P separate. We then describe common problems being 3.1 What is P2P? attacked by the two communities and compare so- lutions from each camp, and discuss problems that Unintentionally, Shirky describes p2p much like seem to be truly disjoint. We conclude with a call to a Grid: “Peer-to-peer is a class of applications that action for the p2p community to examine the Grid take advantage of resources — storage, cycles, con- needs and consider future research problems in that tent, human presence — available at the edges of context. the Internet” [19]. Stoica et al. offer a more restric- tive definition focusing on decentralized and non- 2 Grids hierarchical node organization [54]. 2.1 What is the Grid? P2p’s focus on decentralization, instability, and fault tolerance exemplify areas that essentially have Buyya defines the Grid as “a type of parallel and been omitted from emerging Grid standards, but distributed system that enables the sharing, selec- will become more significant as the system grows. tion, and aggregation of resources distributed across multiple administrative domains based on their (re- 3.2 Goals of P2P sources) availability, capability, performance, cost, The goal of p2p is to take advantage of the idle and users’ quality-of-service requirements” [8]. cycles and storage of the edge of the Internet, effec- The Grid is “distinguished from conventional dis- tively utilizing its “dark matter” [19]. This overar- tributed computing by its focus on large-scale re- ching goal introduces issues including decentraliza- source sharing, innovative applications, and in some tion, anonymity and pseudonymity, redundant stor- cases, high performance orientation” [27]. age, search, locality, and authentication. 2.2 Goals of the Grid 3.3 Manifestations The Grid aims to be self-configuing, self-tuning, O’Reilly’s p2p site [46] divides p2p systems into and self-healing, similar to the goals in autonomic nineteen categories, primarily offering file sharing computing [2]. It aims to fulfill the vision of Cor- (e.g., Gnutella [28], KaZaA [36]), distributed com- bato’s Multics [13]: like a utility company, a mas- putation (e.g., distributed.net [17]), and anonymity sive resource to which a user gives his or her com- (e.g., Freenet [12], Publius [44]). Seti@home is an- putational or storage needs. The Grid’s goal is to other p2p application, although the Grid community utilize the shared storage and cycles from the mid- considers it one of theirs as well [51]. dle and edges of the Internet. Prominant research instances of p2p include tally anonymous file sharing) impose special re- Chord, Pastry, and Tapestry [48, 53, 56]. File quirements for assorted (not always technical) rea- systems built on these include CFS, PAST, and sons. But even in these situations, a general aware- Oceanstore [14, 18, 37]. File systems, however, ness of technical approaches taken by the other are not applications and, while a variety of forms community may help solve “physically private” of storage have been built, there exists no driving problems. application in this space. 4.3 “Grid projects do not have the flexibility to 4 Three Falacies That Have Kept the Com- try new algorithms/ideas because they have munities Separate to get real work done. P2P research is all about this flexibility.” This paper argues that the p2p and Grid commu- P2p research is very flexible: one version can nities are “natural” partners in research. The follow- obsolete the previous and new algorithms can be ing are some objections and responses to this claim. developed without conforming to any standard. 4.1 “The technical problems in Grid systems Within the emerging Grid standards, however, there are different from those in p2p systems.” is room for flexible research too.
Recommended publications
  • A Distributed, Cooperative Citeseer
    OverCite: A Distributed, Cooperative CiteSeer Jeremy Stribling, Jinyang Li,† Isaac G. Councill,†† M. Frans Kaashoek, Robert Morris MIT Computer Science and Artificial Intelligence Laboratory †New York University and MIT CSAIL, via UC Berkeley ††Pennsylvania State University Abstract common for non-commercial Web sites to be popular yet to lack the resources to support their popularity. Users of CiteSeer is a popular online resource for the computer sci- such sites are often willing to help out, particularly in the ence research community, allowing users to search and form of modest amounts of compute power and network browse a large archive of research papers. CiteSeer is ex- traffic. Examples of applications that thrive on volunteer pensive: it generates 35 GB of network traffic per day, re- resources include SETI@home, BitTorrent, and volunteer quires nearly one terabyte of disk storage, and needs sig- software distribution mirror sites. Another prominent ex- nificant human maintenance. ample is the PlanetLab wide-area testbed [36], which is OverCite is a new digital research library system that made up of hundreds of donated machines over many dif- aggregates donated resources at multiple sites to provide ferent institutions. Since donated resources are distributed CiteSeer-like document search and retrieval. OverCite en- over the wide area and are usually abundant, they also al- ables members of the community to share the costs of run- low the construction of a more fault tolerant system using ning CiteSeer. The challenge facing OverCite is how to a geographically diverse set of replicas. provide scalable and load-balanced storage and query pro- In order to harness volunteer resources at many sites, a cessing with automatic data management.
    [Show full text]
  • Donut: a Robust Distributed Hash Table Based on Chord
    Donut: A Robust Distributed Hash Table based on Chord Amit Levy, Jeff Prouty, Rylan Hawkins CSE 490H Scalable Systems: Design, Implementation and Use of Large Scale Clusters University of Washington Seattle, WA Donut is an implementation of an available, replicated distributed hash table built on top of Chord. The design was built to support storage of common file sizes in a closed cluster of systems. This paper discusses the two layers of the implementation and a discussion of future development. First, we discuss the basics of Chord as an overlay network and our implementation details, second, Donut as a hash table providing availability and replication and thirdly how further development towards a distributed file system may proceed. Chord Introduction Chord is an overlay network that maps logical addresses to physical nodes in a Peer- to-Peer system. Chord associates a ranges of keys with a particular node using a variant of consistent hashing. While traditional consistent hashing schemes require nodes to maintain state of the the system as a whole, Chord only requires that nodes know of a fixed number of “finger” nodes, allowing both massive scalability while maintaining efficient lookup time (O(log n)). Terms Chord Chord works by arranging keys in a ring, such that when we reach the highest value in the key-space, the following key is zero. Nodes take on a particular key in this ring. Successor Node n is called the successor of a key k if and only if n is the closest node after or equal to k on the ring. Predecessor Node n is called the predecessor of a key k if and only n is the closest node before k in the ring.
    [Show full text]
  • On the Feasibility of Peer-To-Peer Web Indexing and Search
    University of Pennsylvania ScholarlyCommons Departmental Papers (CIS) Department of Computer & Information Science October 2003 On the Feasibility of Peer-to-Peer Web Indexing and Search Jiyang Li Massachusetts Institute of Technology Boon Thau Loo University of Pennsylvania, [email protected] Joseph M. Hellerstein University of California M Frans Kaashoek Massachusetts Institute of Technology David Karger Massachusetts Institute of Technology See next page for additional authors Follow this and additional works at: https://repository.upenn.edu/cis_papers Recommended Citation Jiyang Li, Boon Thau Loo, Joseph M. Hellerstein, M Frans Kaashoek, David Karger, and Robert Morris, "On the Feasibility of Peer-to-Peer Web Indexing and Search", . October 2003. Postprint version. Published in Lecture Notes in Computer Science, Volume 2735, Peer-to-Peer Systems II, 2003, pages 207-215. Publisher URL: http://dx.doi.org/10.1007/b11823 NOTE: At the time of publication, author Boon Thau Loo was affiliated with the University of California at Berkeley. Currently (April 2007), he is a faculty member in the Department of Computer and Information Science at the University of Pennsylvania. This paper is posted at ScholarlyCommons. https://repository.upenn.edu/cis_papers/328 For more information, please contact [email protected]. On the Feasibility of Peer-to-Peer Web Indexing and Search Abstract This paper discusses the feasibility of peer-to-peer full-text keyword search of the Web. Two classes of keyword search techniques are in use or have been proposed: flooding of queries vo er an overlay network (as in Gnutella), and intersection of index lists stored in a distributed hash table.
    [Show full text]
  • CHR: a Distributed Hash Table for Wireless Ad Hoc Networks
    CHR: a Distributed Hash Table for Wireless Ad Hoc Networks∗ Filipe Araujo´ Lu´ıs Rodrigues Universidade de Lisboa University of Lisboa fi[email protected] [email protected] J¨org Kaiser Changling Liu University of Ulm University of Ulm [email protected] [email protected] Carlos Mitidieri University of Ulm [email protected] Abstract This paper focuses on the problem of implementing a distributed hash table (DHT) in wireless ad hoc networks. Scarceness of resources and node mobility turn routing into a challenging problem and therefore, we claim that building a DHT as an overlay network (like in wired environments) is not the best option. Hence, we present a proof-of-concept DHT, called Cell Hash Routing (CHR), designed from scratch to cope with problems like limited available energy, communication range or node mobility. CHR overcomes these problems, by using position information to organize a DHT of clusters instead of individual nodes. By using position-based routing on top of these clusters, CHR is very efficient. Furthermore, its localized routing and its load sharing schemes, make CHR very scalable in respect to network size and density. For these reasons, we believe that CHR is a simple and yet powerful adaptation of the DHT concept for wireless ad hoc environments. ∗Selected sections of this report were published in the Proceedings of the Fourth Interna- tional Workshop on Distributed Event-Based Systems (DEBS'05), in conjuction with the 25th International Conference on Distributed Computing Systems (ICDCS-25), Columbus, Ohio, USA, June 2005.
    [Show full text]
  • Mast Cornellgrad 0058F 12280.Pdf (2.365Mb)
    Protocols for Building Secure and Scalable Decentralized Applications A Dissertation Presented to the Faculty of the Graduate School of Cornell University in Partial Fulfillment of the Requirements for the Degree of Doctor of Philosophy by Kai Mast December 2020 © 2020 Kai Mast ALL RIGHTS RESERVED Protocols for Building Secure and Scalable Decentralized Applications Kai Mast, Ph.D. Cornell University 2020 Decentralized ledger technologies distribute data and execution across a public peer- to-peer network, which allows for more democratic governance of distributed systems and enables tolerating Byzantine failures. However, current protocols for such decen- tralized ledgers are limited in performance as they require every participant of the protocol to execute and validate every operation. Because of this, systems such as Bitcoin or Ethereum are limited in their throughput to around 10 transaction per sec- ond. Additionally, current implementations provide virtually no privacy to individual users, which precludes decentralized ledgers from being used in many real-world ap- plications. This thesis analyses the scalability and privacy limitations of current protocols and discusses means to improve them in detail. It then outlines two novel protocols for building decentralized ledgers, their implementation, and evaluates their performance under realistic workloads. First, it introduces the BitWeave, a blockchain protocol enabling parallel transac- tion validation and serialization while maintaining the same safety and liveness guaran- tees provided by Bitcoin. BitWeave partitions the system’s workload across multiple distinct shards, each of which then executes transactions mostly independently, while allowing for serializable cross-shard transactions. Second, it discusses DataPods, which is a database architecture and programming abstraction that combines the safety properties of decentralized systems with the scala- bility and confidentiality of centralized systems.
    [Show full text]
  • JAIME TEEVAN [email protected]
    JAIME TEEVAN [email protected] http://www.teevan.org WORK: MICROSOFT RESEARCH HOME One Microsoft Way 13109 NE 38th Place Redmond, WA 98052 Bellevue, WA 98005 (425) 421-9299 (425) 556-9753 EDUCATION Massachusetts Institute of Technology, Cambridge, MA Ph.D., Electrical Engineering and Computer Science, January 2007. Thesis: Supporting Finding and Re-Finding Through Personalization Advisor: Prof. David R. Karger Committee Members: Prof. Mark S. Ackerman, Dr. Susan T. Dumais, Prof. Robert C. Miller S.M., Electrical Engineering and Computer Science, June 2001. Thesis: Improving Information Retrieval with Textual Analysis: Bayesian Models and Beyond Advisor: Prof. David R. Karger Yale University, New Haven, CT B.S., Computer Science, May 1998. Cum laude, with distinction in major. Senior thesis: Automatically Creating High Quality Internet Directories Advisor: Prof. Gregory D. Hager PROFESSIONAL EXPERIENCE Researcher, Microsoft Research, 2006 – Present. Studied the history of digital content (how the content has changed and how people have interacted with it over time) and built tools to take advantage of that history. Developed a community of researchers interested in personal information management. Published numerous papers, presented research to the academic community, and influenced product decisions. Research Assistant, Massachusetts Institute of Technology, Computer Science and Artificial Intelligence Laboratory, 1999 – 2006. Explored re-finding in dynamic information environments. Developed the Re:Search Engine, a tool to support re-finding by preserving previous search contexts. Devised a generative document model for information retrieval that more closely matches real text data than previous naïve Bayesian models. Research Intern, Microsoft Research, Spring 2004. Investigated the variation in goals of people using Web search engines.
    [Show full text]
  • Building Distributed Systems Using Programmable Networks
    c Copyright 2020 Ming Liu Building Distributed Systems Using Programmable Networks Ming Liu A dissertation submitted in partial fulfillment of the requirements for the degree of Doctor of Philosophy University of Washington 2020 Reading Committee: Arvind Krishnamurthy, Chair Luis Ceze Ratul Mahajan Program Authorized to Offer Degree: Computer Science & Engineering University of Washington Abstract Building Distributed Systems Using Programmable Networks Ming Liu Chair of the Supervisory Committee: Professor Arvind Krishnamurthy Computer Science & Engineering The continuing increase of data center network bandwidth, coupled with a slower improvement in CPU performance, has challenged our conventional wisdom regarding data center networks: how to build distributed systems that can keep up with the network speeds and are high-performant and energy-efficient? The recent emergence of a programmable network fabric (PNF) suggests a poten- tial solution. By offloading suitable computations to a PNF device (i.e., SmartNIC, reconfigurable switch, or network accelerator), one can reduce request serving latency, save end-host CPU cores, and enable efficient traffic control. In this dissertation, we present three frameworks for building PNF-enabled distributed systems: (1) IncBricks, an in-network caching fabric built with network accelerators and programmable switches; (2) iPipe, an actor-based framework for offloading distributed applications on Smart- NICs; (3) E3, an energy-efficient microservice execution platform for SmartNIC-accelerated servers. This dissertation presents how to make efficient use of in-network heterogeneous computing re- sources by employing new programming abstractions, applying approximation techniques, co- designing with end-host software layers, and designing efficient control-/data-planes. Our pro- totyped systems using commodity PNF hardware not only show the feasibility of such an approach but also demonstrate that it is an indispensable technique for efficient data center computing.
    [Show full text]
  • Chord: a Scalable Peer-To-Peer Lookup Service for Internet
    Chord: A Scalable Peer-to-peer Lookup Service for Internet Applications ¡ Ion Stoica, Robert Morris, David Karger, M. Frans Kaashoek, Hari Balakrishnan MIT Laboratory for Computer Science [email protected] http://pdos.lcs.mit.edu/chord/ Abstract and involves relatively little movement of keys when nodes join A fundamental problem that confronts peer-to-peer applications is and leave the system. to efficiently locate the node that stores a particular data item. This Previous work on consistent hashing assumed that nodes were paper presents Chord, a distributed lookup protocol that addresses aware of most other nodes in the system, making it impractical to this problem. Chord provides support for just one operation: given scale to large number of nodes. In contrast, each Chord node needs a key, it maps the key onto a node. Data location can be easily “routing” information about only a few other nodes. Because the implemented on top of Chord by associating a key with each data routing table is distributed, a node resolves the hash function by item, and storing the key/data item pair at the node to which the communicating with a few other nodes. In the steady state, in an ¢ -node system, each node maintains information only about £¥¤§¦©¨ key maps. Chord adapts efficiently as nodes join and leave the £¥¤§¦©¨ ¢ system, and can answer queries even if the system is continuously ¢ other nodes, and resolves all lookups via mes- changing. Results from theoretical analysis, simulations, and ex- sages to other nodes. Chord maintains its routing information as nodes join and leave the system; with high probability each such periments show that Chord is scalable, with communication cost £¥¤§¦©¨ and the state maintained by each node scaling logarithmically with event results in no more than ¢ messages.
    [Show full text]
  • Overcite: a Cooperative Digital Research Library
    OverCite: A Cooperative Digital Research Library by Jeremy Stribling Submitted to the Department of Electrical Engineering and Computer Science in partial fulfillment of the requirements for the degree of Master of Science in Computer Science and Engineering at the MASSACHUSETTS INSTITUTE OF TECHNOLOGY September 2005 ) Massachusetts Institute of Technology 2005. All rights reserved. Author ......... ................. Department of Electri Engineering and CIter Science September 2, 2005 ^ /,\ la /7 Certified by ......................... 1 K/ I M. Frans Kaashoek Professor Thesis Supervisor 7pb., - ,,, -.. 7.B- Z-111" Acceptedby......... ( .............. Arthur C. Smith Chairman, Department Committee on Graduate Students MASSACHUSEtITSINSmTE OF TECHNOLOGY ARCHIVfS F71·· YTZAAAA MAKRL ZUUb LIBRARIES 2 OverCite: A Cooperative Digital Research Library by Jeremy Stribling Submitted to the Department of Electrical Engineering and Computer Science on September 2, 2005, in partial fulfillment of the requirements for the degree of Master of Science in Computer Science and Engineering Abstract CiteSeer is a well-known online resource for the computer science research community, allowing users to search and browse a large archive of research papers. Unfortunately, its current centralized incarnation is costly to run. Although members of the community would presumably be willing to donate hardware and bandwidth at their own sites to assist CiteSeer, the current architecture does not facilitate such distribution of resources. OverCite is a design for a new architecture for a distributed and cooperative research library based on a distributed hash table (DHT). The new architecture harnesses donated resources at many sites to provide document search and retrieval service to researchers worldwide. A preliminary evaluation of an initial OverCite prototype shows that it can service more queries per second than a centralized system, and that it increases total storage capacity by a factor of n/4 in a system of n nodes.
    [Show full text]
  • Overcite: a Distributed, Cooperative Citeseer
    OverCite: A Distributed, Cooperative CiteSeer Jeremy Stribling, Jinyang Li,† Isaac G. Councill,†† M. Frans Kaashoek, Robert Morris MIT Computer Science and Artificial Intelligence Laboratory †New York University and MIT CSAIL, via UC Berkeley ††Pennsylvania State University Abstract common for non-commercial Web sites to be popular yet to lack the resources to support their popularity. Users of CiteSeer is a popular online resource for the computer sci- such sites are often willing to help out, particularly in the ence research community, allowing users to search and form of modest amounts of compute power and network browse a large archive of research papers. CiteSeer is ex- traffic. Examples of applications that thrive on volunteer pensive: it generates 35 GB of network traffic per day, re- resources include SETI@home, BitTorrent, and volunteer quires nearly one terabyte of disk storage, and needs sig- software distribution mirror sites. Another prominent ex- nificant human maintenance. ample is the PlanetLab wide-area testbed [36], which is OverCite is a new digital research library system that made up of hundreds of donated machines over many dif- aggregates donated resources at multiple sites to provide ferent institutions. Since donated resources are distributed CiteSeer-like document search and retrieval. OverCite en- over the wide area and are usually abundant, they also al- ables members of the community to share the costs of run- low the construction of a more fault tolerant system using ning CiteSeer. The challenge facing OverCite is how to a geographically diverse set of replicas. provide scalable and load-balanced storage and query pro- In order to harness volunteer resources at many sites, a cessing with automatic data management.
    [Show full text]
  • Algorithmic Engineering Towards More Efficient Key-Value Systems Bin
    Algorithmic Engineering Towards More Efficient Key-Value Systems Bin Fan CMU-CS-13-126 December 18, 2013 School of Computer Science Carnegie Mellon University Pittsburgh, PA 15213 Thesis Committee: David G. Andersen, Chair Michael Kaminsky, Intel Labs Garth A. Gibson Edmund B. Nightingale, Microsoft Research Submitted in partial fulfillment of the requirements for the degree of Doctor of Philosophy. Copyright © 2013 Bin Fan The views and conclusions contained in this document are those of the author and should not be interpreted as representing the official policies, either expressed or implied, of any sponsoring institution, the U.S. government or any other entity. Keywords: Memory Efficiency, Hash Table, Bloom Filter, Caching, Load Bal- ancing Dedicated to my family iv Abstract Distributed key-value systems have been widely used as elemental components of many Internet-scale services at sites such as Amazon, Facebook and Twitter. This thesis examines a system design approach to scale existing key-value systems, both horizontally and vertically, by carefully engineering and integrating techniques that are grounded in re- cent theory but also informed by underlying architectures and expected workloads in practice. As a case study, we re-design FAWN-KV—a dis- tributed key-value cluster consisting of “wimpy” key-value nodes—to use less memory but achieve higher throughput even in the worst case. First, to improve the worst-case throughput of a FAWN-KV system, we propose a randomized load balancing scheme that can fully utilize all the nodes regardless of their query distribution. We analytically prove and empirically demonstrate that deploying a very small but ex- tremely fast load balancer at FAWN-KV can effectively prevent uneven or dynamic workloads creating hotspots on individual nodes.
    [Show full text]
  • P43-Balakrishnan.Pdf
    LOOKING UP in P2P DATA Systems By Hari Balakrishnan, M. Frans Kaashoek, David Karger, Robert Morris, and Ion Stoica he main challenge in P2P computing is to design and imple- ment a robust and scalable distributed system composed of inexpensive, individually unreliable computers in unrelated administrative domains. The participants in a typical P2P Tsystem might include computers at homes, schools, and businesses, and can grow to several million concurrent participants. P2P systems are attractive for P2P computing raises many several reasons: interesting research problems in distributed systems. In this article • The barriers to starting and we will look at one of them, the growing such systems are low, lookup problem. How do you find since they usually don’t require any given data item in a large P2P any special administrative or system in a scalable manner, with- financial arrangements, out any centralized servers unlike centralized or hierarchy? This problem facilities; is at the heart of any P2P • P2P systems offer a way system. It is not addressed to aggregate and make use well by most popular sys- of the tremendous com- tems currently in use, and it putation and storage provides a good example of resources on computers across how the challenges of designing the Internet; and P2P systems can be addressed. • The decentralized and distrib- The recent algorithms devel- uted nature of P2P systems oped by several research groups for gives them the potential to be the lookup problem present a sim- robust to faults or intentional ple and general interface, a distrib- attacks, making them ideal for uted hash table (DHT).
    [Show full text]