Designing Distributed Database Systems for Efficient Operation
Sangkyu Rho, Salvatore T. March
Abstract
Sangkyu Rho, Salvatore T. March
Abstract
Distributed database systems can yield significant cost and performance advantages over centralized systems for geographically distributed organizations. The efficiency of a distributed database depends primarily on the data allocation (data replication and placement) and the operating strategies (where and how retrieval and update query processing operations are performed). We develop a distributed database design approach that comprehensively treats data allocation and operating strategies, explicitly modeling their interdependencies for both retrieval and updateprocessing. Wedemonstratethatdatareplication,joinnodeselection,anddatareductionbysemijoinare important design and operating decisions that have significant impact on both the cost and response time of a distributed database system.
OpenAlex reports 5 citations for this work. Citation counts describe recorded attention and do not establish research quality.
A contribution statement is not available in the OpenAlex record.
Method details are not available in the OpenAlex metadata.
Findings are not separately available in the OpenAlex metadata.
Limitations are not available in the OpenAlex metadata.
Application details are not available in the OpenAlex metadata.
Distributed database systems can yield significant cost and performance advantages over centralized systems for geographically distributed organizations. The efficiency of a distributed database depends primarily on the data allocation (data replication and placement) and the operating strategies (where and how retrieval and update query processing operations are performed). We develop a distributed database design approach that comprehensively treats data allocation and operating strategies, explicitly modeling their interdependencies for both retrieval and updateprocessing. Wedemonstratethatdatareplication,joinnodeselection,anddatareductionbysemijoinare important design and operating decisions that have significant impact on both the cost and response time of a distributed database system.
Key concepts: Computer science, Distributed database, Replication (statistics), Distributed computing, Database, Database testing, Database tuning, Distributed data store