Title :
Estimating temporary files sizes in distributed realtional database systems
Author :
Tian-Jy Chao;Csaba J. Egyhazy
Author_Institution :
VPI &
Abstract :
The estimated sizes of temporary files are one of the most important statistics used by an Optimizer in generating a minimum cost processing strategy. Statistical information is the primary input to the estimation technique. The amount of statistics kept concerning the key attributes and the distribution of data values within a domain in a relation will greatly affect the accuracy of the estimates. However, the cost of storing this information may outweigh its value. Ultimately such a determination is left to those responsible for designing and implementing distributed realtional DBMSs. This paper presents a new method to calculate temporary files sizes. It also describes specific data structures and algorithms to implement the proposed method. The major tools suggested are a Log File and several specialized data matrices. The latter contain information unique to each relation. The Log Files is a posting files that records the number of occurrences of different values in the database.
Keywords :
"Estimation","Arrays","Binary trees","Database systems","Color"
Conference_Titel :
Data Engineering, 1986 IEEE Second International Conference on
Print_ISBN :
978-0-8186-0655-7
DOI :
10.1109/ICDE.1986.7266200