TY - GEN
T1 - Fast data analysis with integrated statistical metadata in scientific datasets
AU - Liu, Jialin
AU - Chen, Yong
PY - 2012
Y1 - 2012
N2 - Scientific datasets, such as HDF5 and PnetCDF, have been used widely in many scientific applications. These data formats and libraries provide essential support for data analysis in scientific discovery and innovations. In this research, we present an approach to boost data analysis, namely Fast Analysis with Statistical Metadata (FASM), via data sub setting and integrating a small amount of statistics into datasets. We discuss how the FASM can improve data analysis performance. It is currently evaluated with the PnetCDF on synthetic and real data, but can also be implemented in other libraries. The FASM can potentially lead to a new dataset design and can have an impact on data analysis.
AB - Scientific datasets, such as HDF5 and PnetCDF, have been used widely in many scientific applications. These data formats and libraries provide essential support for data analysis in scientific discovery and innovations. In this research, we present an approach to boost data analysis, namely Fast Analysis with Statistical Metadata (FASM), via data sub setting and integrating a small amount of statistics into datasets. We discuss how the FASM can improve data analysis performance. It is currently evaluated with the PnetCDF on synthetic and real data, but can also be implemented in other libraries. The FASM can potentially lead to a new dataset design and can have an impact on data analysis.
KW - FASM
KW - big data
KW - data intensive computing
KW - high performance computing
KW - statistical techniques
UR - http://www.scopus.com/inward/record.url?scp=84871148080&partnerID=8YFLogxK
U2 - 10.1109/ICPPW.2012.89
DO - 10.1109/ICPPW.2012.89
M3 - Conference contribution
AN - SCOPUS:84871148080
SN - 9780769547954
T3 - Proceedings of the International Conference on Parallel Processing Workshops
SP - 602
EP - 603
BT - Proceedings - 41st International Conference on Parallel Processing Workshops, ICPPW 2012
T2 - 41st International Conference on Parallel Processing Workshops, ICPPW 2012
Y2 - 10 September 2012 through 13 September 2012
ER -