References
[1]
Rubin, D. B., “Inference and Missing
Data,” Biometrika, Vol. 63, No. 3, 1976, pp. 581–592. https://doi.org/10.2307/2335739
[2]
Youssef, M., Mohammed, L., and Karim, M.,
“Optimizing Data Pipelines for Green AI: A Comparative Analysis of
Pandas, Polars, and PySpark for CO2 Emission Prediction,”
Computers, Vol. 14, No. 8, 2025, p. 319.
[3]
Dasu, T., and Johnson, T., “Exploratory
Data Mining and Data Cleaning,” Journal of Statistical
Software, Vol. 59, No. 10, 2003, pp. 1–23.
[4]
Wickham, H., “Tidy Data,”
Journal of Statistical Software, Vol. 59, No. 10, 2014, pp.
1–23.
[5]
Codd, E. F., “The Relational Model for
Database Management,” Journal of Statistical Software,
Vol. 59, No. 10, 1990, pp. 1–23.
[6]
Wickham, H., “The Split-Apply-Combine
Strategy for Data Analysis,” Journal of Statistical
Software, Vol. 45, No. 1, 2011, pp. 1–19. https://www.jstatsoft.org/article/view/v045i10
[7]
Kandel, S., Heer, J., Plaisant, C., Kennedy,
J., Van Ham, F., Riche, N. H., Weaver, C., Lee, B., Brodbeck, D., and
Buono, P., “Research Directions in Data Wrangling: Visualizations
and Transformations for Usable and Credible Data,”
Information Visualization, Vol. 10, No. 4, 2011, pp. 271–288.
https://doi.org/10.1177/1473871611415994
[8]
맥키니, “파이썬 라이브러리를 활용한
데이터 분석,” 한빛미디어, 서울, 2023.
[9]
패스캐버, “판다스 인 액션,”
한빛미디어, 서울, 2022.
[10]
브루스, 브루스, and 게데크, “데이터
과학을 위한 통계,” 한빛미디어, 서울, 2018.
[11]
윌케, “데이터 시각화 교과서: 데이터
분석의 본질을 살리는 그래프와 차트 제작의 기본 원리와 응용,”
책만, 서울, 2020.
[12]
M.
웡, 도나, “월스트리트저널 인포그래픽 가이드: 데이터, 사실, 수치를
시각적으로 표현하는 법,” 인사이트, 서울, 2014.
[13]
황재진, and 윤영진, “사례 분석으로 배우는
데이터 시각화: 막대 차트부터 대시보드까지 태블로로 실습하며 배우는
인사이트 도출법,” 한빛미디어, 서울, 2022.
[14]
McKinney, W., “Pandas: A
Foundational Python Library for Data Analysis and
Statistics,” 2011.
[15]
JetBrains, “The
State of Data Science 2024: 6 Key Data Science Trends,”
2024.
[16]
Kestra, “Embedded Databases in
2026: DuckDB, SQLite, Polars, and chDB,” 2026.