A Systematic Review of Common Beginner Programming Mistakes in Data Engineering

Neuwinger M, Riehle D (2025)


Publication Language: English

Publication Type: Conference contribution

Publication year: 2025

Journal

Publisher: Institute of Electrical and Electronics Engineers Inc.

Pages Range: 170-181

Conference Proceedings Title: 2025 IEEE/ACM 37th International Conference on Software Engineering Education and Training (CSEE&T)

Event location: Ottawa, ON CA

ISBN: 979-8-3315-3710-4

DOI: 10.1109/CSEET66350.2025.00024

Abstract

The design of effective programming languages, libraries, frameworks, tools, and platforms for data engineering strongly depends on their ease and correctness of use. Anyone who ignores that it is humans who use these tools risks building tools that are useless, or worse, harmful. To ensure our data engineering tools are based on solid foundations, we performed a systematic review of common programming mistakes in data engineering. We focus on programming beginners (students) by analyzing both the limited literature specific to data engineering mistakes and general programming mistakes in languages commonly used in data engineering (Python, SQL, Java). Through analysis of 21 publications spanning from 2003 to 2024, we synthesized these complementary sources into a comprehensive classification that captures both general programming challenges and domain-specific data engineering mistakes. This classification provides an empirical foundation for future tool development and educational strategies. We believe our systematic categorization will help researchers, practitioners, and educators better understand and address the challenges faced by novice data engineers.

Authors with CRIS profile

How to cite

APA:

Neuwinger, M., & Riehle, D. (2025). A Systematic Review of Common Beginner Programming Mistakes in Data Engineering. In 2025 IEEE/ACM 37th International Conference on Software Engineering Education and Training (CSEE&T) (pp. 170-181). Ottawa, ON, CA: Institute of Electrical and Electronics Engineers Inc..

MLA:

Neuwinger, Max, and Dirk Riehle. "A Systematic Review of Common Beginner Programming Mistakes in Data Engineering." Proceedings of the 37th IEEE/ACM International Conference on Software Engineering Education and Training, CSEE and T 2025, Ottawa, ON Institute of Electrical and Electronics Engineers Inc., 2025. 170-181.

BibTeX: Download