- Kuala Lumpur, Kuala Lumpur Kuala Lumpur WP Kuala Lumpur Malaysia
Working Location
Job Description
Responsibilities
Position Objective:
Support the design, development, and optimization of modern data pipelines and analytics solutions using Databricks, Python, SQL, and GitHub Copilot. The role provides hands-on exposure to enterprise data engineering practices, cloud-based data platforms, and AI-assisted software development.
Collaborate with data engineers and solution architects to build scalable data workflows, improve data quality and accessibility, and leverage AI tools to accelerate development productivity. Through this internship, the candidate will gain practical experience in delivering end-to-end data & AI solutions that enable data-driven decision-making and business insights.
Roles and Responsibilities:
(1) Data Engineering & Pipeline Development
Assist in designing, developing, and maintaining data pipelines using Databricks.
Build and optimize ETL/ELT processes to ingest, transform, and load data from multiple sources.
Support the development of scalable data models and data products.
Perform data quality validation, monitoring, and troubleshooting activities.
Document data pipelines, workflows, and technical solutions.
(2) Cloud & Analytics Solutions
Work with cloud-based data platforms and modern data architectures.
Assist in integrating data from databases, APIs, files, and enterprise applications.
Support implementation of best practices for data governance, security, and performance optimization.
Participate in testing and deployment activities.
(3) AI-Assisted Development
Leverage GitHub Copilot to accelerate code development and improve productivity.
Explore the use of AI-powered development tools for coding, documentation, testing, and troubleshooting.
Contribute to proof-of-concepts involving Generative AI for data engineering use cases.
Learn and apply prompt engineering techniques to improve development outcomes.
(4) Collaboration & Continuous Learning
Participate in Agile ceremonies, team meetings, and technical discussions.
Collaborate with data engineers, architects, analysts, and product stakeholders.
Research emerging technologies and recommend innovative approaches to improve development efficiency.
Present internship project outcomes and lessons learned to the team.
Qualifications:
Currently pursuing a Bachelor's or Master's degree in Computer Science, Data Science, Software Engineering, Information Technology, Artificial Intelligence, or a related field.
Basic knowledge of Python, SQL, and software development concepts.
Understanding of data engineering fundamentals, including databases, ETL/ELT processes, and data pipelines.
Familiarity with Git/GitHub and an interest in AI-assisted development tools such as GitHub Copilot.
Exposure to Databricks, Apache Spark, cloud platforms, or analytics technologies is an advantage.
Strong analytical thinking, problem-solving abilities, and eagerness to learn new technologies
Important: Please include your internship start & end dates in your application. Only shortlisted candidates will be contacted.
Important Information
Never provide your bank or credit card details when applying for jobs. Do not transfer any money or complete unrelated online surveys. If you see something suspicious, Report this Job ad.