We are looking for a Data Engineer to join our analytics team. You will design, build and maintain scalable data pipelines, data warehouses and cloud-based data platforms that support analytics and business decision-making.
You will work closely with cross-functional teams to transform raw data into reliable, accessible datasets while improving the performance, quality and efficiency of our data systems.
Responsibilities
- Design, build, test and maintain databases, data warehouses, data lakes and large-scale data processing systems.
- Develop reliable data pipelines to support analytics, data modelling and production workloads.
- Build and continuously improve ETL/ELT processes based on business requirements.
- Collect, clean and transform structured, semi-structured and unstructured data.
- Prepare data for descriptive, predictive and prescriptive analytics.
- Develop and optimise Oracle SQL, PL/SQL, stored procedures and batch jobs.
- Improve data quality, reliability, availability and processing efficiency.
- Organise and maintain data assets and catalogues for easy access and retrieval.
- Support routine and ad-hoc data requirements from analytics teams, stakeholders and business users.
- Work closely with engineering, analytics and business teams to deliver practical data solutions.
Requirements
- Bachelor’s degree in Computer Science, Information Technology, Engineering or a related field
- At least 3 years of relevant experience in data engineering, data warehousing, data processing, data modelling or ETL/ELT development.
- Strong experience in SQL and Oracle PL/SQL development.
- Hands-on experience building and maintaining data pipelines and data warehouses.
- Experience using Python, Java or another programming language for data processing and automation.
- Experience working with cloud-based data platforms, preferably AWS.
- Familiarity with AWS services such as S3, EC2, Redshift, Athena, EMR or Kinesis.
- Experience with one or more data engineering technologies, such as Apache Airflow, Apache Spark, Hadoop, HDFS, Scala or Hive.
- Good understanding of database architecture, schema design and performance optimisation.
- Strong analytical, problem-solving and communication skills.
- Ability to work effectively with both technical and business teams.
Preferred Skills
- Experience with real-time data streaming solutions.
- Familiarity with Kubernetes, containerisation and CI/CD pipelines.
- Experience deploying data services or microservices.
- Experience in data crawling, data lake development or large-scale data processing.
- Knowledge of machine learning, statistical modelling or algorithm development.
- Relevant certifications in cloud computing, data engineering, software development, databases, data science or machine learning.