We are looking for a Senior Data Engineer to build and maintain scalable data platforms that support AI, analytics, and backend applications. You will be responsible for designing data architectures, developing ETL pipelines, integrating data from multiple sources, and ensuring high-quality, reliable datasets for downstream systems.
Key Responsibilities
- Design and develop scalable ETL pipelines.
- Integrate, clean, and standardize data from multiple sources (databases, APIs, files, logs, etc.).
- Design and maintain Data Lake, Data Warehouse, and data models.
- Build and optimize data pipeline architecture for scalability and reliability.
- Implement data quality processes, including data validation, duplicate detection, and missing value handling.
- Build Feature Store or Feature Pipelines to support AI and machine learning.
- Implement data versioning to ensure reproducibility and traceability.
- Monitor, troubleshoot, and optimize production data pipelines.
- Collaborate with AI Engineers, Backend Engineers, and Data Analysts to deliver reliable data solutions.
Requirements
- Bachelor‘s degree in Computer Science, Information Technology, Information Systems, or a related field.
- Minimum 5 years of experience as a Data Engineer or in a similar role.
- Strong SQL and Python programming skills.
- Proven experience building ETL/ELT pipelines and processing data from multiple sources.
- Hands-on experience designing Data Lake, Data Warehouse, data pipeline architecture, and data models.
- Experience implementing data validation, duplicate detection, and missing value handling.
- Experience building Feature Store or Feature Pipelines for AI/ML applications.
- Experience with data versioning and production pipeline monitoring.
- Basic English reading skills for technical documentation.
Preferred Skills
- Apache Airflow, Spark, Kafka.
- PostgreSQL.
- Microsoft Excel.
- NLP Data for AI.