Job description
Years of Experience: 4+ years
Primary Skills: Python, SQL, ETL/ELT, Data Pipelines, Data Modeling, PostgreSQL, MySQL, MongoDB, Cloud (AWS/GCP/Azure)
Preferred Skills: Apache Spark, Databricks, Airflow, Prefect, RAG Pipelines, Embeddings, GenAI Data Pipelines
Responsibilities: Design, develop, and maintain scalable and reliable data pipelines.
Build and optimize ETL/ELT workflows across multiple data sources.
Design data models and data architectures to support analytics, AI, and business applications.
Develop data solutions using Python, SQL, and modern data engineering technologies.
Work with cloud-based data platforms and ensure scalability, reliability, and cost efficiency.
Process and manage large volumes of structured and unstructured data.
Ensure data quality, consistency, security, and availability across pipelines.
Collaborate with AI engineers, software engineers, data scientists, and product teams to support AI and data-driven solutions.
Troubleshoot pipeline and data-related issues and take ownership through resolution.
Contribute to improving data engineering practices, automation, monitoring, and observability.
Evaluate and adopt new technologies that improve the way we build and manage data systems.