← к выдаче
разработка

Data Engineer - IT

Guardian7 дн. назад

Коротко

Builds scalable data pipelines for structured and semi-structured data using Databricks, PySpark/Apache Spark, and advanced SQL; handles ETL, data warehousing · Needs: Based in Chennai or Gurgaon.

Описание от работодателя

Job Description: Required Skills & Experience Strong hands-on experience in: Databricks (Lakehouse platform, notebooks, jobs, clusters) PySpark / Apache Spark Advanced SQL (joins, window functions, performance tuning) Solid understanding of ETL concepts, data warehousing, and data modeling Experience building scalable data pipelines for structured and semi-structured data Familiarity with data lake / delta lake architecture Good understanding of performance optimization techniques in Spark and SQL Experience with version control (Git) and deployment practices Strong problem-solving and analytical skills Good to Have Experience with cloud platforms (AWS) Exposure to CI/CD pipelines (Jenkins) Knowledge of Delta Lake, Unity Catalog, or similar governance tools Understanding of data quality frameworks / monitoring tools Experience working in Agile / Scrum teams Location: This position can be based in any of the following locations: Chennai, Gurgaon Current Guardian Colleagues: Please apply through the internal Jobs Hub in Workday