Jobs Companies Egen Lead Data Engineer (Databricks, PySpark & GCP)

À propos de ce poste Lead Data Engineer (Databricks, PySpark & GCP) chez Egen

Egen · Hybride · Hyderabad

Job Overview:

           
We are looking for a skilled and motivated Lead Data Engineer with strong experience in Python programming, PySpark, Databricks and Google Cloud Platform (GCP) to join our data engineering team. The ideal candidate will be responsible for requirements gathering, designing, architecting the solution, developing, and maintaining robust and scalable ETL (Extract, Transform, Load) & ELT data pipelines. The role involves working with customers directly, gathering requirements, discovery phase,  designing, architecting the solution, using various GCP services, implementing data transformations, data ingestion, data quality, and consistency across systems, and post post-delivery support.


Experience Level:

10 to 16 years of relevant IT experience

Key Responsibilities:

  • Design, develop, test, and maintain scalable ETL data pipelines using Python, PySpark, Databricks & GCP / Azure.

    • Architect the enterprise solutions with various technologies like GCP, Azure, Databricks, PySpark and Spark SQL.

    • Work extensively on Google Cloud Platform (GCP) services such as:

      • Dataflow for real-time and batch data processing

      • Cloud Functions for lightweight serverless compute

      • BigQuery for data warehousing and analytics

      • Cloud Composer for orchestration of data workflows (on Apache Airflow)

      • Google Cloud Storage (GCS) for managing data at scale

      • IAM for access control and security

      • Cloud Run for containerized applications

Should have experience in the following areas :

  • Develop production-grade Databricks notebooks and workflows.

  • Build data transformation pipelines using PySpark and Spark SQL.

  • Implement Delta Lake architecture.

  • Design Bronze, Silver, and Gold data layers using the Medallion Architecture.

  • Implement Databricks Workflows/Jobs and dependency management.

  • Tune Spark jobs for large-scale data processing.

  • Optimize cluster configuration and compute utilization.

  • Implement appropriate partitioning, caching, and file-size optimization strategies.

  • Perform data ingestion from various sources and apply transformation and cleansing logic to ensure high-quality data delivery.

    • Implement and enforce data quality checks, validation rules, and monitoring.

      • Collaborate with data scientists, analysts, and other engineering teams to understand data needs and deliver efficient data solutions.

        • Manage version control using GitHub and participate in CI/CD pipeline deployments for data projects.

          • Write complex SQL queries for data extraction and validation from relational databases such as SQL Server, Oracle, or PostgreSQL.

            • Document pipeline designs, data flow diagrams, and operational support procedures.

Required Skills:

  • 10+ years of hands-on experience in Python for backend or data engineering projects.

    • Strong understanding and working experience with GCP cloud services (especially Dataflow, BigQuery, Cloud Functions, Cloud Composer, etc.).

    • Working experience with Azure Data Factory (ADF), Azure Databricks, Azure Data Lake Storage Gen2 (ADLS).

      • Solid understanding of data pipeline architecture, data integration, and transformation techniques.

        • Experience in working with version control systems like GitHub and knowledge of CI/CD practices.

        • Experience in Apache Spark, Kafka, Redis, Fast APIs, Airflow, GCP Composer DAGs.

          • Strong experience in SQL with at least one enterprise database (SQL Server, Oracle, PostgreSQL, etc.).

          • Experience with PySpark is required.

          • Experience in data migrations from on-premise data sources to Cloud platforms.

          • Good to Have (Optional Skills):

            • Experience with AWS services.

            • Additional Details:

              • Excellent problem-solving and analytical skills.

              • Strong communication skills and ability to collaborate in a team environment.

              • Education:

                ● Bachelor's degree in Computer Science, a related field, or equivalent experience.

Prêt à postuler chez Egen ?
Postuler chez Egen

Emplois similaires

NO
Scientific Computing Engineer - Drug Product Process Modeling & Data Science
Novartis
⚡ Postuler tôt Hyderabad (Office) Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 20 h
Amgen
Associate Data Engineer
Amgen
⚡ Postuler tôt India - Hyderabad Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 1 j
Micron Technology
Sr Engineer, Data
Micron Technology
⚡ Postuler tôt Hyderabad - Phoenix Aquila, In... Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 1 j
PG
Senior Data Engineer
Procter & Gamble
⚡ Postuler tôt HYDERABAD OFFICE INDIA PSC PGH Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 1 j
MSD
Lead Semantic Data Engineer
MSD
⚡ Postuler tôt CZE - Central Bohemian - Pragu... Sur site 🛂 Parrainage de visa
● Nouveau 👁 Vu ✓ Postulé il y a 1 j
NO
Data Engineer
Novartis
⚡ Postuler tôt Hyderabad (Office) Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 1 j
Vanguard
Data Engineer, Specialist
Vanguard
⚡ Postuler tôt Hyderabad, India Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 1 j
Synchrony
AVP Senior Data Engineer L11
Synchrony
⚡ Postuler tôt Hyderabad IN Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 1 j
PwC
IN_Senior Associate_Data Engineer_Data and Analytics_Advisory_Hyderabad
PwC
⚡ Postuler tôt Hyderabad - Salarpuria Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 1 j

Inscrivez-vous pour des suggestions adaptées aux emplois que vous ouvrez et aux recherches que vous enregistrez.

Plus d’emplois chez Egen

Voir tous les emplois chez Egen →

Postuler maintenant
🤖

Doucement — un instant

JobsRadar a été conçu pour de vraies personnes qui traversent une période difficile dans leur recherche d’emploi — pas pour des requêtes automatisées. Vous cliquez beaucoup trop vite et vous êtes maintenant temporairement bloqué.

Revenez plus tard. Si vous cherchez réellement un emploi, nous sommes de votre côté — agissez simplement comme un être humain.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Prenez une longueur d’avance dans votre recherche d’emploi.

Rejoignez notre canal Telegram pour ce qui vous aide à décrocher le poste — références salariales, le pouls hebdomadaire du marché et les annonces de nouveautés. Pas de spam, que du signal.

Rejoindre le canal — c’est gratuit