Jobs Companies Caterpillar Senior Data Engineer

À propos de ce poste Senior Data Engineer chez Caterpillar

Caterpillar · Sur site · Bangalore, Karnataka

Career Area:

Technology, Digital and Data

Job Description:

Your Work Shapes the World at Caterpillar Inc.

When you join Caterpillar, you're joining a global team who cares not just about the work we do – but also about each other.  We are the makers, problem solvers, and future world builders who are creating stronger, more sustainable communities. We don't just talk about progress and innovation here – we make it happen, with our customers, where we work and live. Together, we are building a better world, so we can all enjoy living in it.

Position Overview 

We are seeking a highly skilled and experienced Senior Data Engineer with strong expertise in PySpark, Azure Databricks, ETL/ELT, Microsoft Azure, and Microsoft Fabric. The candidate will design, develop, deploy, and maintain scalable enterprise data solutions, with a strong focus on Medallion Architecture, Lakehouse patterns, secure data sharing, Fabric capacity utilization, data quality, performance, governance, and operational reliability. 

Key Responsibilities 

  • Design, develop, and maintain scalable batch and streaming data pipelines using PySpark, Azure Databricks, and Microsoft Fabric. 
  • Architect and implement Medallion Architecture across Bronze, Silver, and Gold layers for enterprise data processing and analytics. 
  • Build robust ETL/ELT solutions for data ingestion, transformation, validation, reconciliation, and delivery across multiple source systems. 
  • Develop and optimize PySpark and Spark SQL workloads for high-volume structured, semi-structured, and unstructured data. 
  • Design and maintain Lakehouse and data lake solutions using Azure Data Lake Storage Gen2, Delta Lake, Microsoft Fabric OneLake, Fabric Lakehouse, and Warehouse. 
  • Implement integration solutions using Azure Data Factory, Fabric Data Factory, data pipelines, notebooks, and Dataflows Gen2. 
  • Design secure and governed data-sharing solutions across workspaces, domains, business units, and approved external consumers. 
  • Implement reusable data products and cross-workspace sharing patterns using OneLake, OneLake shortcuts, Lakehouse, Warehouse, and semantic models. 
  • Contribute to Microsoft Fabric capacity planning, workspace-to-capacity assignment, workload monitoring, utilization analysis, and performance optimization. 
  • Monitor Fabric workloads using available capacity and workload metrics, identify resource contention, and recommend workload or scheduling improvements. 
  • Design domain-aligned Fabric workspace structures with appropriate separation for development, testing, production, security, and ownership boundaries. 
  • Implement data quality controls, monitoring, observability, lineage, error handling, reconciliation, and auditability across data pipelines. 
  • Apply security best practices using managed identities, role-based access control, workspace roles, row-level or object-level controls where applicable, and secure secrets management. 
  • Integrate data engineering solutions with Git-based source control and CI/CD pipelines for automated testing and deployment. 
  • Optimize performance, scalability, reliability, and cost across Azure Databricks and Microsoft Fabric workloads. 
  • Collaborate with Data Architects, Product Owners, Business Analysts, Data Scientists, QA engineers, governance teams, and platform teams in an Agile/Scrum environment. 
  • Provide technical leadership, conduct design and code reviews, establish engineering standards, and mentor data engineers. 

Required Skills & Qualifications 

  • Bachelor's degree in Computer Science, Engineering, Information Systems, or a related field. 
  • 6+ years of proven Data Engineering experience, including delivery of enterprise-scale cloud data platforms. 
  • Excellent hands-on expertise in PySpark, Spark SQL, DataFrame APIs, debugging, reusable framework development, and performance optimization. 
  • Strong hands-on experience with Azure Databricks, including notebooks, jobs/workflows, clusters, Delta Lake, and production deployment patterns. 
  • Practical implementation experience with Medallion Architecture, including Bronze, Silver, and Gold data layers. 
  • Strong knowledge of ETL/ELT processes, ingestion patterns, incremental processing, transformation frameworks, and workflow orchestration. 
  • Hands-on Microsoft Fabric experience, including Fabric Data Factory, notebooks, Lakehouse, Warehouse, OneLake, data pipelines, Dataflows Gen2, and semantic models. 
  • Experience with Fabric capacity concepts, capacity assignment, utilization monitoring, workload analysis, performance tuning, and capacity-aware solution design. 
  • Experience designing Fabric workspaces across domains and environments, including access, ownership, deployment, and workload-isolation considerations. 
  • Strong experience implementing enterprise data-sharing patterns using OneLake shortcuts, shared Lakehouse or Warehouse data, governed data products, and semantic models. 
  • Hands-on Azure experience, including Azure Data Lake Storage Gen2, Azure Data Factory, Azure Synapse Analytics, Azure Key Vault, and Azure Monitor. 
  • Advanced SQL skills with experience in data modelling, data warehousing, query optimization, and performance tuning. 
  • Experience with Delta Lake, Parquet, schema evolution, partitioning, and modern Lakehouse architecture patterns. 
  • Strong understanding of data quality, metadata management, governance, lineage, security, privacy, and operational monitoring. 
  • Experience with Azure DevOps or equivalent Git-based repositories and CI/CD pipelines for data engineering solutions. 
  • Excellent problem-solving, troubleshooting, communication, and stakeholder-management skills. 

Preferred Skills 

  • Experience delivering Microsoft Fabric solutions in enterprise environments with multiple workspaces and shared capacity. 
  • Experience interpreting capacity and workload metrics and recommending performance, scheduling, scaling, or workload-distribution improvements. 
  • Knowledge of streaming architectures, event-driven processing, Eventstreams, and real-time analytics. 
  • Azure Data Engineer, Azure Databricks, or Microsoft Fabric certifications are preferred. 

 

Posting Dates:

September 14, 2026 - September 20, 2026

Caterpillar is an Equal Opportunity Employer.  Qualified applicants of any age are encouraged to apply

Not ready to apply? Join our Talent Community.

Prêt à postuler chez Caterpillar ?
Postuler chez Caterpillar

À propos de Caterpillar

There’s more to work at Caterpillar than just the work itself. We hire smart, friendly people and it shows in our culture. We hold ourselves to high standards and make sure our values of integrity, excellence, teamwork, commitment and sustainability come to life in the way we work. We make sure our employees feel continuously challenged while also supported. We provide professional growth opportunities, including leadership programs. We celebrate the diversity of our team, while also working together as one Caterpillar. Our culture, like everything at our company, is made possible by each employee’s contribution. Person by person, we create the environment we work in, and we are proud of the

Voir tous les emplois chez Caterpillar →

Emplois similaires

Take-Two Interactive Software, Inc.
Data Engineer
Take-Two Interactive Software, Inc.
⚡ Postuler tôt Bangalore, Karnataka, India Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 9 h
Quince
Software Development Engineer II (Data Engineering)
Quince
⚡ Postuler tôt Bengaluru, Karnataka, India Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 3 j
Take-Two Interactive Software, Inc.
Senior Data Engineer
Take-Two Interactive Software, Inc.
⚡ Postuler tôt Bangalore, Karnataka, India Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 4 j
Caterpillar
IT Analyst Applications ( Data Engineer – AI & Analytics )
Caterpillar
⚡ Postuler tôt Bangalore, Karnataka Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 4 j
Caterpillar
Senior Data Specialist (Lead Data Engineer)
Caterpillar
⚡ Postuler tôt Bangalore, Karnataka Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 4 j
2K
Data Quality Engineer
2K
⚡ Postuler tôt Bangalore, Karnataka, India Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 4 j
Caterpillar
Data Specialist (Senior Data Engineer)
Caterpillar
⚡ Postuler tôt Bangalore, Karnataka Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 6 j
Kyndryl
Senior Data Engineer
Kyndryl
⚡ Postuler tôt Greater Noida, Uttar Pradesh,... Hybride
● Nouveau 👁 Vu ✓ Postulé il y a 1 sem.
Kyndryl
Data Analyst / Engineer
Kyndryl
⚡ Postuler tôt Bangalore, Karnataka, India Hybride
● Nouveau 👁 Vu ✓ Postulé il y a 1 sem.

Inscrivez-vous pour des suggestions adaptées aux emplois que vous ouvrez et aux recherches que vous enregistrez.

Plus d’emplois chez Caterpillar

Voir tous les emplois chez Caterpillar →

Postuler maintenant
🤖

Doucement — un instant

JobsRadar a été conçu pour de vraies personnes qui traversent une période difficile dans leur recherche d’emploi — pas pour des requêtes automatisées. Vous cliquez beaucoup trop vite et vous êtes maintenant temporairement bloqué.

Revenez plus tard. Si vous cherchez réellement un emploi, nous sommes de votre côté — agissez simplement comme un être humain.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Prenez une longueur d’avance dans votre recherche d’emploi.

Rejoignez notre canal Telegram pour ce qui vous aide à décrocher le poste — références salariales, le pouls hebdomadaire du marché et les annonces de nouveautés. Pas de spam, que du signal.

Rejoindre le canal — c’est gratuit