CareerScanCareerScan
JobsCompanies
BlogContact
For Employers
Sign InRegister Free
CareerScanCareerScan

India's verified job platform connecting candidates directly with employers. 100% free applications with instant ATS resume scoring.

Chennai, Bengaluru & Hyderabad
Jobs by location
Jobs in ChennaiJobs in BengaluruJobs in HyderabadJobs in PuneJobs in Mumbai
Popular roles
AR Caller JobsHealthcare Medical CodingReact / Full Stack DeveloperData & Power BI AnalystCustomer Support Executive
Top companies
TCS CareersCognizant JobsInfosys OpeningsApollo HospitalsOmega Healthcare
Career services
Free ATS Resume CheckerAI Resume Builder (Free)AI Job MatcherSalary Guide & BenchmarksJob Alerts on WhatsApp
© 2026 CareerScan India. All rights reserved.256-bit SSL encrypted & verified
Back to all jobs
  1. Home
  2. Jobs
  3. Lead Data Engineer (Databricks, PySpark & GCP)
E
Egen

Lead Data Engineer (Databricks, PySpark & GCP)

Hyderabad
10+ years exp
Full-time
Posted 5d ago
2 views
Actively Hiring Urgent Opening Direct 1-Click Apply

Check Your Resume Match Score

Scan your resume against ATS criteria for this Lead Data Engineer (Databricks, PySpark & GCP) role at Egen.

Apply for this position

Apply on Company Website
Notice a broken link or wrong info?

Job Description

  • Job Overview:

We are looking for a skilled and motivated Lead Data Engineer with strong experience in Python programming, PySpark, Databricks and Google Cloud Platform (GCP) to join our data engineering team. The ideal candidate will be responsible for requirements gathering, designing, architecting the solution, developing, and maintaining robust and scalable ETL (Extract, Transform, Load) & ELT data pipelines. The role involves working with customers directly, gathering requirements, discovery phase, designing, architecting the solution, using various GCP services, implementing data transformations, data ingestion, data quality, and consistency across systems, and post post-delivery support.

  • Experience Level:

10 to 16 years of relevant IT experience

Key Responsibilities:

  • Design, develop, test, and maintain scalable ETL data pipelines using Python, PySpark, Databricks & GCP / Azure.

    • Architect the enterprise solutions with various technologies like GCP, Azure, Databricks, PySpark and Spark SQL.

    • Work extensively on Google Cloud Platform (GCP) services such as:

      • Dataflow for real-time and batch data processing
      • Cloud Functions for lightweight serverless compute
      • BigQuery for data warehousing and analytics
      • Cloud Composer for orchestration of data workflows (on Apache Airflow)
      • Google Cloud Storage (GCS) for managing data at scale
      • IAM for access control and security
      • Cloud Run for containerized applications

Should have experience in the following areas :

  • Develop production-grade

Databricks

notebooks and workflows.

  • Build data transformation pipelines using

PySpark

and

Spark SQL

.

  • Implement Delta Lake architecture.

  • Design Bronze, Silver, and Gold data layers using the Medallion Architecture.

  • Implement Databricks Workflows/Jobs and dependency management.

  • Tune Spark jobs for large-scale data processing.

  • Optimize cluster configuration and compute utilization.

  • Implement appropriate partitioning, caching, and file-size optimization strategies.

  • Perform data ingestion from various sources and apply transformation and cleansing logic to ensure high-quality data delivery.

    • Implement and enforce data quality checks, validation rules, and monitoring.

      • Collaborate with data scientists, analysts, and other engineering teams to understand data needs and deliver efficient data solutions.

        • Manage version control using GitHub and participate in CI/CD pipeline deployments for data projects.

          • Write complex SQL queries for data extraction and validation from relational databases such as SQL Server, Oracle, or PostgreSQL.

            • Document pipeline designs, data flow diagrams, and operational support procedures.

Required Skills:

  • 10+ years of hands-on experience in Python for backend or data engineering projects.

    • Strong understanding and working experience with GCP cloud services (especially Dataflow, BigQuery, Cloud Functions, Cloud Composer, etc.).

    • Working experience with Azure Data Factory (ADF), Azure Databricks, Azure Data Lake Storage Gen2 (ADLS).

      • Solid understanding of data pipeline architecture, data integration, and transformation techniques.

        • Experience in working with version control systems like GitHub and knowledge of CI/CD practices.

        • Experience in Apache Spark, Kafka, Redis, Fast APIs, Airflow, GCP Composer DAGs.

          • Strong experience in SQL with at least enterprise database (SQL Server, Oracle, PostgreSQL, etc.).
          • Experience with

PySpark

is required.
- Experience in data migrations from on-premise data sources to Cloud platforms.

  • Good to Have (Optional Skills):

        - Experience with AWS services.
    
  • Additional Details:

          - Excellent problem-solving and analytical skills.
          - Strong communication skills and ability to collaborate in a team environment.
    
  • Education:

  • Bachelor's degree in Computer Science, a related field, or equivalent experience.

Key Requirements & Skills

  • 10+ years of hands-on experience in Python for backend or data engineering projects.

Frequently Asked Questions

How to apply for Lead Data Engineer (Databricks, PySpark & GCP) at Egen?

Click the "Apply on Company Website" button on this page to submit your application directly on the employer's official portal.

What is the salary for this role?

Salary details will be discussed during the interview.

What experience is required?

10+ years of experience is required.

Is this position still open?

Yes, currently active and accepting applications.

ApplicationActively Hiring
Apply on Company Website
Broken link or expired?
E

Egen

Visit Company Website

More jobs at Egen

Client Partner - Early Velocity

Remote

Client Partner- Commercial West

San Francisco, CA

Delivery Lead

Remote

Share this Opening

Job Alerts for data_engineering

Receive email alerts whenever new data_engineering roles in Hyderabad are posted.

Set Free Alert →

Similar Openings

Explore related active roles in data_engineering

View all
UrgentActively Hiring
Momentum Financial Services Group
Lead Data Engineer
Momentum Financial Services Group Verified
10+ years
Salary not disclosed
Hyderabad (Remote)
data_engineeringFull-timeRemote
Posted 20h ago
Apply Now
UrgentActively Hiring
aecom2
Data Engineer
aecom2 Verified
0-2 Yrs
₹111/mo
Bristol, 3 RIVERGATE, gb
data_engineeringFull-time
Posted 20h ago
Apply Now
UrgentActively Hiring
Magnals
Lead Data Engineer
Magnals Verified
5+ years
₹10.7L – ₹12.1L/mo
Remote
data_engineeringFull-timeRemote
Posted 20h ago
Apply Now

Lead Data Engineer (Databricks, PySpark & GCP)

Egen · Hyderabad

Apply on Company Website