MSD

Specialist , Data Engineering

Hyderabad, Telangana
5 years exp
Day Shift
Posted 18h ago
0 views
Actively Hiring Direct 1-Click Apply

Check Your Resume Match Score

Scan your resume against ATS criteria for this Specialist , Data Engineering role at MSD.

Apply for this position

Apply on Company Website

Job Description

Job Description


Specialist:

Data Engineering


The Opportunity:

Join a global biopharma company with a 130-year legacy and mission to achieve new milestones in healthcare. Be part of a technology-driven, data-led organization supporting a diversified portfolio of medicines, vaccines, and animal health products. Work alongside passionate teams that use data, analytics, and insights to drive decisions and tackle some of the world’s greatest health threats.


Our Technology Centers are globally distributed hubs that enable our digital transformation and business outcomes across IT. They bring together diverse teams to collaborate, share best practices, and deliver solutions that save and improve lives.


This role is based at our Hyderabad Tech Center and follows a hybrid working model (3 days onsite, 2 days remote). Candidates are expected to reside within commuting distance of the Hyderabad office.


Role Overview


We are hiring ahands-on Data Engineer who can design, build, and operate production-grade data platforms and pipelines end to end. You will deliver reliable, governed, secure, and analytics-ready data by implementing moderndata warehousing andLakehouse patterns onAWS andDatabricks, with strong focus ondata quality,dimensional modeling, and scalableETL/ELT. This role partners closely with analytics, data science, and business stakeholders to translate requirements into robust datasets, while applying engineering best practices such as testing, code reviews, CI/CD, and observability.


What will you do in this role


  • Design, build, and operatebatch and streaming data pipelines to ingest data from multiple sources into anAWS data lake / lakehouse anddata warehouse.
  • Develop and maintainETL/ELT transformations usingPython,PySpark, and SQL; optimize jobs for performance, cost, and reliability.
  • Partner with Data Analysts, Data Scientists, and business stakeholders to understand use cases and delivercurated, analytics-ready datasets and features.
  • Implementdata quality controls (validation rules, reconciliation, anomaly checks), defineSLAs/SLOs, and contribute tometadata, lineage, anddata catalog practices.
  • Use orchestration and observability to run pipelines reliably (e.g.,Databricks Workflows,AWS Step Functions, scheduling, logging, monitoring, alerting).
  • Apply engineering best practices: unit/integration testing, automateddata tests, code reviews, and quality gates withinCI/CD.
  • Model and publish data for BI/analytics usingdimensional modeling (star/snowflake), facts dimensions, andslowly changing dimensions (SCD).
  • Write and tuneadvanced SQL for profiling, transformations, and performance troubleshooting across large datasets.
  • Build onAWS using services such asS3,Glue,Lambda,Step Functions,EMR, and CloudWatch; follow security best practices (IAM, encryption, least privilege).
  • Provision and manage cloud resources usingInfrastructure as Code (e.g.,Terraform) across dev/test/prod environments.
  • Package and deploy workloads usingDocker (and where applicable ECS/Fargate); manage dependencies and runtime configurations.
  • UseGitHub for version control (branching strategies, pull requests, code reviews) and set upCI/CD for automated build, test, and deployment.
  • Develop scalable processing onDatabricks /Apache Spark usingPySpark and lakehouse concepts (e.g.,Delta Lake, ACID, schema evolution).
  • Use notebooks (e.g., Jupyter/Databricks) for exploration and PoCs, then productionize solutions with reusable modules, tests, and deployment pipelines.
  • Work in anAgile delivery model (planning, daily sync, reviews, retros), providing accurate estimates and proactively managing risks/dependencies.
  • Create and maintain technical documentation (data contracts, pipeline specs, runbooks) and support operational handoffs.
What Should you have:
  • 5 years of hands-on experience in data engineering building production pipelines and data5 years of hands-on experience in data engineering building production pipelines and data platforms.
  • StrongAWS experience: S3,Glue,Lambda,Step Functions,EMR (and/or ECS/Fargate), plus CloudWatch; solid grasp of IAM and encryption.
  • Nice to have: AWS certification (Developer/Architect) or equivalent demonstrated expertise.
  • Experience working inAgile teams; strong collaboration, communication, and stakeholder management skills.
  • Experience withDatabricks and lakehouse capabilities (e.g., Delta Lake, job/workflow orchestration, cluster tuning) is strongly preferred.
  • StrongSQL skills including complex joins/window functions, data profiling, and performance tuning; understanding ofdimensional modeling concepts.
  • Proficient inPython andPySpark with solid Spark fundamentals (partitioning, shuffle, caching, file formats) and ability to debug/optimize.
  • Strong withGitHub,CI/CD concepts, and engineering practices (code reviews, branching, release management); working knowledge ofDocker andTerraform.
  • Demonstrated ability to work across teams, drive alignment, and take ownership to deliver outcomes (including production support/on call as needed).
  • Nice to have experience in ETL tools such as DBT, Matillion and data quality/testing frameworks (e.g., Collibra) and data governance tools such as Immuta
  • Nice to have experience with orchestration tools (e.g., Airflow), streaming (Kafka/Kinesis), and modern table formats (Delta/Iceberg/Hudi).
  • Bachelor’s degree in computer science, Engineering, or related field (or equivalent practical experience).
Primary Skills:

Python, PySpark, SQL, AWS, Databricks, GitHub, Data Lake, ETL/ELT and CI/CD


Secondary Skills:

Dimensional modeling, Docker and Terraform


Who we are


We are known as well-known org Inc., Rahway, New Jersey, USA in the United States and Canada and MSD everywhere else. For more than a century, bringing forward medicines and vaccines for many of the worlds most challenging diseases. Today, our company continues to be at the forefront of research to deliver innovative health solutions and advance the prevention and treatment of diseases that threaten people and animals around the world.


What we look for


Imagine getting up in the morning for a job as important as helping to save and improve lives around the world. Here, you have that opportunity. You can put your empathy, creativity, digital mastery, or scientific genius to work in collaboration with a diverse group of colleagues who pursue and bring hope to countless people who are battling some of the most challenging diseases of our time. Our team is constantly evolving, so if you are among the intellectually curious, join us—and start making your impact today.


Required Skills:

Amazon Web Services (AWS), CI/CD, Databricks Platform, Data ETL, Data Lake, GitHub, PySpark, Python (Programming Language), Structured Query Language (SQL)


Preferred Skills:

Dimensional Modeling, Docker (Software), Terraform


Current Employees applyHERE


Current Contingent Workers applyHERE


Secondary Language(s) Job Description:

Search Firm Representatives Please Read Carefully


Merck Co., Inc., Rahway, NJ, USA, also known as Merck Sharp Dohme LLC, Rahway, NJ, USA, does not accept unsolicited assistance from search firms for employment opportunities. All CVs / resumes submitted by search firms to any employee at our company without a valid written search agreement in place for this position will be deemed the sole property of our company. No fee will be paid in the event a candidate is hired by our company as a result of an agency referral where no pre-existing agreement is in place. Where agency agreements are in place, introductions are position specific. Please, no phone calls or emails.


Employee Status:

Regular


Relocation:

Domestic


VISA Sponsorship:

No


Travel Requirements:

No Travel Required


Flexible Work Arrangements:

Hybrid


Shift:

Not Indicated


Valid Driving License:

No


Hazardous Material(s):

n/a


Job Posting End Date:

09/16/2026*A job posting is effective until 11:59:59PM on the dayBEFORE the listed job posting end date. Please ensure you apply to a job posting no later than the dayBEFORE the job posting end date.


Requisition ID:

R414162

Frequently Asked Questions

How to apply for Specialist , Data Engineering at MSD?

Click the "Apply via CareerScan" button on this page.

What is the salary for this role?

Salary details will be discussed during the interview.

What experience is required?

5 years of experience is required.

Is this position still open?

Yes, currently active and accepting applications.

Similar Openings

Explore related active roles in email support

View all
Actively Hiring
6–8 years
Salary not disclosed
Bengaluru, Karnataka
email supportDay Shift
Posted 18h ago
Apply Now
Actively Hiring
0-2 Yrs
Salary not disclosed
Hyderabad, Telangana
email supportDay Shift
Posted 18h ago
Apply Now
Actively Hiring
1 year
Salary not disclosed
Hyderabad, Telangana
email supportDay Shift
Posted 18h ago
Apply Now

Specialist , Data Engineering

MSD · Hyderabad