MSD

Associate Specialist, Data Engineering

Hyderabad, Telangana
3 years exp
Day Shift
Posted 21h ago
0 views
Actively Hiring Direct 1-Click Apply

Check Your Resume Match Score

Scan your resume against ATS criteria for this Associate Specialist, Data Engineering role at MSD.

Apply for this position

Apply on Company Website

Job Description

Job Description


Associate Specialist:

Data Engineering


The Opportunity:

Join a global biopharma company with a 130-year legacy and mission to achieve new milestones in healthcare. Be part of a technology-driven, data-led organization supporting a diversified portfolio of medicines, vaccines, and animal health products. Work alongside passionate teams that use data, analytics, and insights to drive decisions and tackle some of the world’s greatest health threats.


Our Technology Centers are globally distributed hubs that enable our digital transformation and business outcomes across IT. They bring together diverse teams to collaborate, share best practices, and deliver solutions that save and improve lives.


This role is based at our Hyderabad Tech Center and follows a hybrid working model (3 days onsite, 2 days remote). Candidates are expected to reside within commuting distance of the Hyderabad office.


Role Overview


We are looking for aData Engineer with2–4 years of hands-on experience in building and supporting data pipelines, ETL/ELT workflows, and analytics-ready datasets. The ideal candidate should have strong fundamentals inPython,PySpark,SQL,AWS, andDatabricks, with practical exposure to data lakes, lakehouse patterns, data warehousing, data quality, and production support. This role is best suited for a hands-on engineer who can work from defined requirements, contribute to reliable data solutions, collaborate with cross-functional teams, and grow into larger ownership over time.


What will you do in this role


  • Build, enhance, and supportbatch and streaming data pipelines using defined technical designs and backlog requirements.
  • Develop and maintainETL/ELT transformations usingPython,PySpark, and SQL across data lake, lakehouse, and warehouse environments.
  • Work closely with Data Analysts, Data Scientists, senior engineers, tech leads, and product managers to understand requirements and delivercurated, analytics-ready datasets.
  • Implementdata quality checks, validations, reconciliations, and basic anomaly checks to improve trust and usability of data outputs.
  • Run, monitor, and troubleshoot pipelines using orchestration and observability tools such asDatabricks Workflows,AWS Step Functions, scheduling, logging, monitoring, and alerting.
  • Follow engineering practices includingunit testing,integration testing, automateddata tests, code reviews, and quality gates withinCI/CD.
  • Support BI and analytics use cases by applyingdimensional modeling concepts such as facts, dimensions, star/snowflake schemas, andslowly changing dimensions (SCD).
  • Write and tuneSQL queries for data profiling, transformations, validations, debugging, and performance improvements.
  • UseAWS services such asS3,Glue,Lambda,Step Functions,EMR, and CloudWatch to support data engineering workloads while following security practices such as IAM, encryption, and least privilege.
  • Contribute to cloud resource provisioning and environment configuration usingTerraform, with guidance from senior engineers.
  • Package, deploy, and support workloads usingDocker and related runtime configurations, including ECS/Fargate where applicable.
  • UseGitHub for version control, branching, pull requests, code reviews, and contribution toCI/CD pipelines.
  • Develop scalable data processing logic onDatabricks /Apache Spark usingPySpark and lakehouse concepts such asDelta Lake, ACID transactions, and schema evolution.
  • Use Jupyter/Databricks notebooks for exploration, debugging, and PoCs; convert validated logic into reusable modules, tests, and deployment-ready pipelines.
  • Participate inAgile delivery ceremonies, provide task-level estimates, share progress updates, and raise risks or dependencies early.
  • Create and maintain technical documentation such as pipeline specifications, data contracts, runbooks, and support notes.
What Should you have:
  • Bachelor’s degree in computer science, Engineering, or a related field, or equivalent practical experience.
  • 2–4 years of hands-on experience in data engineering, including building or supporting production data pipelines and ETL/ELT workflows.
  • Practical experience withAWS services such asS3,Glue,Lambda,Step Functions,EMR, and CloudWatch; understanding of IAM, encryption, and cloud security basics.
  • Hands-on experience withDatabricks,Apache Spark,PySpark, and lakehouse concepts such asDelta Lake.
  • StrongSQL skills for joins, window functions, data profiling, transformations, validations, and performance tuning.
  • Good working knowledge ofPython andPySpark, including Spark fundamentals such as partitioning, shuffle, caching, file formats, debugging, and optimization.
  • Understanding ofdimensional modeling concepts including facts, dimensions, star/snowflake schemas, andslowly changing dimensions (SCD).
  • Exposure toGitHub,CI/CD, code reviews, branching, release practices, and engineering quality standards.
  • Working knowledge ofDocker andTerraform for deployment, runtime configuration, and cloud environment support.
  • Ability to work in Agile teams, communicate clearly, collaborate with cross-functional stakeholders, and take ownership of assigned deliverables.
Primary Skills:

Python, PySpark, SQL, AWS, Databricks, GitHub, Data Lake, ETL/ELT and CI/CD


Secondary Skills:

Dimensional modeling, Docker and Terraform


Good to Have


  • Exposure to data quality or testing frameworks such as Collibra, Immuta, and basic awareness of data governance practices including catalog, lineage, and access controls.
  • Exposure to orchestration tools such as Airflow or modern table formats such as Delta, Iceberg.
  • AWS certification such as Developer or Solutions Architect, or equivalent demonstrated cloud experience.

Who we are


We are known as well-known org Inc., Rahway, New Jersey, USA in the United States and Canada and MSD everywhere else. For more than a century, bringing forward medicines and vaccines for many of the worlds most challenging diseases. Today, our company continues to be at the forefront of research to deliver innovative health solutions and advance the prevention and treatment of diseases that threaten people and animals around the world.


What we look for


Imagine getting up in the morning for a job as important as helping to save and improve lives around the world. Here, you have that opportunity. You can put your empathy, creativity, digital mastery, or scientific genius to work in collaboration with a diverse group of colleagues who pursue and bring hope to countless people who are battling some of the most challenging diseases of our time. Our team is constantly evolving, so if you are among the intellectually curious, join us—and start making your impact today.


Required Skills:

Amazon Web Services (AWS), CI/CD, Databricks Platform, Data Lake, ETL Development, GitHub, PySpark, Python (Programming Language), Structured Query Language (SQL)


Preferred Skills:

Dimensional Modeling, Docker (Software), Terraform


Current Employees applyHERE


Current Contingent Workers applyHERE


Secondary Language(s) Job Description:

Search Firm Representatives Please Read Carefully


Merck Co., Inc., Rahway, NJ, USA, also known as Merck Sharp Dohme LLC, Rahway, NJ, USA, does not accept unsolicited assistance from search firms for employment opportunities. All CVs / resumes submitted by search firms to any employee at our company without a valid written search agreement in place for this position will be deemed the sole property of our company. No fee will be paid in the event a candidate is hired by our company as a result of an agency referral where no pre-existing agreement is in place. Where agency agreements are in place, introductions are position specific. Please, no phone calls or emails.


Employee Status:

Regular


Relocation:

Domestic


VISA Sponsorship:

No


Travel Requirements:

No Travel Required


Flexible Work Arrangements:

Hybrid


Shift:

Not Indicated


Valid Driving License:

No


Hazardous Material(s):

n/a


Job Posting End Date:

09/16/2026*A job posting is effective until 11:59:59PM on the dayBEFORE the listed job posting end date. Please ensure you apply to a job posting no later than the dayBEFORE the job posting end date.


Requisition ID:

R414160

Frequently Asked Questions

How to apply for Associate Specialist, Data Engineering at MSD?

Click the "Apply via CareerScan" button on this page.

What is the salary for this role?

Salary details will be discussed during the interview.

What experience is required?

3 years of experience is required.

Is this position still open?

Yes, currently active and accepting applications.

Similar Openings

Explore related active roles in email support

View all
Actively Hiring
6–8 years
Salary not disclosed
Bengaluru, Karnataka
email supportDay Shift
Posted 21h ago
Apply Now
Actively Hiring
0-2 Yrs
Salary not disclosed
Hyderabad, Telangana
email supportDay Shift
Posted 21h ago
Apply Now
Actively Hiring
1 year
Salary not disclosed
Hyderabad, Telangana
email supportDay Shift
Posted 21h ago
Apply Now

Associate Specialist, Data Engineering

MSD · Hyderabad