Scan your resume against ATS criteria for this Spcialist , Data Engineering role at Merck.
Join a global biopharma company with a 130-year legacy and mission to achieve new milestones in healthcare. Be part of a technology-driven, data-led organization supporting a diversified portfolio of medicines, vaccines, and animal health products. Work alongside passionate teams that use data, analytics, and insights to drive decisions and tackle some of the world’s greatest health threats.
Our Technology Centers are globally distributed hubs that enable our digital transformation and business outcomes across IT. They bring together diverse teams to collaborate, share best practices, and deliver solutions that save and improve lives.
This role is based at our Hyderabad Tech Center and follows a hybrid working model (3 days, 2 days remote). Candidates are expected to reside within commuting distance of the Hyderabad office.
We are hiring a
who can design, build, and operate production-grade data platforms and pipelines end to end. You will deliver reliable, governed, secure, and analytics-ready data by implementing modern
and
patterns on
and
, with strong focus on
,
, and scalable
. This role partners closely with analytics, data science, and business stakeholders to translate requirements into robust datasets, while applying engineering best practices such as testing, code reviews, CI/CD, and observability.
data pipelines to ingest data from multiple sources into an
and
.
transformations using
,
, and SQL; optimize jobs for performance, cost, and reliability.
and features.
controls (validation rules, reconciliation, anomaly checks), define
, and contribute to
, and
practices.
,
, scheduling, logging, monitoring, alerting).
, automated
, code reviews, and quality gates within
.
(star/snowflake), facts & dimensions, and
.
for profiling, transformations, and performance troubleshooting across large datasets.
using services such as
,
,
,
,
, and CloudWatch; follow security best practices (IAM, encryption, least privilege).
(e.g.,
) across dev/test/prod environments.
(and where applicable ECS/Fargate); manage dependencies and runtime configurations.
for version control (branching strategies, pull requests, code reviews) and set up
for automated build, test, and deployment.
/
using
and lakehouse concepts (e.g.,
, ACID, schema evolution).
delivery model (planning, daily sync, reviews, retros), providing accurate estimates and proactively managing risks/dependencies.
of hands-on experience in data engineering building production pipelines and data platforms.
experience:
,
,
,
,
teams; strong collaboration, communication, and stakeholder management skills.
and lakehouse capabilities (e.g., Delta Lake, job/workflow orchestration, cluster tuning) is strongly preferred.
skills including complex joins/window functions, data profiling, and performance tuning; understanding of
concepts.
and
with solid Spark fundamentals (partitioning, shuffle, caching, file formats) and ability to debug/optimize.
,
concepts, and engineering practices (code reviews, branching, release management); working knowledge of
and
.
experience with orchestration tools (e.g., Airflow), streaming (Kafka/Kinesis), and modern table formats (Delta/Iceberg/Hudi).
Python, PySpark, SQL, AWS, Databricks, GitHub, Data Lake, ETL/ELT and CI/CD
Dimensional modeling, Docker and Terraform
We are known as well-known org Inc., Rahway, New Jersey, USA in the United States and Canada and MSD everywhere else. For more than a century, bringing forward medicines and vaccines for many of the world's most challenging diseases. Today, our company continues to be at the forefront of research to deliver innovative health solutions and advance the prevention and treatment of diseases that threaten people and animals around the world.
Imagine getting up in the morning for a job as important as helping to save and improve lives around the world. Here, you have that opportunity. You can put your empathy, creativity, digital mastery, or scientific genius to work in collaboration with a diverse group of colleagues who pursue and bring hope to countless people who are battling some of the most challenging diseases of our time. Our team is constantly evolving, so if you are among the intellectually curious, join us—and start making your impact today.
Amazon Web Services (AWS), CI/CD, Data ETL, Data Lake, GitHub, PySpark, Python (Programming Language), Structured Query Language (SQL)
Dimensional Modeling, Docker (Software), Terraform
Current Employees apply HERE
Current Contingent Workers apply HERE
Merck & Co., Inc., Rahway, NJ, USA, also known as Merck Sharp & Dohme LLC, Rahway, NJ, USA, does not accept unsolicited assistance from search firms for employment opportunities. All CVs / resumes submitted by search firms to any employee at our company without a valid written search agreement in place for this position will be deemed the sole property of our company. No fee will be paid in the event a candidate is hired by our company as a result of an agency referral where no pre-existing agreement is in place. Where agency agreements are in place, introductions are position specific. Please, no phone calls or emails.
Regular
Domestic
No
No Travel Required
Hybrid
Not Indicated
No
n/a
09/16/2026
*A job posting is effective until 11:59:59PM on the day BEFORE the listed job posting end date. Please ensure you apply to a job posting no later than the day BEFORE the job posting end date.
vision and manage cloud resources using
(e.g.,
) across dev/test/prod environments.
How to apply for Spcialist , Data Engineering at Merck?
Click the "Apply on Company Website" button on this page to submit your application directly on the employer's official portal.
What is the salary for this role?
Salary details will be discussed during the interview.
What experience is required?
5+ years of experience is required.
Is this position still open?
Yes, currently active and accepting applications.
Explore related active roles in data_engineering
Spcialist , Data Engineering
Merck · IND - Telangana - Hyderabad (Hitec City Raidurg)