CareerScanCareerScan
JobsCompanies
BlogContact
For Employers
Sign InRegister Free
CareerScanCareerScan

India's verified job platform connecting candidates directly with employers. 100% free applications with instant ATS resume scoring.

Chennai, Bengaluru & Hyderabad
Jobs by location
Jobs in ChennaiJobs in BengaluruJobs in HyderabadJobs in PuneJobs in Mumbai
Popular roles
AR Caller JobsHealthcare Medical CodingReact / Full Stack DeveloperData & Power BI AnalystCustomer Support Executive
Top companies
TCS CareersCognizant JobsInfosys OpeningsApollo HospitalsOmega Healthcare
Career services
Free ATS Resume CheckerAI Resume Builder (Free)AI Job MatcherSalary Guide & BenchmarksJob Alerts on WhatsApp
© 2026 CareerScan India. All rights reserved.256-bit SSL encrypted & verified
Back to all jobs
  1. Home
  2. Jobs
  3. Director, Site Reliability Engineering
jobgether
jobgether

Director, Site Reliability Engineering

US
₹16.9L/mo
10+ years exp
Full-time
Posted 3d ago
0 views
Actively Hiring Urgent Opening Direct 1-Click Apply

Check Your Resume Match Score

Scan your resume against ATS criteria for this Director, Site Reliability Engineering role at jobgether.

Apply for this position

Apply on Company Website
Notice a broken link or wrong info?

Job Description

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Director, Site Reliability Engineering based in United States.

As Director, Site Reliability Engineering, you will lead the teams and technical strategy responsible for keeping critical infrastructure reliable, scalable, and secure for millions of users. You will work across software, systems, automation, cloud infrastructure, and operational processes to solve complex reliability challenges at scale. The role combines strategic leadership with hands-on technical depth, including troubleshooting production systems and partnering closely with software engineering teams. You will help shape the future architecture and deployment practices of a large-scale, privacy-focused technology environment. You will also drive improvements in automation, observability, incident response, and engineering efficiency. This is a remote-first leadership opportunity with significant ownership, autonomy, and impact.

Accountabilities:

Lead and develop Site Reliability Engineering teams responsible for the reliability, scalability, performance, and operational health of large-scale systems.

Define and execute the technical direction for infrastructure, deployment, reliability engineering, automation, and operational practices.

Lead high-impact and complex initiatives from initial proposal and planning through implementation, measurement, and postmortem.

Investigate and resolve sources of instability across high-traffic, distributed systems, identifying root causes and implementing sustainable remediation.

Establish and improve tools, services, monitoring, alerts, incident-response processes, and operational practices that identify and mitigate reliability risks.

Partner closely with software engineers to troubleshoot production issues, evaluate performance considerations, and implement appropriate code-level or infrastructure-level solutions.

Drive automation for infrastructure provisioning and configuration management to improve efficiency, scalability, consistency, and reliability.

Leverage cloud-native architectures and services to strengthen system resilience and support continued growth.

Help ensure products and infrastructure meet established reliability standards while minimizing user impact during failures and incidents.

Identify emerging technical needs and opportunities to guide the long-term evolution of deployment and infrastructure architecture.

Support a culture of ownership, continuous improvement, measurable outcomes, and effective post-incident learning.

Requirements:

10+ years of relevant professional experience in Site Reliability Engineering, platform engineering, infrastructure engineering, software engineering, or related fields.

4+ years of experience leading SRE or comparable engineering teams.

Experience participating in or managing 24/7 on-call operations for large-scale production environments.

Advanced programming experience and the ability to read, write, troubleshoot, and deploy software across production systems.

Strong experience with Linux administration and troubleshooting, web technologies, distributed systems, and high-traffic production environments.

Demonstrated ability to lead complex technical projects from ambiguous initial requirements through execution and postmortem.

Experience developing effective reliability tooling, services, monitoring, alerting, and incident-response capabilities.

Strong investigative and root-cause analysis skills, particularly within distributed and high-scale systems.

Experience designing and implementing infrastructure automation, provisioning, and configuration-management solutions.

Hands-on experience with cloud-native services and architectures, including application packaging and deployment using Docker and Docker Compose.

Experience with high-level programming languages such as Go, Perl, TypeScript, Python, or comparable technologies.

Experience with AI-driven software development, including the design and implementation of agentic workflows.

Strong ability to turn ambiguous or complex problems into practical, innovative solutions with measurable outcomes.

Strategic thinking and technical foresight, with the ability to anticipate future infrastructure and reliability requirements.

Excellent communication and collaboration skills, with the ability to work effectively across engineering teams and technical disciplines.

Strong sense of ownership, autonomy, and accountability in a remote-first working environment.

Benefits:

Annual compensation of

$243,800 USD

, plus stock options.

Transparent compensation structure, with team members at the same professional level and within the same global region receiving the same compensation regardless of functional team, location, gender, educational background, or years of experience.

Fully remote, flexible working arrangement with no core working hours.

Average full-time commitment of approximately 40 hours per week.

Company-sponsored health benefits for eligible team members based in the United States; these benefits do not extend to team members based in Canada or other countries.

Paid parental leave.

Support for home-office setup.

Co-working allowances.

Opportunities to participate in company-wide and team gatherings, with travel expected at least twice per year for an all-hands meeting and a team retreat.

Remote-first environment centered on trust, inclusivity, ownership, and empowered project management.

Equal employment opportunities and a commitment to an accessible, inclusive hiring process.

Reasonable accommodations are available for candidates who require support during the application process.

Successful candidates must complete a background check as a condition of employment.

The role requires participation in video meetings with cameras enabled.

  • How Jobgether works:
    We use an

AI-powered matching process

to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team.
We appreciate your interest and wish you the best!
https://jobgether.com/how-jobgether-works  Why Apply Through Jobgether?   

  • Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time.

 
 
#LI-CL1

Benefits & Perks

Benefits:

Annual compensation of

$243,800 USD

, plus stock options.

Frequently Asked Questions

How to apply for Director, Site Reliability Engineering at jobgether?

Click the "Apply on Company Website" button on this page to submit your application directly on the employer's official portal.

What is the salary for this role?

The salary for this role is $243,800 per annum.

What experience is required?

10+ years of experience is required.

Is this position still open?

Yes, currently active and accepting applications.

ApplicationActively Hiring
Apply on Company Website
Broken link or expired?
jobgether

jobgether

Looking for a work-from-home job? Get matched with the best flexible and remote jobs in the world with Jobgether.

Visit Company Website

More jobs at jobgether

Sr. Sales Specialist, Govt. HighQ & Partnership Sales

US

Sr. Symitar Systems Analyst

US

Staff Backend Engineer

Canada

Share this Opening

Job Alerts for devops

Receive email alerts whenever new devops roles in US are posted.

Set Free Alert →

Similar Openings

Explore related active roles in devops

View all
Actively Hiring
Finc
Platform Engineer
Finc Verified
7+ years
Salary not disclosed
Atlanta
devopsFull-time
Posted 1d ago
Apply Now
Actively Hiring
Supabase
AI Platform Engineer
Supabase Verified
0-2 Yrs
₹7/mo
Remote, Global
devopsFull-timeRemote
Posted 1d ago
Apply Now
UrgentActively Hiring
Eurofins
IT Infrastructure Service Desk Agent
Eurofins Verified
3+ years
Salary not disclosed
Coimbatore, TN, in
devopsFull-time
Posted 1d ago
Apply Now

Director, Site Reliability Engineering

jobgether · US

Apply on Company Website