jobgether

Platform Engineer

India
1+ years exp
Full-time
Posted 4d ago
1 views
Actively Hiring Urgent Opening Direct 1-Click Apply

Check Your Resume Match Score

Scan your resume against ATS criteria for this Platform Engineer role at jobgether.

Apply for this position

Apply on Company Website

Job Description

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Platform Engineer based in India.

Join a technology-driven environment where platform reliability, automation, and scalable infrastructure are central to delivering high-quality digital solutions. In this role, you will help maintain stable and resilient systems while responding proactively to incidents and operational challenges. You will work across Kubernetes, Docker, Linux, cloud infrastructure, CI/CD, and backend technologies. The position combines hands-on troubleshooting with continuous improvements to monitoring, deployment, and support processes. You will collaborate with technical teams and stakeholders to meet demanding service levels and minimize operational dependencies. This is an opportunity to deepen your platform engineering expertise while contributing to highly automated, production-grade environments.

Accountabilities

  • Monitor production environments, system performance, alerts, and operational metrics to maintain stability and meet defined service levels.

  • Resolve Tier-1 incidents using established runbooks, including system restarts, network connectivity checks, model resets, data-flow checks, alert reprocessing, and other standard fixes.

  • Continuously monitor PagerDuty alerts and respond to incidents according to documented procedures, escalating unresolved issues to L2/L3 teams when required.

  • Perform regular server and infrastructure checks, including Redis cron jobs, security systems, real-time overlays, and other operational components.

  • Review root-cause analyses and resolution requests, document incident outcomes, and ensure appropriate follow-up actions are completed.

  • Create, maintain, and execute runbooks for recurring incidents while identifying opportunities to reduce manual intervention and move systems toward greater automation.

  • Support SLA management by prioritizing incidents according to severity and ensuring timely response and resolution.

  • Contribute to Kubernetes platform engineering by building and maintaining scalable container infrastructure and improving deployment reliability.

  • Enhance CI/CD, monitoring, alerting, health checks, deployment strategies, and overall platform observability.

  • Troubleshoot Kubernetes, Docker, Linux, networking, and distributed-system issues across production environments.

  • Develop and maintain production-grade Python applications and services, including web applications, database integrations, and streaming pipelines.

  • Support technologies such as Flask, Gunicorn, Redis, Kafka, Docker, Kubernetes, and Linux shell environments.

  • Provide regular daily and weekly reporting on incidents, system performance, service levels, and key operational metrics.

  • Coordinate planned system downtime, upgrades, and maintenance activities while minimizing disruption to operations.

  • Maintain clear and consistent communication with internal teams and other stakeholders throughout incident resolution and operational activities.

    Requirements

    • 1–6 years of relevant experience in platform engineering, DevOps, site reliability, backend engineering, cloud operations, or a related technical field.

    • Strong expertise in Linux, Docker, Kubernetes, container infrastructure, and production system troubleshooting.

    • Hands-on experience with Kubernetes resources including Deployments, Pods, Jobs, StatefulSets, ConfigMaps, Services, NodePort, Ingress, Volumes, and Custom Resource Definitions.

    • Experience creating and maintaining scalable Kubernetes architectures and using Helm Charts to template deployments.

    • Knowledge of multi-container pod architectures, including sidecar and init containers, as well as probes and health checks.

    • Strong understanding of CI/CD implementation and deployment strategies such as Blue/Green, Canary, and rolling deployments.

    • Experience using container logs, monitoring tools, and alerting systems to identify and resolve production issues.

    • Familiarity with hybrid-cloud environments is an advantage.

    • Strong Python programming skills, including iterators, exception handling, file handling, data structures, object-oriented programming, and software design patterns.

    • Experience developing production-grade Python applications rather than scripts, with knowledge of Flask and Gunicorn.

    • Understanding of database integrations and streaming technologies such as Redis and Kafka.

    • Experience with multiprocessing architectures and strong Git/GitHub knowledge.

    • Familiarity with video streaming and image-processing technologies such as GStreamer, FFmpeg, and OpenCV is highly desirable.

    • Knowledge of cloud services and Linux shell scripting.

    • Kubernetes certification such as CKAD is an advantage.

    • Strong analytical and problem-solving skills, with the ability to troubleshoot incidents methodically and work effectively under operational pressure.

    • Strong communication and collaboration skills, with a proactive approach to incident management and stakeholder coordination.

    • Willingness to work rotational shifts, including overnight coverage, and to join immediately.

      Benefits

      • Fully remote working arrangement from India, with base locations associated with Mumbai, Bengaluru, or Trivandrum.

      • Rotational shift schedule designed to provide continuous operational coverage:

        • Shift A: 6:00 AM–2:00 PM IST
        • Shift B: 2:00 PM–10:00 PM IST
        • Shift C: 10:00 PM–6:00 AM IST
        • Two consecutive days off each week.
        • Opportunity to work with modern cloud, Kubernetes, DevOps, automation, and backend technologies.
        • Exposure to production-scale systems, distributed architectures, and highly automated operational environments.
        • Opportunities for continuous technical learning and professional development.
        • Collaborative, diverse, and growth-oriented work culture.
        • Opportunity to develop expertise across platform engineering, DevOps, backend development, and cloud technologies.
  • How Jobgether works:

We use an

AI-powered matching process

to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team.

We appreciate your interest and wish you the best!

Why Apply Through Jobgether?

  • Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time.

#LI-CL1

Key Requirements & Skills

  • 1-6 years of relevant experience in platform engineering, DevOps, site reliability, backend engineering, cloud operations, or a related field
  • Strong expertise in Linux, Docker, Kubernetes, container infrastructure, and production system troubleshooting
  • Hands-on experience with Kubernetes resources: Deployments, Pods, Jobs, StatefulSets, ConfigMaps, Services, NodePort, Ingress, Volumes, CRDs
  • Experience creating scalable Kubernetes architectures and using Helm Charts to template deployments
  • Knowledge of multi-container pod architectures (sidecar, init containers), probes, and health checks
  • Strong understanding of CI/CD implementation and deployment strategies (Blue/Green, Canary, rolling)
  • Experience using container logs, monitoring tools, and alerting systems to resolve production issues
  • Familiarity with hybrid-cloud environments
  • Strong Python programming skills (iterators, exception handling, file handling, data structures, OOP, design patterns)
  • Experience developing production-grade Python applications with knowledge of Flask and Gunicorn
  • Understanding of database integrations and streaming technologies such as Redis and Kafka
  • Experience with multiprocessing architectures and strong Git/GitHub knowledge
  • Familiarity with video streaming and image-processing technologies such as GStreamer, FFmpeg, and OpenCV
  • Knowledge of cloud services and Linux shell scripting
  • Kubernetes certification such as CKAD
  • Strong analytical and problem-solving skills; methodical troubleshooting under operational pressure
  • Strong communication and collaboration skills with a proactive approach to incident management
  • Willingness to work rotational shifts including overnight coverage and to join immediately

Benefits & Perks

Benefits

- Fully remote working arrangement from India, with base locations associated with Mumbai, Bengaluru, or Trivandrum.
- Rotational shift schedule designed to provide continuous operational coverage:
  
  - Shift A: 6:00 AM–2:00 PM IST
  - Shift B: 2:00 PM–10:00 PM IST
  - Shift C: 10:00 PM–6:00 AM IST
  - Two consecutive days off each week.
  - Opportunity to work with modern cloud, Kubernetes, DevOps, automation, and backend technologies.
  - Exp

Frequently Asked Questions

How to apply for Platform Engineer at jobgether?

Click the "Apply via CareerScan" button on this page.

What is the salary for this role?

Salary details will be discussed during the interview.

What experience is required?

1+ years of experience is required.

Is this position still open?

Yes, currently active and accepting applications.

Platform Engineer

jobgether · India