JobTarget Logo

Platform Engineer in India at Jobgether

NewJob Function: Engineering
Jobgether
India, India
Posted on
New job! Apply early to increase your chances of getting hired.

Explore Related Opportunities

Job Description

Platform Engineer

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Platform Engineer based in India.

Join a technology-driven environment where platform reliability, automation, and scalable infrastructure are central to delivering high-quality digital solutions. In this role, you will help maintain stable and resilient systems while responding proactively to incidents and operational challenges. You will work across Kubernetes, Docker, Linux, cloud infrastructure, CI/CD, and backend technologies. The position combines hands-on troubleshooting with continuous improvements to monitoring, deployment, and support processes. You will collaborate with technical teams and stakeholders to meet demanding service levels and minimize operational dependencies. This is an opportunity to deepen your platform engineering expertise while contributing to highly automated, production-grade environments.

Accountabilities
  • Monitor production environments, system performance, alerts, and operational metrics to maintain stability and meet defined service levels.
  • Resolve Tier-1 incidents using established runbooks, including system restarts, network connectivity checks, model resets, data-flow checks, alert reprocessing, and other standard fixes.
  • Continuously monitor PagerDuty alerts and respond to incidents according to documented procedures, escalating unresolved issues to L2/L3 teams when required.
  • Perform regular server and infrastructure checks, including Redis cron jobs, security systems, real-time overlays, and other operational components.
  • Review root-cause analyses and resolution requests, document incident outcomes, and ensure appropriate follow-up actions are completed.
  • Create, maintain, and execute runbooks for recurring incidents while identifying opportunities to reduce manual intervention and move systems toward greater automation.
  • Support SLA management by prioritizing incidents according to severity and ensuring timely response and resolution.
  • Contribute to Kubernetes platform engineering by building and maintaining scalable container infrastructure and improving deployment reliability.
  • Enhance CI/CD, monitoring, alerting, health checks, deployment strategies, and overall platform observability.
  • Troubleshoot Kubernetes, Docker, Linux, networking, and distributed-system issues across production environments.
  • Develop and maintain production-grade Python applications and services, including web applications, database integrations, and streaming pipelines.
  • Support technologies such as Flask, Gunicorn, Redis, Kafka, Docker, Kubernetes, and Linux shell environments.
  • Provide regular daily and weekly reporting on incidents, system performance, service levels, and key operational metrics.
  • Coordinate planned system downtime, upgrades, and maintenance activities while minimizing disruption to operations.
  • Maintain clear and consistent communication with internal teams and other stakeholders throughout incident resolution and operational activities.
Requirements
  • 1–6 years of relevant experience in platform engineering, DevOps, site reliability, backend engineering, cloud operations, or a related technical field.
  • Strong expertise in Linux, Docker, Kubernetes, container infrastructure, and production system troubleshooting.
  • Hands-on experience with Kubernetes resources including Deployments, Pods, Jobs, StatefulSets, ConfigMaps, Services, NodePort, Ingress, Volumes, and Custom Resource Definitions.
  • Experience creating and maintaining scalable Kubernetes architectures and using Helm Charts to template deployments.
  • Knowledge of multi-container pod architectures, including sidecar and init containers, as well as probes and health checks.
  • Strong understanding of CI/CD implementation and deployment strategies such as Blue/Green, Canary, and rolling deployments.
  • Experience using container logs, monitoring tools, and alerting systems to identify and resolve production issues.
  • Familiarity with hybrid-cloud environments is an advantage.
  • Strong Python programming skills, including iterators, exception handling, file handling, data structures, object-oriented programming, and software design patterns.
  • Experience developing production-grade Python applications rather than scripts, with knowledge of Flask and Gunicorn.
  • Understanding of database integrations and streaming technologies such as Redis and Kafka.
  • Experience with multiprocessing architectures and strong Git/GitHub knowledge.
  • Familiarity with video streaming and image-processing technologies such as GStreamer, FFmpeg, and OpenCV is highly desirable.
  • Knowledge of cloud services and Linux shell scripting.
  • Kubernetes certification such as CKAD is an advantage.
  • Strong analytical and problem-solving skills, with the ability to troubleshoot incidents methodically and work effectively under operational pressure.
  • Strong communication and collaboration skills, with a proactive approach to incident management and stakeholder coordination.
  • Willingness to work rotational shifts, including overnight coverage, and to join immediately.
Benefits
  • Fully remote working arrangement from India, with base locations associated with Mumbai, Bengaluru, or Trivandrum.
  • Rotational shift schedule designed to provide continuous operational coverage:
    • Shift A: 6:00 AM–2:00 PM IST
    • Shift B: 2:00 PM–10:00 PM IST
    • Shift C: 10:00 PM–6:00 AM IST
  • Two consecutive days off each week.
  • Opportunity to work with modern cloud, Kubernetes, DevOps, automation, and backend technologies.
  • Exposure to production-scale systems, distributed architectures, and highly automated operational environments.
  • Opportunities for continuous technical learning and professional development.
  • Collaborative, diverse, and growth-oriented work culture.
  • Opportunity to develop expertise across platform engineering, DevOps, backend development, and cloud technologies.
How Jobgether works:
We use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team.
We appreciate your interest and wish you the best!
Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time.
#LI-CL1

Job Location

India, India

Frequently asked questions about this position

Continue to apply
Enter your email to continue. You’ll be redirected to the employer’s application.
By clicking Continue, you understand and agree to JobTarget's Terms of Use and Privacy Policy.