Verified current Job

Site Reliability Engineer - India

All roles at JumpCloud® are Remote unless otherwise specified in the Job Description.

Job Remote Full source details
Jumpcloud Bangalore, Bangalore, India - Remote Source published Sep 29, 2026 Verified 2 hours ago
✓ 100% verification score · Source: Jumpcloud (lever) · Always confirm final requirements on the original source.
Complete source information imported The available role or programme description, requirements, benefits and source facts were imported from the public official endpoint and formatted for reading.
EmploymentFull Time
Work modeRemote / location-flexible

Overview

All roles at JumpCloud® are Remote unless otherwise specified in the Job Description.

Full job description

All roles at JumpCloud® are Remote unless otherwise specified in the Job Description. About JumpCloud® JumpCloud® is the AI-powered unified IT management platform designed to secure the modern workforce. By consolidating identity, device, and access management, JumpCloud provides intelligent, secure IT that scales from human users to autonomous AI agents. We help organizations around the globe eliminate complexity and turn AI risk into an optimized advantage, ensuring the right people and agents have secure access to the right resources at all times.

JumpCloud is Intelligent, Secure IT.

About the role:

We are seeking a Software Engineer 3 (SRE) to join our Infrastructure & Reliability Engineering team. This role sits at the center of platform resilience—ensuring high availability, performance, recoverability, and operational maturity across JumpCloud’s production systems. This is not a traditional operations role.

Our SREs are engineers first: designing automation, building observability frameworks, defining reliability standards, and reducing operational toil through code. You will build and scale cloud-native infrastructure, participate in incident management, and implement reliability best practices across our directory platform and microservices.

Design, deploy, and maintain the reliability, availability, and performance of critical JumpCloud systems and APIs across AWS and GCP. Operationalize SLIs, SLOs, and error budgets in direct partnership with core application teams. Build and refine end-to-end observability across microservices and cloud infrastructure using tools like Datadog. Implement actionable monitoring across Golden Signals (Latency, Traffic, Errors, Saturation) to optimize detection (MTTD) and minimize alert fatigue. Participate in on-call rotations, incident response, and blameless post-incident reviews to drive continuous systemic improvements. Manage and operationalize production Kubernetes (EKS) clusters utilizing GitOps delivery workflows (Argo CD, Kargo). Provision and secure multi-cloud infrastructure using modular Terraform (Infrastructure-as-Code). Develop and maintain Disaster Recovery (DR) dashboards, runbooks, multi-region failover automation, and validation tests to ensure alignment with defined RTO and RPO targets. Eliminate operational toil by writing production-grade Python or Go scripts and automation tools. Leverage AI-assisted development tools (Cursor, Claude Code, GitHub Copilot) to accelerate scripting, runbook generation, and incident triage.

5+ years of professional software engineering experience in SRE, DevOps, or Platform Engineering operating 24/7 mission-critical systems. Python/Go Proficiency: Hands-on capabilities writing code for SRE tools, custom automation, and cloud integrations. Kubernetes Ecosystem: Production experience with Kubernetes cluster operations, container orchestration, and GitOps pipelines (Argo CD). Infrastructure as Code: Solid experience writing, maintaining, and modularizing Terraform configurations. Cloud Architecture: Direct experience operating cloud workloads on AWS (EKS, IAM, VPC networking, Route53, ALB/NLB) or GCP. FinOps & Cost Visibility: Practical experience setting up cost-allocation tagging, resource right-sizing, and building FinOps dashboards to visualize cloud spend. Disaster Recovery & Monitoring: Experience building DR dashboards, running failover drills, and configuring monitoring tools to track system health and recovery metrics Observability & Incident Management: Practical experience with Datadog (or similar), PagerDuty, alerting hygiene, and working within SLI/SLO frameworks. Solid operational experience configuring and troubleshooting production service meshes (Istio or similar) and managing high-availability proxy solutions (HAProxy, NGINX, or similar). Problem Solving & Mindset: Strong troubleshooting skills, effective collaboration, and a track record of driving operational efficiency through code. A strong team player who helps us live by our core values: building connections, thinking big, and getting 1% better every day.

Experience with CI/CD tools such as GitHub Actions or GitLab Pipelines. Basic understanding of chaos engineering principles or testing resilience in staging/production. Familiarity with secrets management tools (HashiCorp Vault, AWS Secrets Manager, External Secrets Operator). Basic knowledge of DevSecOps tools and scanning/fixing infrastructure-as-code vulnerabilities.

Tips for this job

Practical JobOpportunity guidance. These tips do not replace official rules or create new eligibility requirements.

  1. Tailor the CV and application to the responsibilities and required skills stated on the official employer page.
  2. Use concrete evidence of relevant work, projects and measurable results rather than generic claims.
  3. Confirm location, work authorization, remote restrictions and sponsorship terms before applying.
  4. Apply through the original employer or official recruitment destination shown on this page.

Verification notes

laptop-ats-crawler v3

Original authoritative source

JobOpportunity is the discovery and verification layer. Confirm eligibility, dates, salary/funding and application instructions on the original source before submitting anything.

Apply through JobOpportunity →

Browse current JobOpportunity listings from Jumpcloud (lever) →

More ways to save

Discover deals, coupons and free courses on our sister site.

Explore DealVorio
Save more with DealVorio: deals, coupons, free courses, apps and books