JP Morgan Services India Pvt Ltd

Site Reliability Engineer III

JP Morgan Services India Pvt Ltd
Mumbai/Bombay
Not disclosed
Work from OfficeWork from Office
Full TimeFull Time
Min. 3 yearsMin. 3 years

Job Description

Site Reliability Engineer III

Join us at the center of a rapidly growing technology field, where your work helps modernize complex, mission-critical systems. We’ll value your ideas, support your growth, and empower you to make reliability improvements that matter. Guidelines.docx Raw Posting.docx

Job summary

As a Site Reliability Engineer III at JPMorgan Chase within the Corporate Technology, you solve broad business problems with simple, straightforward solutions while improving the availability, reliability, and scalability of your application or platform. You use code and cloud infrastructure to configure, maintain, monitor, and optimize applications and their associated infrastructure, and you contribute meaningfully by sharing end-to-end operational knowledge across the team. We work collaboratively, communicate clearly during incidents, and focus on iterative improvements that reduce toil and improve outcomes.

Job responsibilities

  • Design appropriate-level reliability designs, guide and assist others, and build consensus with peers while supporting adoption of site reliability engineering best practices within your team.
  • Collaborate with software engineers and partner teams to design, develop, test, and implement deployment and reliability approaches using automated continuous integration and continuous delivery (CI/CD) pipelines.
  • Implement infrastructure, configuration, and network as code for the applications and platforms in your remit, and iteratively improve solutions by decomposing problems into smaller, actionable changes.
  • Operate and optimize applications and their associated infrastructure by configuring, maintaining, and monitoring services to meet availability, reliability, and scalability expectations.
  • Resolve complex problems with technical experts, key stakeholders, and team members by using service level indicators (SLIs) and service level objectives (SLOs) to proactively address issues before they impact customers.
  • Use enterprise-authorized AI capabilities to accelerate incident triage, troubleshooting, and post-incident analysis, validating outputs and handling operational data according to sensitivity and security requirements.
  • Identify patterns in operational signals that indicate reliability risk or recurring toil, prioritize reuse-first improvements tied to SLO outcomes, and recognize roadblocks while exploring new technologies where appropriate.

Required qualifications, capabilities, and skills

  • Formal training or certification on site reliability engineering concepts and 3+ years applied experience (country-specific requirements apply: NAMR/APAC—India/LATAM/Hong Kong; EMEA/LATAM—Brazil; Singapore follows local country guidance).
  • Proficiency in site reliability engineering culture and principles, including how to implement site reliability engineering within an application or platform.
  • Proficiency in at least one programming language such as Python, Java/Spring Boot, and .NET, with experience developing, debugging, and maintaining code in a large corporate environment.
  • Working knowledge of using enterprise-authorized AI capabilities within the work environment to support site reliability engineering workflows, including strong validation habits and awareness of data sensitivity.
  • Ability to validate AI-assisted operational recommendations before applying changes, escalating when uncertain and following data sensitivity requirements.
  • Experience with observability practices such as white-box and black-box monitoring, SLO alerting, and telemetry collection, with familiarity troubleshooting common networking technologies and issues.
  • Experience with continuous integration and continuous delivery tooling, plus familiarity with containers and container orchestration.

Preferred qualifications, capabilities, and skills

  • Experience  in site reliability engineering / production support / DevOps / platform roles with real on-call exposure, including improving reliability through incident response, root-cause analysis (RCA)/postmortems, problem management, and measurable toil reduction (reduced pages, automated repetitive tasks, improved MTTR/MTBF); ability to lead parts of incident calls, write clear RCAs, mentor others, and drive operational standards (runbooks, alerts, SLO definitions).
  • Strong observability experience aligned to your stack, including Dynatrace (APM, alerting, dashboards) and/or Grafana + Prometheus; log analysis in Splunk (queries, dashboards, troubleshooting); ability to correlate metrics, logs, and traces to isolate issues in microservices; plus advanced observability concepts such as OpenTelemetry/distributed tracing, alert-as-code, SLO tooling, synthetic monitoring, and capacity planning using telemetry.
  • Hands-on Kubernetes operations/troubleshooting (deployments, services/ingress, configmaps/secrets, HPA, node/pod debugging) and solid AWS experience supporting containerized workloads (EKS preferred where applicable, plus IAM/VPC basics); experience with CI/CD and infrastructure as code (Terraform modules, state management, pipeline integration; Jenkins/GitLab CI; blue/green and canary deployments) and advanced platform tooling (Helm/Kustomize, service mesh such as Istio/Linkerd, policy such as OPA/Gatekeeper, secrets tools such as Vault, GitOps such as ArgoCD/Flux).
  • Strong automation mindset and microservices reliability depth: Java microservices plus scripting (Python and/or shell) to automate operational tasks (self-healing, runbook automation, unit tests) while following secure coding practices; Linux/Unix fundamentals (process/memory/disk troubleshooting, networking basics, log analysis, performance triage); operational experience supporting Kafka and/or MQ (including IBM MQ), and databases (Oracle and/or MongoDB) with awareness of performance symptoms and connection pools; certificate management in distributed systems (TLS/SSL, rotation/renewal, keystores/truststores, outage prevention); understanding of microservice failure modes (retries/timeouts, circuit breakers, rate limiting, backpressure, dependency mapping); 

Experience Level

Senior Level

Job role

Work location
Work locationMumbai, Maharashtra, India
Department
DepartmentIT & Information Security
Role / Category
Role / CategoryIT Security
Employment type
Employment typeFull Time
Shift
ShiftDay Shift

Job requirements

Experience
ExperienceMin. 3 years

About company

Name
NameJP Morgan Services India Pvt Ltd
Job posted by JP Morgan Services India Pvt Ltd

Similar jobs you can apply for

Accounts / Finance

Lighting Automation Engineer

Green World Technology
Girgaon, Mumbai/Bombay
₹25,000 - ₹35,000
Work from Office
Full Time
Night Shift
Min. 2 years
Good (Intermediate / Advanced) English
Cinepolis India Private Limited

Counter Staff

Cinepolis India Private Limited
Bhandup West, Mumbai/Bombay
₹14,000 - ₹19,000*
Work from Office
Full Time
Night Shift
Any experience
No English Required
Navion Electronics Private Ltd

Purchase Executive

Navion Electronics Private Ltd
Mahalaxmi Nagar, Mumbai/Bombay
₹18,000 - ₹20,000
Work from Office
Full Time
Min. 6 months
Good (Intermediate / Advanced) English
Deepak & Sahil Engcon Private Limited

Project Manager

Deepak & Sahil Engcon Private Limited
Pali Hills, Mumbai/Bombay
₹80,000 - ₹90,000
Field Job
Full Time
Min. 3 years
Good (Intermediate / Advanced) English
Teamspace Financial Services Private Limited

CA Firm

Teamspace Financial Services Private Limited
Wadala, Mumbai/Bombay
₹20,000 - ₹40,000
Work from Office
Full Time
Min. 2 years
Good (Intermediate / Advanced) English

Sous Chef

Deliure
Byculla, Mumbai/Bombay
₹35,000 - ₹45,000
Work from Office
Full Time
Night Shift
Min. 5 years
Basic English

You can expect a minimum salary of 0 INR. The salary offered will depend on your skills, experience and performance in the interview.

The candidate should have completed the required education and people who have 3 to 31 years are eligible to apply for this job. You can apply for more jobs in Mumbai/Bombay to get hired quickly.

The candidate should have sound communication skills and sound communication skills for this job.

Both Male and Female candidates can apply for this job.

No, it's not a work from home job and can't be done online. You can explore and apply for other work from home jobs in Mumbai/Bombay at apna.

No work-related deposit needs to be made during your employment with the company.

Go to the apna app and apply for this job. Click on the apply button and call HR directly to schedule your interview.

The last date to apply for this job is . For more details, download apna app and find Full Time jobs in Mumbai/Bombay . Through apna, you can find jobs in 64 cities across India. Join NOW!