JPMorganChase Logo

JPMorganChase

Site Reliability Engineer III - Python, Grafana, Splunk, AWS, Jenkins

Posted 3 Days Ago
Be an Early Applicant
Hybrid
Hyderabad, Telangana, IND
Mid level
Hybrid
Hyderabad, Telangana, IND
Mid level
Builds and improves reliable, scalable applications and cloud infrastructure through automation, monitoring, infrastructure as code, and CI/CD. The role configures systems, implements observability and SLO-based alerting, investigates incidents, applies authorized AI capabilities to operational workflows, and collaborates with engineering teams to resolve issues and reduce recurring toil. It also guides peers in adopting SRE practices and iteratively improving application availability, reliability, and scalability.
The summary above was generated by AI

There’s nothing more exciting than being at the center of a rapidly growing field in technology and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. 
As a Site Reliability Engineer III at JPMorgan Chase within the Employee Platforms Team, you will solve complex and broad business problems with simple and straightforward solutions. Through code and cloud infrastructure, you will configure, maintain, monitor, and optimize applications and their associated infrastructure to independently decompose and iteratively improve on existing solutions. You are a significant contributor to your team by sharing your knowledge of end-to-end operations, availability, reliability, and scalability of your application or platform. 
Job responsibilities

  • Guides and assists others in the areas of building appropriate level designs and gaining consensus from peers where appropriate, supporting adoption of site reliability engineering best practices within your team
  • Collaborates with other software engineers and teams to design, develop, test, and implement deployment and reliability approaches using automated continuous integration and continuous delivery pipelines
  • Uses enterprise-authorized AI capabilities within the work environment to accelerate incident triage, troubleshooting, and post-incident analysis, validating outputs and handling operational data according to sensitivity and security requirements.
  • Implements infrastructure, configuration, and network as code for the applications and platforms in your remit
  • Collaborates with technical experts, key stakeholders, and team members to resolve complex problems and proactively address issues using service level indicators and objectives before they impact customers
  • Applies enterprise-authorized AI capabilities within the work environment to identify patterns in operational signals that indicate reliability risk or recurring toil, prioritizing reuse-first improvements tied to SLO outcomes.
  • Familiar with availability, reliability, scalability, and solutions in their applications and works with partners to improve these outcomes iteratively
  • Proactively recognizes road blocks and identifies improvements to solve business problems, including exploring new technologies where appropriate

Required qualifications, capabilities, and skills 

  • Formal training or certification on site reliability engineering concepts and 3+ years applied experience 
  • Proficient in site reliability culture and principles and familiarity with how to implement site reliability within an application or platform
  • Proficient in at least one programming language such as Python, Java/Spring Boot, and .Net
  • Working knowledge of using enterprise-authorized AI capabilities within the work environment to support SRE workflows with strong validation habits and awareness of data sensitivity
  • Experience in Python, Grafana , Prometheus, Dynatrace, Datadog and Splunk, AWS, CICD experience, Jenkins, Github, Terraform
  • Ability to validate AI-assisted operational recommendations before applying changes, escalating when uncertain and following data sensitivity requirements
  • Proficient knowledge of software applications and technical processes within a given technical discipline (e.g., Cloud, AI, Android, etc.)
  • Experience in observability such as white and black box monitoring, service level objective alerting, and telemetry collection
  • Experience with continuous integration and continuous delivery tooling


Preferred  qualifications, capabilities, and skills 

  • Familiarity with AI coding assistant tools
  • Familiarity with container and container orchestration and troubleshooting common networking technologies and issues

JPMorganChase Hyderabad, Telangana, IND Office

JP Morgan Tower, Salarpuria Sattva Knowledge City, HITEC City, Raidurgam, Hyderabad, Telangana, India, 500081

Similar Jobs

10 Minutes Ago
Hybrid
Hyderabad, Telangana, IND
Senior level
Senior level
Financial Services
Leads software engineering for JPMorganChase’s Payments Trust and Safety Technology team, developing secure, scalable fraud protection solutions. Responsibilities include designing and delivering Java and Spring Boot microservices on AWS, reviewing code, troubleshooting production systems, improving resiliency and automation, and guiding responsible AI-assisted engineering practices. The role also evaluates vendor and architectural solutions, coaches engineers, and drives technical outcomes across agile teams.
Top Skills: Ai Coding AssistantsAmazon Ec2Amazon S3Apache KafkaAws EcsAws EksAws LambdaCi/CdDatabasesGithub CopilotJavaMicroservicesNetwork Load BalancerSpring Boot
11 Minutes Ago
Hybrid
Hyderabad, Telangana, IND
Senior level
Senior level
Financial Services
Leads platform and application support engineering for secure, stable, scalable development and test environments. Responsibilities include environment health management, incident triage and resolution, root-cause analysis, monitoring and automation, observability dashboards, tooling maintenance, post-change validation, and environment hygiene. The role also drives responsible adoption of AI-assisted engineering practices, CI/CD, secure coding, automated testing, and reliability improvements across distributed teams.
Top Skills: Ai-Assisted Software Development ToolsCi/CdContainerizationDynatraceEvent-Driven SystemsKibanaMicroservicesObservabilitySplunk
2 Hours Ago
In-Office
Hyderabad, Telangana, IND
Senior level
Senior level
Big Data • Fintech • Information Technology • Insurance • Financial Services
Leads medium to large technical projects and small to medium programs from planning through delivery. Manages stakeholders, roadmaps, budgets, resources, risks, requirements, reporting, and project artifacts. Facilitates Agile and Kanban practices, removes delivery barriers, supports product owners, and drives process improvement. The role requires autonomous leadership across technology, financial services, cybersecurity, information risk, and third-party vendor environments.
Top Skills: AgileClarity PpmConfluenceJIRAKanbanScrumSdlc

What you need to know about the Hyderabad Tech Scene

Because of its proximity to leading research institutions and a government committed to the city's growth, Hyderabad's tech scene is booming. With plans to establish India's first "AI city," the city is on track to become one of the world's most anticipated tech hubs, with companies like TransUnion, Schrödinger and Freshworks, among others, already calling the city home.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account