Site Reliability Engineer II - Java/Python, Kubernetes, AWS, Terraform

Department Icon Data Science Analytics & Machine Learning
149+ Applicants
Posted: 5 hours ago
0-1 years
Bengaluru / Bangalore, Karnataka
work from office

Posted: 5 hours ago
|
Applicants: 149+
Job Description
About Company
Similar Jobs
Please verify your account first! Send OTP

Job Description

Play a key role in ensuring system reliability at one of the worlds most iconic and largest financial institutions.
As a Site Reliability Engineer II at JPMorgan Chase within the Commercial Investment Bank team, you will use technology to solve business problems and leverage software engineering best practices as we strive towards excellence. This role often works independently to execute small to medium projects, but youll also have the opportunity to collaborate with cross functional teams to continually improve your level of knowledge about JPMorgan Chases business and relevant technologies.
Job responsibilities
  • Executes small to medium projects independently with initial direction and graduates to designing and delivering projects independently
  • Leverages technology to solve business problems by writing high quality, maintainable, and robust code following best practices in software engineering
  • Uses enterprise-authorized AI capabilities within the work environment to speed up incident triage, troubleshooting, and post-incident analysis, validating outputs and handling operational data according to sensitivity and security requirements.
  • Participates in triaging, examining, diagnosing, and resolving incidents and works with others to solve problems at their root
  • Recognizes toil within the role and proactively works towards eliminating it through systems engineering or updating application code
  • Understands observability patterns and strives to implement and improve service level indicators, objectives monitoring, and alerting solutions for optimal transparency and analysis
  • Applies enterprise-authorized AI capabilities within the work environment to identify recurring toil and reliability risks from operational signals, prioritizing reuse-first improvements and measurable SLO outcomes.
Required qualifications, capabilities, and skills
  • Formal training or certification on site reliability engineering concepts and 2+ years applied experience
  • Proficiency in reliability, scalability, performance, security, enterprise system architecture, toil reduction, and other site reliability best practices with the ability to implement these practices within an application or platform
  • Fluency in at least one programming language such as (e.g., Java/Python, Java Spring Boot, etc.)
  • Deep knowledge of software applications and technical processes with emerging depth in one or more technical disciplines
  • Proficiency and experience in observability such as white and black box monitoring, SLO alerting, and telemetry collection using tools such as Grafana, Dynatrace, Prometheus, Datadog, Splunk, etc. Proficiency in continuous integration and continuous delivery tools (e.g., Jenkins, GitLab, Terraform, etc.)
  • Grasp of SDLC, secure development, DevOps/CI/CD tooling capable of implementing top-tier continuous improvement with root-cause analysis and auto-remediation.
  • Demonstrated experience using enterprise-authorized AI capabilities within the work environment to improve SRE workflows (e.g., incident investigation support and knowledge capture) with strong validation habits and awareness of data sensitivity.
  • Use evaluate AI-assisted operational recommendations for correctness and risk, define appropriate guardrails for team usage, and ensure outcomes align to resiliency and security expectations.

    Looking to get Placed? Try our Placement Guarantee Plan

  • Experience with monitoring/logging tools (e.g., Splunk, AppDynamics) and dashboard technologies.
  • Experience with container and container orchestration (e.g., ECS, Kubernetes, Docker, etc.)
  • Experience with troubleshooting common networking technologies and issues

    Preferred qualifications, capabilities, and skills

  • Tools / Agentic AI to solve common opportunities in SRE Domain
  • Certification towards CI/CD and AI skills. Splunk Administrator certification desired.
  • Drive to self-educate and evaluate new technology

Skills

PythonSplunkAiGuardrailsAgenticAgentic Ai

If a job posting appears fraudulent, asks for payment, contains misleading information, or violates our guidelines, please report it immediately. Our team will review it promptly, Jobaaj does not charge any fee from the applicants.

About Company

JPMorgan Chase & Co., often referred to as Chase, is an American multinational investment bank and financial services holding company headquartered in New York City.

Important dates & deadlines?

Application Deadline

25 Nov 26, 03:39 PM IST

Similar Jobs

View All
Loading...
Bag Logo
Jobaaj
Don't Miss out any Updates

Subscribe now for the latest job alerts
and never miss an update

Job Alert
Google hiring for Specific Roles Apply Now!
1 min ago
New Opportunity
Amazon is hiring freshers Apply Now!
5 min ago
Featured Jobs
Microsoft opening 50+ positions Apply Now!
10 min ago

Site Reliability Engineer II - Java/Python, Kubernetes, AWS, Terraform

Share with