JPMorganChase fournit des services bancaires, de financement, de gestion d’actifs et de traitement des paiements aux particuliers, entreprises, institutions et gouvernements.
Site Reliability Engineer II - AI & Corporate Risk Tech
Am I a fit — voir ma compatibilitéJPMorganChase recherche un Site Reliability Engineer II au sein de Corporate Risk Technology à Glasgow. Le poste consiste à configurer, maintenir, surveiller et optimiser des applications et leur infrastructure via le code et le cloud. Une expérience en SRE, en observabilité, en CI/CD et avec un langage comme Python, Java/Spring Boot ou .NET est requise.
Repères sur JPMorganChase
- Domaine officiel
- jpmorganchase.com
- Offres ouvertes
- 303
Détails de l’offre
La description complète publiée par JPMorganChase.
Description de l’offre
There's nothing more exciting than being at the center of a rapidly growing field in technology and applying your skills to drive innovation and modernize some of the world's most complex and mission-critical systems. At JPMorganChase, you'll be part of a team that values curiosity, collaboration, and continuous improvement — where your contributions directly shape the reliability and resilience of platforms that matter. As a Site Reliability Engineer II at JPMorganChase within Corporate Risk Technology, you will solve complex and broad business problems with simple, straightforward solutions.
Through code and cloud infrastructure, you will configure, maintain, monitor, and optimize applications and their associated infrastructure — independently decomposing and iteratively improving on existing solutions. You are a meaningful contributor to your team, sharing your knowledge of end-to-end operations, availability, reliability, and scalability of your application or platform.
Job responsibilities
Guide and support team members in building appropriate-level designs, gaining peer consensus, and driving adoption of site reliability engineering best practices across the team
- Collaborate with software engineers and cross-functional teams to design, develop, test, and implement deployment and reliability approaches using automated continuous integration and continuous delivery pipelines
- Implement infrastructure, configuration, and network as code for the applications and platforms within your scope
- Partner with technical experts, key stakeholders, and team members to resolve complex problems and proactively address issues using service level indicators and objectives before they impact customers
- Identify and address roadblocks, propose improvements to solve business problems, and explore new technologies where appropriate
- Apply familiarity with availability, reliability, and scalability principles to iteratively improve outcomes in collaboration with partners
- Uses enterprise-authorized AI capabilities within the work environment to accelerate incident triage, troubleshooting, and post-incident analysis, validating outputs and handling operational data according to sensitivity and security
requirements
Applies enterprise-authorized AI capabilities within the work environment to identify patterns in operational signals that indicate reliability risk or recurring toil, prioritizing reuse-first improvements tied to service level objective outcomes
Required qualifications
, capabilities, and skills
- Formal training or certification on site reliability engineering concepts and proficient applied experience
- Proficiency in site reliability culture and principles, with the ability to implement site reliability practices within an application or platform
- Proficiency in at least one programming language such as Python, Java/Spring Boot, or .NET
- Experience in observability practices such as white and black box monitoring, service level objective alerting, and telemetry collection
- Proficient knowledge of software applications and technical processes within a given technical discipline (e.g., cloud, AI, mobile platforms)
- Working knowledge of using enterprise-authorized AI capabilities within the work environment to support site reliability engineering workflows, with strong validation habits and awareness of data sensitivity
- Ability to review and validate AI-assisted operational recommendations before applying changes, escalating when uncertain and following security and data handling requirements Preferred qualifications, capabilities, and skills
- Experience with continuous integration and continuous delivery tooling
- Familiarity with container technologies and container orchestration platforms
- Experience troubleshooting common networking technologies and issues
- There’s nothing more exciting than being at the center of a rapidly growing field in technology and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.
Prérequis
- Formation ou certification en concepts de site reliability engineering
- Expérience appliquée en site reliability engineering
- Principes et pratiques de site reliability engineering
- Observabilité
- Surveillance white box et black box
- Alertes fondées sur des objectifs de niveau de service
- Collecte de télémétrie
- Cloud
- AI
- Validation des recommandations opérationnelles assistées par IA