Lead technical design and resiliency for Electronic Trading Services, champion SRE practices, set SLOs/SLO indicators, lead incident response, mentor engineers, reduce toil, and implement observability, CI/CD, containerization, and networking improvements to ensure reliability and low-latency performance.
Assume a critical role in defining the future of a globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability.
As a Lead Site Reliability Engineer at JPMorgan Chase within the Infrastructure Platforms, Electronic Trading Services , you hold a leadership role in your team, demonstrate strong knowledge across multiple technical domains, and advise others on the technical and business issues facing them. Take lead and conduct resiliency design reviews, break up complex problems into digestible work for other engineers, act as a technical lead for medium to large-sized products, and provide advice and mentoring to other engineers.
Job responsibilities
- Demonstrates and champions site reliability culture and practices and exerts technical influence throughout your team
- Leads initiatives to improve the reliability and stability of your team’s applications and platforms using data-driven analytics to improve service levels
- Collaborates with team members to identify comprehensive service level indicators and stakeholders to establish reasonable service level objectives and error budgets with customers
- Demonstrates a high level of technical expertise within one or more technical domains and proactively identifies and solves technology-related bottlenecks in your areas of expertise
- Acts as the main point of contact during major incidents for your application and demonstrates the skills to identify and solve issues quickly to avoid financial losses
- Documents and shares knowledge within your organization via internal forums and communities of practice
Required qualifications, capabilities, and skills
- Bachelor’s Degree in Computer Science, Cybersecurity, Data Science, or related disciplines
-
Formal training or certification with 5+ years supporting critical security-focused applications in large-scale environments
-
Deep proficiency in reliability, scalability, performance, security, enterprise system architecture, toil reduction, and other site reliability best practices with the ability to implement these practices within an application or platform
- Fluency in at least one programming language such as (e.g., Python, Java Spring Boot, .Net, etc.) and experience working on low latency networks for Electronic Trading Services
- Deep knowledge of software applications and technical processes with emerging depth in one or more technical disciplines
- Proficiency and experience in observability such as white and black box monitoring, SLO alerting, and telemetry collection using tools such as Grafana, Dynatrace, Prometheus, Datadog, Splunk, etc.
- Proficiency in continuous integration and continuous delivery tools (e.g., Jenkins, GitLab, Terraform, etc.)
- Experience with container and container orchestration (e.g., ECS, Kubernetes, Docker, etc.)
- Experience with troubleshooting common networking technologies and issues
- Drive to self-educate and evaluate new technology & ability to teach new programming languages to team members
Preferred qualifications, capabilities, and skills
- Experience with Arista, Cisco, F5, and Fortinet devices
- Familiarity with network automation tools and techniques, such as Ansible
- Experience with Corvil and Wireshark
We recognize that our people are our strength and the diverse talents they bring to our global workforce are directly linked to our success. We are an equal opportunity employer and place a high value on diversity and inclusion at our company. We do not discriminate on the basis of any protected attribute, including race, religion, color, national origin, gender, sexual orientation, gender identity, gender expression, age, marital or veteran status, pregnancy or disability, or any other basis protected under applicable law. We also make reasonable accommodations for applicants’ and employees’ religious practices and beliefs, as well as mental health or physical disability needs. Visit our FAQs for more information about requesting an accommodation.
JPMorganChase Singapore, Singapore, SGP Office
One@Changi City, Changi Business Park Central 1, Singapore, Singapore, 486036
Similar Jobs
Financial Services
The Site Reliability Engineer will develop technical expertise, own production deployments, improve performance and reliability, and manage operational risk while collaborating across teams.
Top Skills:
CC++GoLinuxRust
Security • Software
The Site Reliability Engineer is responsible for maintaining application systems' high availability and stability, ensuring compliance with laws, managing changes, and driving automation for infrastructure optimization.
Top Skills:
AWSAzureDockerGitlab CiGCPJenkinsKubernetes
Database • Analytics
The Senior Site Reliability Engineer will ensure the reliability and performance of ClickHouse Cloud, managing incidents and collaborating with engineering teams to optimize systems.
Top Skills:
AnsibleAWSAzureClickhouseDockerGoGoogle Cloud PlatformKubernetesPuppetPythonSQLTerraform
What you need to know about the Singapore Tech Scene
The digital revolution has driven a constant demand for tech professionals across industries like software development, data analytics and cybersecurity. In Singapore, one of the largest cities in Southeast Asia, the demand for tech talent is so high that the government continues to invest millions into programs designed to develop a talent pipeline directly from universities while also scaling efforts in pre-employment training and mid-career upskilling to expand and elevate its workforce.



