Kyndryl Logo

Kyndryl

Cloud Platform Resilience Engineer (Azure / AWS)

Posted 13 Days Ago
Be an Early Applicant
In-Office
South West, SGP
Senior level
In-Office
South West, SGP
Senior level
Operate and improve enterprise Azure and AWS cloud platforms, using Terraform and IaC. Monitor and troubleshoot platform health, support Kubernetes/container environments, run disaster recovery testing, automate operational tasks, collaborate across teams, and contribute to incident RCA and documentation to ensure platform resilience and service continuity.
The summary above was generated by AI

Who We Are

At Kyndryl, we run and reimagine the mission-critical technology systems that drive advantage for the world’s leading businesses.  We are at the heart of progress; with proven expertise and a continuous flow of AI-powered insight, enabling smarter decisions, faster innovation, and a lasting competitive edge. For our people—Kyndryls—that means doing purposeful work that powers human progress. Join us and experience a flexible, supportive environment where your well-being is prioritized and your potential can thrive.


The Role

As a Cloud Platform Resilience Engineer, you will play a key role in ensuring the reliability, availability and operational excellence of enterprise cloud platforms across Azure and AWS environments.

This is a hands-on engineering role focused on cloud operations, infrastructure automation, platform resilience and service continuity.

What You Will Do

  • Support and operate enterprise Azure and AWS cloud environments in a production setting.

  • Build, maintain and enhance cloud infrastructure using Terraform and Infrastructure as Code practices.

  • Monitor platform health, investigate service disruptions and restore services during incidents.

  • Support Kubernetes and container-based platforms, including operational maintenance and troubleshooting.

  • Participate in disaster recovery testing, backup validation and platform resilience initiatives.

  • Automate operational tasks and improve platform reliability through scripting and engineering solutions.

  • Collaborate with networking, security, infrastructure and application teams to resolve complex issues and improve service performance.

  • Contribute to root cause analysis, operational documentation and continuous service improvement activities.


Who You Are

Required Skills and Experience

  • 5+ years of experience supporting enterprise cloud platforms, infrastructure operations, platform engineering or site reliability environments.

  • Hands-on experience supporting Azure and/or AWS environments in production.

  • Practical experience with Terraform or other Infrastructure as Code technologies.

  • Working knowledge of Kubernetes, containers and cloud-native technologies.

  • Experience troubleshooting incidents and supporting highly available enterprise services.

  • Strong understanding of cloud networking, identity and access management, and security fundamentals.

  • Experience working within structured Incident, Change and Problem Management processes.

  • Strong analytical, communication and stakeholder management skills.

Preferred Skills and Experience

  • Experience with monitoring and observability platforms such as Azure Monitor, Amazon CloudWatch, Grafana or Prometheus.

  • Experience automating operational processes using PowerShell, Python, Bash or similar technologies.

  • Microsoft Azure, Amazon Web Services, Kubernetes, Terraform or ITIL certifications.

  • Experience supporting regulated, mission-critical or large-scale enterprise environments.

  • Exposure to AI-assisted operations, intelligent monitoring or automated incident analysis.


Being You

The “Kyn” in Kyndryl means kinship, which represents the strong bonds we have with each other, our customers and our communities. We focus on ensuring all Kyndryls feel included and we welcome people of all cultures, backgrounds, and experiences. Even if you don’t meet every requirement, we encourage you to apply. We believe in growth, and we’re excited to see what you can bring. At Kyndryl, employee feedback has told us that our number one driver of employee engagement is belonging. That sense of belonging — being a valued, respected, trusted member of the team — is fundamental to our culture and fueling great experiences for our customers. This dedication to welcoming everyone into our company means that Kyndryl gives you the ability to thrive and contribute to our culture of empathy and shared success. That’s The Kyndryl Way.

What You Can Expect

Your career with us isn’t just a job—it’s an adventure with purpose.  We offer a dynamic, hybrid-friendly culture that supports your well-being and empowers you to grow. Our Be Well programs are thoughtfully designed to support your financial, mental, physical, and social health—because we know that when you feel your best, you do your best.
From your very first day, you’ll dive into impactful work that powers the systems our customers rely on every day. You won’t just contribute—you’ll make a difference, tackling meaningful projects that sharpen your skills and fuel your growth.
We’re here to champion your journey. With powerful tools to chart your career path, personalized development goals aligned with your ambitions, and continuous feedback to keep you inspired and on track, you’ll have everything you need to thrive and evolve. You’ll develop in-demand skills to grow your career and achieve your ambitions with access to cutting-edge learning opportunities—from certifications with Microsoft, Google, and Amazon to coaching and hands-on experiences. And through it all, you’ll be part of a culture that values empathy, restless learning, and a devotion to shared success.
We want you to thrive here—and we’re committed to helping you do just that. Ready to make an impact? Join us and help shape what’s next.

Get Referred!

If you know someone that works at Kyndryl, when asked ‘How Did You Hear About Us’ during the application process, select ‘Employee Referral’ and enter your contact's Kyndryl email address.

Similar Jobs

Entry level
Artificial Intelligence • Hardware • Information Technology • Machine Learning
The Engineer for Product Development will evaluate new equipment, integrate AI tools for efficiency, and manage supplier engagement while supporting continuous improvement initiatives.
Top Skills: Ai-Assisted ToolsAutomationDigital FluencyInternet Of ThingsProcess Control System
An Hour Ago
In-Office
Singapore, SGP
Entry level
Entry level
Artificial Intelligence • Hardware • Information Technology • Machine Learning
Support New Product Introduction (NPI) execution activities in semiconductor assembly, ensuring timely builds and issue tracking across teams.
Top Skills: ExcelManufacturing SystemsSAP
An Hour Ago
In-Office
Singapore, SGP
Senior level
Senior level
Artificial Intelligence • Hardware • Information Technology • Machine Learning
The Manager - Global Engineering Labs will lead a team to setup and maintain engineering labs, ensuring compliance and efficiency in lab operations, while managing budgeting and collaborating with stakeholders.
Top Skills: Electrical Failure AnalysisPhysical Failure AnalysisSemiconductor Engineering

What you need to know about the Singapore Tech Scene

The digital revolution has driven a constant demand for tech professionals across industries like software development, data analytics and cybersecurity. In Singapore, one of the largest cities in Southeast Asia, the demand for tech talent is so high that the government continues to invest millions into programs designed to develop a talent pipeline directly from universities while also scaling efforts in pre-employment training and mid-career upskilling to expand and elevate its workforce.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account