Lifted, an Upwork Company Logo

Lifted, an Upwork Company

#116807 - Senior Software Engineer / SRE (Observability Focus)

Sorry, this job was removed at 02:19 a.m. (SGT) on Saturday, Aug 01, 2026
Be an Early Applicant
In-Office or Remote
Hiring Remotely in Singapore, SGP
Senior level
In-Office or Remote
Hiring Remotely in Singapore, SGP
Senior level

Similar Jobs

41 Minutes Ago
Remote or Hybrid
Singapore, SGP
Mid level
Mid level
Cloud • Information Technology • Security • Software • Cybersecurity
The Customer Engineer will serve as a technical advisor, managing customer relationships, validating technical solutions, and ensuring long-term adoption and growth of Cloudflare's services while leveraging AI for workflow automation.
Top Skills: AIAWSAzureDeveloper PlatformsGCPIsacaIsc2JavaScriptNetworkingPalo AltoPythonSecurity
41 Minutes Ago
Remote or Hybrid
Singapore, SGP
Mid level
Mid level
Cloud • Information Technology • Security • Software • Cybersecurity
The Enterprise Customer Engineer at Cloudflare acts as a technical advisor managing customer relationships, technical validations, and business expansions through AI-augmented workflows.
Top Skills: AICloud InfrastructureCloudflareJavaScriptNetworkingPythonSaaSWeb Security
41 Minutes Ago
Remote or Hybrid
Singapore, SGP
Senior level
Senior level
Cloud • Information Technology • Security • Software • Cybersecurity
Lead rapid-response and proactive reliability work for high-severity, customer-facing incidents. Debug across edge, network, DNS, transport, and customer stacks; own on-call, postmortems, and remediation. Build telemetry, detectors, diagnostic tooling, and automation with Product Engineering. Mentor support teams, define reliability metrics, and ship AI-assisted diagnostics to reduce detection and resolution time.
Top Skills: BashBgpDistributed TracingDnsElasticsearchGrafanaGreHttp/SIpsecKibanaNtpOspfPythonSmtpSnmpTcpdumpTls/SslWiresharkZero Trust
Develop and operate observability-driven platform reliability: integrate APIs, manage Kubernetes deployments, build dashboards/alerts/tracing/metrics/logging, automate monitoring tasks (Python preferred), configure Datadog and integrations in AWS, and drive reliability, scalability, and performance improvements.
The summary above was generated by AI
Job Description

We are seeking a Senior Software Engineer / SRE to support platform reliability, monitoring, and modernization initiatives. This role combines software engineering and site reliability engineering, with a strong emphasis on Kubernetes, observability, and cloud infrastructure. You will help improve system reliability, automate operational workflows, and support the organization’s Datadog-based observability environment.

Key Responsibilities
- Support platform reliability, monitoring, and modernization initiatives
- Design, consume, and integrate APIs across internal systems and services
- Work in Kubernetes environments across deployment, operations, and monitoring
- Build and improve observability capabilities across dashboards, alerts, tracing, metrics, and logging
- Monitor containerized and microservices-based architectures
- Integrate observability tooling into AWS environments
- Support CI/CD observability integrations
- Automate monitoring and operational tasks using scripting, with Python preferred
- Help own and operate internal engineering platform capabilities, with extra emphasis on observability platforms
- Drive proactive maintenance efforts and platform improvements focused on reliability, scalability, and performance
- Install and configure Datadog agents and integrations
- Manage API keys and secure configuration practices for observability tooling
- Manage user roles and access controls within observability platforms
- Provide operational and training support related to Datadog

Qualifications

Must-Have Skills
- Strong proficiency in at least one of the following: Python, JavaScript (Node.js), or Java
- Hands-on experience with API integrations
- Strong experience working in Kubernetes environments
- Experience with Datadog or similar observability tools such as Prometheus or Grafana
- Ability to configure dashboards, alerts, and APM
- Experience with tracing, metrics, and logging
- Experience monitoring containerized or microservices-based architectures
- Hands-on experience with AWS
- Experience integrating observability tools into cloud environments
- Experience with CI/CD integrations for observability
- Ability to automate monitoring and operational tasks using scripting
- Demonstrated ownership of reliability, scalability, and performance improvements
- Experience installing and configuring observability agents and integrations
- Experience managing secure configurations, API keys, and user access within observability platforms
- Enterprise experience strongly preferred.

Nice-to-Have Skills
- Familiarity with Go (Golang)
- Experience with New Relic
- Experience with Dynatrace
- Experience with Elastic
- Experience with Splunk Observability

Required Tools & Platforms
- Kubernetes
- Datadog
- AWS
- Prometheus
- Grafana
 

Additional Information

Location, Time & Engagement
- Contract role
- Full-time, 40 hours per week
- Remote within APAC
- Must be able to work hours overlapping U.S. Pacific Time
- 100% allocation

What you need to know about the Singapore Tech Scene

The digital revolution has driven a constant demand for tech professionals across industries like software development, data analytics and cybersecurity. In Singapore, one of the largest cities in Southeast Asia, the demand for tech talent is so high that the government continues to invest millions into programs designed to develop a talent pipeline directly from universities while also scaling efforts in pre-employment training and mid-career upskilling to expand and elevate its workforce.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account