Search Results for "systems-reliability-engineer-sre"
Found 2002 jobs
Senior Staff Cloud Backend Engineer - Observability and Site Reliability
Please complete the attached Internal Transfer Request Form and submit.
Please make sure to apply with your Coupang e-mail address.
Company Introduction
We exist to wow our customers. We know we’re doing the right thing when we hear our customers say, “How did we ever live without Coupang?” Born out of an…
Senior Staff Cloud Backend Engineer - Observability and Site Reliability
Company Introduction
We exist to wow our customers. We know we’re doing the right thing when we hear our customers say, “How did we ever live without Coupang?” Born out of an obsession to make shopping, eating, and living easier than ever, we’re collectively disrupting the multi-billion-dollar e-comm…
Staff Cloud SRE – AI/ML Platform & GPU Compute
The role
This is a rare opportunity to be a founding Staff SRE shaping the reliability of large-scale AI systems and GPU compute infrastructure from the ground up.
As a Staff Cloud Site Reliability Engineer at Wayve, you will build and scale the reliability foundations of our AI cloud platform. This i…
Lead Site Reliability Engineer (SRE)
In the initial months, you'll be working closely with the founding team, gradually taking ownership of central components.
We are on the brink of something monumental, anticipating a tidal wave of traffic. You'll be at the forefront of ensuring our systems not only handle the influx but thrive under…
Software Engineer III, Site Reliability Engineering, GCP Identity
Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google Cloud's services—both our internally critical and our externally-visible systems—have reliability, uptime appropriate to c…
Staff Cloud SRE – AI/ML Platform & GPU Compute
About us
Founded in 2017, Wayve is the leading developer of Embodied AI technology. Our advanced AI software and foundation models enable vehicles to perceive, understand, and navigate any complex environment, enhancing the usability and safety of automated driving systems.
Our vision is to create aut…
Senior DevOps Engineer / SRE (m/f/d)
🎤 Why voize? Because we’re more than just a job!
At voize, we believe the greatest gift to frontline workers is time - time to care, connect, and be present. Today, that time is lost to busywork and complex systems that pull them away from what matters most: people.
Our vision is to change that by buildin…
SRE - Linux
Verisign helps enable the security, stability, and resiliency of the internet. We are a trusted provider of internet infrastructure services for the networked world and deliver unmatched performance in domain name system (DNS) services.
We are a mission focused, values driven company where each indiv…
Site Reliability Engineer - Big Data (7 to 11 years)
About PhonePe Limited:
Headquartered in India, its flagship product, the PhonePe digital payments app, was launched in Aug 2016. As of April 2025, PhonePe has over 60 Crore (600 Million) registered users and a digital payments acceptance network spread across over 4 Crore (40+ million) merchants. Pho…
Manager- Site Reliability Engineering
Secure Every Identity, from AI to HumanIdentity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world st…