Search Results for "it-sr-systems-engineer"

Found 3355 jobs

Senior Site Reliability Engineer -AI Infrastructure Operations

nscale Houston; San Francisco; Seattle

About Nscale
Nscale is the GPU cloud built for AI. We run high-performance, cost-efficient infrastructure for AI-native
startups and global enterprises, from bare metal up through the platform services teams actually build
on. Our culture runs on ownership, accountability, and speed. We move with urgen…

Posted: Aug 19, 2026 1 views
Apply Now

Staff Software Engineer, Platform

assured Remote full_time

Assured is on a mission to modernize insurance. Claims processing (i.e. should we pay this claim?), while often overlooked, is the foundation of the entire industry. It’s currently highly manual, involving phone calls, faxes, and gut instinct, costing tens of billions of dollars a year. We can do be…

Posted: Aug 18, 2026 0 views
Apply Now

Staff Software Engineer, Platform

assured United States remote

Assured is on a mission to modernize insurance. Claims processing (i.e. should we pay this claim?), while often overlooked, is the foundation of the entire industry. It’s currently highly manual, involving phone calls, faxes, and gut instinct, costing tens of billions of dollars a year. We can do be…

Posted: Aug 18, 2026 0 views
aws docker kubernetes typescript
Apply Now

Senior Cloud SRE - AI/ML Platform & GPU Compute

wayve London, United Kingdom full_time

The role

As a Cloud Site Reliability Engineer at Wayve, you will build and scale the reliability foundations of our AI cloud platform. This includes our Model Development Platform (powering end-to-end model development from raw data to on-road experimentation) and our GPU Compute platform (large-scal…

Posted: Aug 18, 2026 0 views
Apply Now

Software Engineer III, Site Reliability Engineering, GCP Identity

google Warsaw, Poland

Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google Cloud's services—both our internally critical and our externally-visible systems—have reliability, uptime appropriate to c…

Posted: Aug 19, 2026 0 views
Apply Now

Senior Software Engineer, Infrastructure

armada Bellevue, United States onsite

About the Company

Armada is the hyperscaler for the edge, delivering modular AI infrastructure from first deployment to AI factory with speed, scale and sovereignty. Named one of Fast Company's Most Innovative Companies and to the CNBC Disruptor 50, Armada’s solutions are deployed in over 60 countrie…

Posted: Aug 18, 2026 1 views
azure bgp go grpc ipsec kubernetes openshift ospf protobuf python
Apply Now

Senior Software Engineer, Infrastructure

armada Bellevue Office, Sunset Corporate Campus Hybrid

About the Company

Armada is the hyperscaler for the edge, delivering modular AI infrastructure from first deployment to AI factory with speed, scale and sovereignty. Named one of Fast Company's Most Innovative Companies and to the CNBC Disruptor 50, Armada’s solutions are deployed in over 60 countrie…

Posted: Aug 19, 2026 1 views
Apply Now

Senior Software Engineer, Site Reliability Engineering

google Warsaw, Poland

Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google Cloud's services—both our internally critical and our externally-visible systems—have reliability, uptime appropriate to c…

Posted: Aug 19, 2026 0 views
Apply Now

Staff Site Reliability Engineer, AI Foundations, F1 Query

google San Jose, CA, USA

Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google Cloud's services—both our internally critical and our externally-visible systems—have reliability, uptime appropriate to c…

Posted: Aug 19, 2026 0 views
Apply Now

Senior Software Engineer, Site Reliability Engineering, Cloud Storage

google Dublin, Ireland

Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google Cloud's services—both our internally critical and our externally-visible systems—have reliability, uptime appropriate to c…

Posted: Aug 19, 2026 0 views
Apply Now
Previous Page 52 of 336 Next