Search Results for "site-reliability-engineer-fleet"

Found 151 jobs

Site Reliability Engineer

asymmetric.re Remote - AMER Remote

Asymmetric Research:

Asymmetric Research ("AR") is a boutique security venture focused on deep partnerships with L1/L2 blockchains and DeFi protocols in an effort to keep them safe. We specialize in four core domains of web3 security: research, engineering, incident response, and infrastructure...

Posted: Apr 7, 2026 6 views
site reliability engineer sre devops infrastructure linux load balancer haproxy ansible chef puppet saltstack go python rust ci/cd grafana loki prometheus alertmanager nomad kubernetes blockchain bitcoin ethereum solana cosmos remote full time
Apply Now

Blockchain Reliability Engineer (DevOps)

nirvana-labs Remote Remote, Full-Time

Role

  • Setup, maintain and monitor fleets of blockchain nodes for Nirvana’s global customer base.
  • Build and maintain scaleable and flexible automations for running blockchain nodes and internal monitoring services.
  • Set and maintain the best-in-class infrastructure practices.
  • Collaborate w...
Posted: Apr 15, 2025 6 views
ansible bash chef devops engineer puppet rpc nodes sre terraform
Apply Now

Senior Site Reliability Engineer

mongodb Gurugram, India hybrid

We are seeking a Senior Site Reliability Engineer to join our growing Gurugram Products & Technology team to provide technical direction, shape architecture, and build key operational foundations of a new platform we are building to make it easier for customers to build AI applications using...

Posted: Aug 18, 2026 1 views
dns istio kubernetes python tcp-ip tls
Apply Now

Software Engineer, Model Deployment- ChatGPT Engineering

openai London, UK full_time

About the Team

ChatGPT relies on a large and growing GPU fleet to serve inference workloads reliably and efficiently. We develop the systems and tools that make it possible to introduce new models, manage production deployments, respond to operational issues, and use infr...

Posted: Aug 18, 2026 1 views
Apply Now

Staff Engineer, Datacenter Server Lifecycle

anthropic London, UK Hybrid

About Anthropic

Anthropic's mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working...

Posted: Apr 29, 2026 11 views
Staff Engineer Datacenter Server Lifecycle Anthropic London UK Python Rust Go Java Kubernetes AWS GCP GPU AI accelerator hardware lifecycle datacenter operations
Apply Now

Senior Platform Engineer

parachute-health United States remote

Parachute Health is transforming post-acute care as the leading digital ordering platform for medical equipment and supplies. We connect major health systems, health plans, and suppliers to help patients get the life-saving products they need at home. Since launching, we've connected 300,000+...

Posted: Aug 18, 2026 1 views
kubernetes python react ruby typescript
Apply Now

Senior Site Reliability Engineer, Fleet Management

mongodb Austin, United States hybrid

The Team

Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that support the broader engineering organization. Among these are our multi-cloud-provider Kubernetes infrastructure, networking, load balancin...

Posted: Aug 18, 2026 0 views
kubernetes python terraform
Apply Now

Site Reliability Engineer

apple Dublin

Become a Site Reliability Engineer in Apple’s Cloud Service Infrastructure team, part of Apple’s Services Engineering organisation, and help scale the cloud that underpins services for billions of Apple users.

We are building and supporting new and existing infrastructure to support the hy...

Posted: Aug 18, 2026 0 views
Apply Now

Senior Site Reliability Engineer

andromeda Global Remote / San Francisco, CA full_time

Senior Site Reliability Engineer

Location: Global Remote / San Francisco · Full-Time

About Andromeda

Andromeda gives AI companies access to the kind of scaled compute once reserved for hyperscalers. Our platform connects 100+ AI customers...

Posted: Aug 18, 2026 0 views
Apply Now

Senior Site Reliability Engineer

braze São Paulo, Brazil hybrid

At Braze, we have found our people. We’re a genuinely approachable, exceptionally kind, and intensely passionate crew.

We seek to ignite that passion by setting high standards, championing teamwork, and creating work-life harmony as we collectively navigate rapid growth on a global scale w...

Posted: Aug 18, 2026 0 views
ansible docker javascript kubernetes linux mongodb python ruby terraform
Apply Now
Page 1 of 16 Next