Search Results for "network-and-site-reliability-engineer"

Found 795 jobs

Site Reliability Engineer

cow-dao Remote (Europe preferred) Remote

About CoW DAO

CoW DAO is on a mission to protect Ethereum users from MEV and optimize trade execution across DeFi. We achieve this through the CoW Protocol, CoW Swap (a leading intent-based DEX aggregator), and the innovative MEV Blocker, which together help secure, aggregate, and route trades f...

Posted: May 7, 2025 21 views
AWS Cloud Docker Engineer Kubernetes Linux SRE
Apply Now

Lead Site Reliability Engineer (Remote)

livepeer Remote Remote

About Livepeer:

Livepeer is on a mission to build the world’s open video infrastructure. Founded in 2017, it is the world’s first open-source protocol for decentralized video streaming, built on Ethereum. The project has empowered developers to create scalable, cost-effective, and censorship-re...

Posted: Jun 12, 2025 20 views
Ansible AWS CI/CD DevOps Docker Engineer Ethereum GCP Grafana Kubernetes Linux Prometheus SRE Terraform Web3
Apply Now

Tech Lead, Site Reliability Engineering (SRE)

edge-node Remote Full-Time

At Edge & Node, we’re focused on building The Graph, a decentralized protocol for accessing and organizing the world’s knowledge and information. Subgraphs, a core technology developed by Edge & Node to access blockchain data, are widely used across web3 to power decentralized applications.

We’re a...

Posted: Apr 14, 2025 13 views
Cloud DevOps Engineer GCP Kubernetes SRE Tech Lead
Apply Now

Site Reliability Engineer (SRE)

baseten San Francisco Office; New York; Remote Hybrid

ABOUT BASETEN

Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the fronti...

Posted: Apr 14, 2026 13 views
site reliability engineer sre kubernetes infrastructure terraform cloudformation pulumi ci/cd github actions gitlab ci circle ci jenkins prometheus elk stack grafana opentelemetry machine learning ml ai infrastructure-as-code hybrid remote
Apply Now

Senior Software Engineer, Infrastructure

decagon New York City Full time

About Decagon

Decagon is the leading conversational AI platform empowering every brand to deliver concierge customer experiences.

Our technology enables industry-defining enterprises like Avis Budget Group, Block’s Cash App and Square, Chime, Oura Health, and Hunter Douglas to deploy AI agents tha...

Posted: Apr 7, 2026 14 views
Senior Software Engineer Infrastructure Decagon conversational AI platform networking data ML serving developer platform real-time voice SLOs low-latency production infrastructure Kubernetes GCP AWS Azure Terraform GitOps CI/CD observability OpenTelemetry Prometheus Grafana Datadog on-prem air-gapped customer-managed deployments equity New York City on-site
Apply Now

Site Reliability Engineer - AI & ML Infrastructure (Kubernetes, AWS & Terraform)

deepgram USA | Remote Remote

Company Overview

Deepgram is the leading platform underpinning the emerging trillion-dollar Voice AI economy, providing real-time APIs for speech-to-text (STT), text-to-speech (TTS), and building production-grade voice agents at scale. More than 200,000 developers and 1,300+ organizations build v...

Posted: Apr 7, 2026 10 views
Site Reliability Engineer SRE AI ML Infrastructure Kubernetes AWS Terraform Slurm GPU HPC DevOps Platform Engineering Python Go Bash CI/CD Hybrid Cloud
Apply Now

Senior DevOps Engineer

olo Belfast, Northern Ireland, Remote Remote

Olo is a leading SaaS platform accelerating digital transformation in the restaurant industry, by helping customers deliver more personalised and profitable guest experiences. As a result, our digital ordering, payment, and guest engagement solutions enable brands to do more with less and make every...

Posted: Apr 7, 2026 19 views
devops senior aws kubernetes eks terraform ci/cd gitops linux bash python .net argo cd flux github actions prometheus kafka postgres soc2 pci
Apply Now

Senior Platform Engineer

attio London; United Kingdom Hybrid

Attio is the CRM built for the AI era. Designed for the most ambitious go-to-market teams, it gives companies the power to understand every customer, automate at scale, and build their go-to-market motion exactly as they need. We've raised $116M from some of the world's best investors: GV (Google Ve...

Posted: Apr 7, 2026 10 views
senior platform engineer devops sre aws gcp azure docker kubernetes typescript go python rust terraform pulumi ci/cd monitoring logging tracing infrastructure automation cloud containerization observability
Apply Now

Software Engineer, Site Reliability (SRE)

sierra San Francisco, CA Full time

About us

  • At Sierra, we’re creating a platform to help businesses build better, more human customer experiences with AI. We are primarily an in-person company based in San Francisco, with growing offices in Atlanta, New York, London, Paris, Madrid, Munich, Singapore, Japan, and Sydney.
  • We are...
Posted: Apr 15, 2026 9 views
software engineer site reliability engineer sre infrastructure terraform aws observability devops ci/cd llm ai saas cloud
Apply Now

Staff Software Engineer, AI Reliability Engineering

anthropic London, UK Hybrid

About Anthropic

Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working...

Posted: Apr 7, 2026 8 views
staff software engineer AI reliability engineering SRE distributed systems infrastructure monitoring observability high-availability incident response ML hardware GPUs TPUs reliability Anthropic London
Apply Now
Page 1 of 80 Next