Search Results for "software-engineer-reliability"

Found 16503 jobs

Site Reliability Engineering, Fabric (Mid, Senior, or Staff)

mongodb Toronto; Vancouver Hybrid

The Team

Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that support the broader engineering organization. Among these are our multi-cloud-provider Kubernetes infrastructure, deployment machinery, and observabilit…

Posted: Aug 19, 2026 0 views
Apply Now

Site Reliability Engineering, Fabric (Mid, Senior, or Staff)

mongodb United States Hybrid

The Team

Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that support the broader engineering organization. Among these are our multi-cloud-provider Kubernetes infrastructure, deployment machinery, and observabilit…

Posted: Aug 19, 2026 0 views
Apply Now

Software Engineering Manager, Site Reliability Engineering, Google Cloud

google Warsaw, Poland

Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google's services—both our internally critical and our externally-visible systems—have reliability, uptime appropriate to users'…

Posted: Aug 19, 2026 0 views
Apply Now

Senior Software Engineer (SMB)

nerdwallet NerdWallet US Remote, Full time

At NerdWallet, we’re on a mission to bring clarity to all of life’s financial decisions and every great mission needs a team of exceptional Nerds. We’ve built an inclusive, flexible, and candid culture where you’re empowered to grow, take smart risks, and be unapologetically yourself (cape optional)…

Posted: Apr 7, 2026 3 views
senior software engineer smb full-stack ruby rails javascript react sql aws azure google cloud restful apis graphql mvc circleci github actions git agile scrum devops ci/cd remote full time
Apply Now

Site Reliability Engineering, Fabric

mongodb Toronto, Canada hybrid

The Team

Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that support the broader engineering organization. Among these are our multi-cloud-provider Kubernetes infrastructure, deployment machinery, and observabilit…

Posted: Aug 18, 2026 0 views
aws azure cloud gcp kubernetes
Apply Now

Tech Lead, Software Engineering

superdial Burlingame, CA full_time

SuperDial is looking for a Tech Lead, Full Stack Engineering to build and scale our AI-powered healthcare platform. This role is ideal for an engineer who is strong across the stack, with deep expertise in Python-based back-end systems and modern JavaScript front-end development.

About the role

  • Own an…

Posted: Aug 18, 2026 0 views
Apply Now

Software Engineer, Test Infrastructure

zipline South San Francisco, California, USA

About Zipline

Zipline is the world’s largest and most experienced drone delivery service. We are on a mission to serve all humans equally by ensuring access to food, medicine and essential goods anytime, anywhere. We design, build, and operate the world’s largest autonomous logistics system, deliveri…

Posted: Aug 19, 2026 0 views
Apply Now

Software Engineer, Observability

airtable San Francisco, CA; New York, NY; Remote (Seattle, WA only) Hybrid/Remote

Airtable is the no-code app platform that empowers people closest to the work to accelerate their most critical business processes. More than 500,000 organizations, including 80% of the Fortune 100, rely on Airtable to transform how work gets done.

The Observability team at Airtable ensures that our…

Posted: Apr 7, 2026 5 views
software engineer observability logging metrics tracing prometheus grafana datadog opentelemetry elk stack kubernetes distributed systems llm ai infrastructure sre site reliability
Apply Now

Software Development Manager II, Site Reliability Developing

google Waterloo, ON, Canada

Site Reliability Development combines software and systems development to build and run large-scale, massively distributed, fault-tolerant systems. Site Reliability Development ensures that Google's services—both our internally critical and our externally-visible systems—have reliability, uptime app…

Posted: Aug 19, 2026 0 views
Apply Now

Staff+ Software Engineer, Observability

anthropic London, UK Hybrid

About Anthropic

Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working toge…

Posted: Apr 7, 2026 3 views
staff software engineer observability infrastructure monitoring telemetry metrics logging tracing python rust go prometheus grafana clickhouse opentelemetry kubernetes ebpf ai llm alerting slo london
Apply Now
Previous Page 65 of 1651 Next