Search Results for "software-engineer-inference-runtime"

Found 142 jobs

Staff+ Software Engineer, Inference Runtime

anthropic Remote-Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NY Remote-Friendly (Travel-Required)

About Anthropic

Anthropic's mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working...

Posted: Jun 13, 2026 7 views
Staff+ Software Engineer Inference Runtime Anthropic Remote San Francisco Seattle New York Rust Python GPU TPU Trainium ML Infrastructure
Apply Now

Staff Software Engineer, Kubernetes Platform

anthropic San Francisco, CA | New York City, NY | Seattle, WA Full-time

About Anthropic

Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working...

Posted: May 6, 2026 4 views
Staff Software Engineer Kubernetes Platform Anthropic San Francisco New York Seattle Go Python Rust C++ Kubernetes GCP AWS Linux eBPF etcd NCCL
Apply Now

Lead Software Platform Engineer, MLOps

tetrascience United States remote

Who We Are

TetraScience is the Scientific Data and AI Cloud company. We are catalyzing the Scientific AI revolution by designing and industrializing AI-native scientific data sets, which we bring to life in a growing suite of next gen lab data management solutions, scientific use cases,...

Posted: Aug 18, 2026 1 views
aws aws-bedrock databricks docker llm mlflow python typescript
Apply Now

AI Engineer

armada Bellevue, United States onsite

About the Company

Armada is the hyperscaler for the edge, delivering modular AI infrastructure from first deployment to AI factory with speed, scale and sovereignty. Named one of Fast Company's Most Innovative Companies and to the CNBC Disruptor 50, Armada’s solutions...

Posted: Aug 18, 2026 1 views
cpp cuda java jax python pytorch tensorflow
Apply Now

Staff+ Software Engineer, Inference Runtime

anthropic San Francisco, United States hybrid

About Anthropic

Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and...

Posted: Aug 18, 2026 0 views
cuda python rust triton
Apply Now

Software Engineer, Productivity - Inference Runtime

openai San Francisco full_time

About the Team

We’re hiring a Developer Productivity engineer to support OpenAI’s Inference Runtime teams. These teams own the systems responsible for serving models reliably, efficiently, and safely across Codex, ChatGPT, API, and internal research workloads. We’re hirin...

Posted: Aug 18, 2026 0 views
Apply Now

Software Engineer, Inference Runtime

lm-studio New York City full_time

LM Studio is used by millions of people around the world to run AI on their own computers, and now with Bionic - also in the cloud. Our values prioritize putting the human in the center, and creating tools that we want to use ourselves, and recommend to our friends and family.

As a team, we...

Posted: Aug 18, 2026 0 views
Apply Now

Software Engineer - Voice AI (Inference Runtime)

baseten San Francisco full_time

ABOUT BASETEN

Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enab...

Posted: Aug 18, 2026 0 views
Apply Now

Software Engineer – AI Inference Engine

friendliai San Francisco full_time

About the Job

We are seeking a highly technical Inference Engine Engineer to optimize the performance and efficiency of our core inference engine. In this role, you will focus on designing, implementing, and optimizing GPU kernels and supporting infrastructure for next-ge...

Posted: Aug 18, 2026 0 views
Apply Now

Software Engineer – AI Inference Engine

friendliai Seoul full_time

About the Job

We are seeking a highly technical Inference Engine Engineer to optimize the performance and efficiency of our core inference engine. In this role, you will focus on designing, implementing, and optimizing GPU kernels and supporting infrastructure for next-ge...

Posted: Aug 18, 2026 0 views
Apply Now
Page 1 of 15 Next