Search Results for "software-engineer-model-inference"

Found 1036 jobs

Software Engineer, Productivity - Model Performance

openai San Francisco full_time

About the Team

We’re hiring software engineers to make OpenAI’s Model Performance teams more productive. These teams work on the systems, tooling, and infrastructure that help improve model performance across OpenAI’s training and inference workloads at frontier scale.

Posted: Aug 18, 2026 0 views

Staff Software Engineer, Inference

anthropic London, United Kingdom hybrid

About Anthropic

Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and...

Posted: Aug 18, 2026 0 views
python rust

Software Engineer, Inference - Multi Modal

openai San Francisco full_time

About the Team

OpenAI’s Inference team powers the deployment of our most advanced models - including our GPT models, 4o Image Generation, and Whisper - across a variety of platforms. Our work ensures these models are available, performant, and scalable in production, and we...

Posted: Aug 18, 2026 0 views

AI Inference Infrastructure Software Engineer (Kubernetes / Cloud)

elastix Seattle full_time

Location: Seattle, WA (Hybrid - 3 days/week in office)

About ElastixAI:

ElastixAI is an early-stage Software startup on a mission to reinvent AI inference infrastructure from the ground up. We're building a next-generation inference platfor...

Posted: Aug 18, 2026 0 views

Software Engineer, Inference Runtime

lm-studio New York City full_time

LM Studio is used by millions of people around the world to run AI on their own computers, and now with Bionic - also in the cloud. Our values prioritize putting the human in the center, and creating tools that we want to use ourselves, and recommend to our friends and family.

As a team, we...

Posted: Aug 18, 2026 0 views

Software Engineer, Inference - Performance Optimization

openai San Francisco full_time

About the Team
Our team analyzes inference stack performance across the application, model, and fleet layers to identify bottlenecks and drive faster, cheaper inference. We combine systems profiling, benchmarking, and analysis to understand where time and cost are spent, the...

Posted: Aug 18, 2026 0 views

Staff Software Engineer, Inference

anthropic Dublin, Ireland hybrid

About Anthropic

Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and...

Posted: Aug 18, 2026 0 views
kubernetes python rust

Software Engineer - Baseten Inference Stack

baseten San Francisco full_time

ABOUT BASETEN

Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enab...

Posted: Aug 18, 2026 0 views

Software Engineer- Model Performance Systems

baseten San Francisco full_time

ABOUT BASETEN

Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enab...

Posted: Aug 18, 2026 0 views

Software Engineer, GPU Inference

cerebras US and Canada Offices full_time

Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.

This order of magnitude increase in s...

Posted: Aug 18, 2026 0 views
Previous Page 12 of 104 Next