Search Results for "ml-engineer-inference-and-optimization"

Found 716 jobs

Staff Software Engineer, Inference Cloud

cerebras Headquarters/Sunnyvale Office full_time

Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.

This order of magnitude increase in s...

Posted: Aug 18, 2026 0 views
Apply Now

Staff Software Engineer, Inference Platform

cerebras Headquarters/Sunnyvale Office full_time

Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.

This order of magnitude increase in s...

Posted: Aug 18, 2026 0 views
Apply Now

Software Engineer, Inference

pulse San Francisco full_time

Overview


Pulse is tackling one of the most persistent challenges in data infrastructure: extracting accurate, structured information from complex documents at scale. We have a breakthrough approach to document understanding that combines intelligent sc...

Posted: Aug 18, 2026 0 views
Apply Now

Software Engineer: ML Infra

generalist San Francisco Bay Area (San Mateo) or Boston (Somerville) full_time

About the Role

Generalist trains very large robot foundation models. This requires utilizing very large numbers of the latest generation GPU hardware and infrastructure (currently Nvidia) to run distributed training jobs and researcher experiments. We have extreme require...

Posted: Aug 18, 2026 0 views
Apply Now

Software Engineer, Inference - Multi Modal

openai San Francisco full_time

About the Team

OpenAI’s Inference team powers the deployment of our most advanced models - including our GPT models, 4o Image Generation, and Whisper - across a variety of platforms. Our work ensures these models are available, performant, and scalable in production, and we...

Posted: Aug 18, 2026 0 views
Apply Now

ML Engineer, Forward Deployed

prior-labs New York full_time

Who we are

Foundation models transformed text and images. Structured data - the largest and most consequential data format in the world - stayed untouched, until now. What LLMs did for language, we're doing for tables.

We pioneered tabular foundation models: TabPFN v2 was a

Sr. AI Inference Platform Engineer

apple Seattle

We are looking for senior engineer to build tooling, automation, and analysis capabilities that strengthen our AI inference platform. This role will focus on developing sophisticated performance benchmarking systems, capacity projection models, and data analysis pipelines that directly inform our...

Posted: Aug 18, 2026 0 views

ML Engineer

ultralytics Shenzhen full_time

About Ultralytics:

At Ultralytics, we commit to relentless innovation in the AI space and seek team members who resonate with our ambition to produce the wor...

Posted: Aug 18, 2026 0 views

Staff + Sr. Software Engineer, Cloud Inference

anthropic San Francisco, United States hybrid

About Anthropic

Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and...

Posted: Aug 18, 2026 0 views
kubernetes python rust

Staff Software Engineer - ML Infrastructure

watney San Francisco full_time

Our Mission

Expand human ambition in the physical world.

Critical infrastructure is constrained by labor shortages, hazardous working conditions, and operational complexity. Watney builds and deploys autonomous robotic systems that increase the speed and...

Posted: Aug 18, 2026 0 views
Previous Page 14 of 72 Next