Search Results for "engineer-inference-optimizations"
Found 1032 jobs
Staff + Sr. Software Engineer, Cloud Inference Launch Engineering
About Anthropic
Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and...
Software Engineer, Productivity - Inference Runtime
About the Team
We’re hiring a Developer Productivity engineer to support OpenAI’s Inference Runtime teams. These teams own the systems responsible for serving models reliably, efficiently, and safely across Codex, ChatGPT, API, and internal research workloads. We’re hirin...
Software Engineer - Baseten Inference Stack
ABOUT BASETEN
Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enab...
ML Engineer - Inference & Model Deployment
Job discovery is broken. Indeed and LinkedIn want to keep it that way. Join our team and help millions of people put food on the table by finding the perfect job.
HiringCafe.com is building a 100× better job search engine — fast, comprehen...
Sr. Machine Learning Engineer, Foundation Models Inference - Cloud OS & Inference
We are the Foundation Model Inference team within Cloud OS and AI Inference organization. We are on a mission to build the most highly performant, secure and private inference stack that powers Siri AI, Apple Intelligence and Apps that are powered with the largest foundation models.
Our sy...
Software Engineer, GPU Inference
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.
This order of magnitude increase in s...
Staff Software Engineer, GPU Inference
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.
This order of magnitude increase in s...
Staff Software Engineer: AI Inference Data Plane
Dive in and do the best work of your career at DigitalOcean. Journey alongside a strong community of top talent who are relentless in their drive to build the simplest scalable cloud. If you have a growth mindset, naturally like to think big and bold, and are energized by the fast-paced environme...
Staff Software Engineer: AI Inference Data Plane
Dive in and do the best work of your career at DigitalOcean. Journey alongside a strong community of top talent who are relentless in their drive to build the simplest scalable cloud. If you have a growth mindset, naturally like to think big and bold, and are energized by the fast-paced environme...
Machine Learning Engineer - Model Optimization
We are seeking experienced ML engineers to optimize and deploy machine learning models on heterogeneous embedded computing platforms. You will work at the intersection of machine learning, compilers, runtime systems, and computer architecture, helping bridge the gap between model...