Search Results for "ml-engineer-inference-and-optimization"
Found 716 jobs
Senior Applied ML Engineer
Why work at Higgsfield AI?
Higgsfield AI is the fastest-scaling generative AI company in history, hitting $500M in annual revenue run rate, 25M+ users worldwide, 6M+ generations per day, and powering 390 of Fortune 500 brands.
We're building at the absolute fron...
Distributed Systems Engineer, Data & Inference Platform
The Role
You'll build and operate the systems that turn raw compute into useful intelligence — the inference services that serve LLMs at scale and the data pipelines that feed them. One week you're hunting a tail-latency regression in a production inference servic...
ML Infrastructure Engineer
TLDR: We are looking for an ML Infrastructure Engineer to build the systems behind our LLM post-training, RL, evaluation, inference, and agentic development workflows. You will work close to researchers, GPUs, training loops, data control systems, evals, inference stacks, and the infrastr...
ML Platform Engineer
Hadrian - Manufacturing the Future
Hadrian is building autonomous factories to reindustrialize America. By combining AI, advanced software, robotics, and full-stack manufacturing, we help aerospace and defense companies build rockets, satellites, aircraft, ships, and othe...
Staff ML Software Engineer (L6) — Platform Systems, AIMS Engineering
At Netflix, our mission is to entertain the world. Together, we are writing the next episode - pushing the boundaries of storytelling, global fandom and making the unimaginable a reality. We are a dream team obsessed with the uncomfortable excitement of discovering what happens when you merge cre...
Principal Engineer, Inference Cloud
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.
This order of magnitude increase in s...
Principal Machine Learning Engineer - Revenue Optimisation
📈 Who We Are:
We are rebuilding the energy transaction, making it transparent and fair.
Our goal is to put power back where it belongs, in the hands of customers and to take on one of the most critical problems of our century, access to low cost elec...
ML Compute Efficiency Automation Engineer, Infrastructure & Planning
Apple’s Platform Acceleration & Compute Efficiency (PACE) is a high-leverage team operating at the intersection of our ML organizations, underlying compute infrastructure, and core platform tooling. Our mission is to empower Apple’s software engineering teams with efficient, scalable compute....
Software Engineer, Model Inference
About the Team
Our Inference team brings OpenAI’s most capable research and technology to the world through our products. We empower consumers, enterprise and developers alike to use and access our start-of-the-art AI models, allowing them to do things that they’ve never be...
ML Infrastructure Engineer
Join Us in Building the Future of Home Robotics
At Sunday, we're developing personal robots to reclaim the hours lost to repetitive tasks. We're focused on an ambitious goal to make generalized robots broadly accessible, enabling households to take back quality time...