Search Results for "engineer-inference-optimizations"

Found 1958 jobs

Senior MLOps Engineer - Edge

hudl Barcelona, Spain; London, United Kingdom; Spain (Remote); United Kingdom (Remote) Remote

At Hudl, we build great teams. We hire the best of the best to ensure you’re working with people you can constantly learn from. You’re trusted to get your work done your way while testing the limits of what’s possible and what’s next. We work hard to provide a culture where everyone feels supported,…

Posted: Aug 19, 2026 0 views
Apply Now

Principal Machine Learning Engineer- LLM Fine-tuning and Optimization

airbnb United States remote

Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences that make it possible fo…

Posted: Aug 18, 2026 0 views
python pytorch
Apply Now

Senior Staff Machine Learning Engineer, LLM/VLM Model Architecture & Optimization

waymo Mountain View, CA, USA; San Francisco, CA, USA; Full-time

Waymo is an autonomous driving technology company with the mission to be the world's most trusted driver. Since its start as the Google Self-Driving Car Project in 2009, Waymo has focused on building the Waymo Driver—The World's Most Experienced Driver™—to improve access to mobility while saving tho…

Posted: Aug 19, 2026 0 views
Apply Now

Senior ML Engineer (Token Factory)

nebius Amsterdam, Netherlands

About Nebius:

Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-hous…

Posted: Aug 19, 2026 0 views
Apply Now

Machine Learning Engineer – Edge AI & On-Device Optimization

apple Herzliya

Join our team as a Machine Learning Engineer and help shape the future of on-device AI. You'll research, design, and deploy cutting-edge deep learning models optimized for Apple silicon edge devices, working across the full ML lifecycle alongside hardware, software, and product teams.

We are looking…

Posted: Aug 18, 2026 0 views
Apply Now

Senior Software Engineer, Infrastructure

decagon San Francisco On-site

About Decagon

Decagon is the leading conversational AI platform empowering every brand to deliver concierge customer experiences.

Our technology enables industry-defining enterprises like Avis Budget Group, Block’s Cash App and Square, Chime, Oura Health, and Hunter Douglas to deploy AI agents that po…

Posted: Apr 12, 2026 8 views
Senior Software Engineer Infrastructure Conversational AI Platform Networking Data ML Serving Developer Platform Real-time Voice SLOs Low Latency High Scale Terraform GitOps Kubernetes GCP AWS Azure Observability CI/CD Full Time On-site
Apply Now

Senior ML Engineer (Token Factory)

nebius Germany; Israel; Netherlands; Prague, Czech Republic; Remote - Europe; United Kingdom Remote

About Nebius:

Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-hous…

Posted: Aug 19, 2026 0 views
Apply Now

AI Infrastructure / ML Engineer

capacloud United States remote

We are hiring an AI Infrastructure / ML Engineer to help optimize CapaCloud for AI workloads, model deployment, training, inference, and scalable GPU utilization.

You will work closely with infrastructure engineers to ensure the platform supports modern AI workflows for startups, researchers, and ent…

Posted: Aug 18, 2026 0 views
cuda docker jax kubernetes python pytorch tensorflow
Apply Now

Member of Technical Staff (Software Engineer)

cerebras Headquarters/Sunnyvale Office full_time

Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.

This order of magnitude increase in speed is transformi…

Posted: Aug 18, 2026 0 views
Apply Now

Expression of Interest: Machine Learning Engineer

moloco Menlo Park, California, United States; New York, New York, United States; Seattle, Washington, United States

About Moloco:

Moloco builds some of the most powerful AI advertising solutions in the world. Our name—short for "machine learning company"—reflects our core mission: democratizing access to the advanced AI that has historically been reserved for tech giants. Led by machine learning pioneers who built…

Posted: Aug 19, 2026 0 views
Apply Now
Previous Page 19 of 196 Next