Search Results for "inference-infrastructure-engineer-serving"
Found 848 jobs
Machine Learning Engineer
Role Overview:
As a Machine Learning Engineer, you will play a central role in translating cutting-edge machine learning research into scalable, production-ready solutions. You will collaborate closely with cross-functional teams to identify opportunities where ML can drive product value, architect r…
Senior Platform Engineer
About Nomic
Nomic is the domain-specific AI platform for the Architecture, Engineering, and Construction (AEC) industry. We help enterprise teams extract structured knowledge from decades of drawings, specs, and project files — combining embedding models, document parsing, and autonomous agents that…
Software Engineer, Inference
About Luma AI
Staff Software Development Test Engineer
About Tekion:
Positively disrupting an industry that has not seen any innovation in over 50 years, Tekion has challenged the paradigm with the first and fastest cloud-native automotive platform that includes the revolutionary Automotive Retail Cloud (ARC) for retailers, Automotive Enterprise Cloud (A…
Staff Software Engineer, Inference Platform
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.
This order of magnitude increase in speed is transformi…
Senior Platform Engineer
Trust is the first of a new breed of banks in Singapore – digitally native and focused on delivering a delightful customer experience. You will work in a fast-paced and collaborative environment to solve new and interesting challenges each day. Together with our Trust team, you will help shape the f…
Senior Software Engineer - Infrastructure Storage
Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligen…
Staff Software Engineer, Inference Cloud
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.
This order of magnitude increase in speed is transformi…
Software Engineer, Inference - Multi Modal
About the Team
OpenAI’s Inference team powers the deployment of our most advanced models - including our GPT models, 4o Image Generation, and Whisper - across a variety of platforms. Our work ensures these models are available, performant, and scalable in production, and we partner closely with Resea…
Staff Machine Learning Engineer
The Opportunity
Do you want to lead projects to build and deploy cutting-edge AI technology to help people get unparalleled value from meetings and conversations? Join our core AI team responsible for ML and work alongside industry-veteran scientists and engineers. As a Staff Machine Learning Enginee…