Search Results for "engineer-inference-optimizations"
Found 1032 jobs
Senior Software Engineer, AI Infrastructure - LVM Inference & Evaluation
Build a safer world with us, one incident at a time.
Ambient.ai is the category creator and leader in Agentic Physical Security. Powered by Ambient Pulsar, the first reasoning Vision-Language Model purpose-built for physical security, our platform seamlessly integrates with...
Staff Software Engineer, Inference
About Anthropic
Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and...
Software Engineer, Inference Runtime
LM Studio is used by millions of people around the world to run AI on their own computers, and now with Bionic - also in the cloud. Our values prioritize putting the human in the center, and creating tools that we want to use ourselves, and recommend to our friends and family.
As a team, we...
Machine Learning Engineer - Inference / Serving
Yobi is a rapidly growing Behavioral AI company on a mission to ethically democratize the benefits of data and AI.
Since 2019, we have built one of the largest consented behavioral datasets in the United States, extending far beyond the walled garden...
Machine Learning Infrastructure Engineer, Model Inference
About Abridge
Abridge was founded in 2018 with the mission of powering deeper understanding in healthcare. Our AI-powered platform was purpose-built for medical conversations, improving clinical documentation efficiencies while enabling clinicians to focus on what matters...
Software Engineer - Performance Optimization
Applied Intuition, Inc. is powering the future of physical AI. Founded in 2017 and now valued at $15 billion, the Silicon Valley company is creating the digital infrastructure needed to bring intelligence to every moving machine on the planet. Applied Intuition services the automotive, defense, t...
AI Inference Infrastructure Software Engineer (Kubernetes / Cloud)
Location: Seattle, WA (Hybrid - 3 days/week in office)
About ElastixAI:
ElastixAI is an early-stage Software startup on a mission to reinvent AI inference infrastructure from the ground up. We're building a next-generation inference platfor...
Software Engineer, Model Inference
About the Team
Our Inference team brings OpenAI’s most capable research and technology to the world through our products. We empower consumers, enterprise and developers alike to use and access our start-of-the-art AI models, allowing them to do things that they’ve never be...
Software Engineer: ML Optimization
About the Role
We internally call this team MBMB (More Big More Better). You will own optimizations on both the training and on-robot inference stacks. We are still in a regime of step-function, not incremental, gains.
You’ll be responsible for:
Principal Engineer, Inference Cloud
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.
This order of magnitude increase in s...