Search Results for "inference-infrastructure-engineer-serving"

Found 848 jobs

Senior Software Engineer - ML Infrastructure

applied Sunnyvale full_time

Applied Intuition, Inc. is powering the future of physical AI. Founded in 2017 and now valued at $15 billion, the Silicon Valley company is creating the digital infrastructure needed to bring intelligence to every moving machine on the planet. Applied Intuition services the automotive, defense, truc…

Posted: Aug 18, 2026 0 views
Apply Now

Solutions Engineer

material San Francisco, CA full_time

About the role

Own the technical strategy, execution, and delivery of customer engagements from pre-sales discovery through deployment. You'll partner with Sales to architect AI infrastructure solutions for Frontier labs, startups, and enterprise ML teams, translating customer requirements into produ…

Posted: Aug 18, 2026 0 views
Apply Now

Lead Software Platform Engineer, MLOps

tetrascience United States remote

Who We Are

TetraScience is the Scientific Data and AI Cloud company. We are catalyzing the Scientific AI revolution by designing and industrializing AI-native scientific data sets, which we bring to life in a growing suite of next gen lab data management solutions, scientific use cases, and AI-enable…

Posted: Aug 18, 2026 1 views
aws aws-bedrock databricks docker llm mlflow python typescript
Apply Now

Senior AI Engineer — Inference & Agent Systems

arcana-analytics United States

Title: Applied AI Engineer — Inference & Agent Systems

Location:
United States

What We're Building

Arcana is building AI agents that synthesize information across heterogeneous sources and deliver structured, reasoned answers in real time. The product only works if the agents are fast, reliable, and corr…

Posted: Aug 19, 2026 0 views
Apply Now

Senior AI Engineer — Inference & Agent Systems

arcana-analytics Bangalore, Hybrid

Title: Senior AI Engineer — Inference & Agent Systems

Location:
- BLR/Remote India

What We're Building

Arcana is building AI agents that synthesize information across heterogeneous sources and deliver structured, reasoned answers in real time. The product only works if the agents are fast, reliable, and…

Posted: Aug 19, 2026 0 views
Apply Now

Staff Software Engineer, Machine Learning Platform

stripe Toronto

Who we are

About Stripe

Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world's largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the G…

Posted: Aug 19, 2026 0 views
Apply Now

Staff Software Engineer, Machine Learning Platform

stripe San Francisco, Seattle

Who we are

About Stripe

Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world's largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the G…

Posted: Aug 19, 2026 0 views
Apply Now

Senior Backend/Platform Engineer

typesafe-ai San Francisco Office full_time

About TypeSafe

TypeSafe is a frontier model lab. We build reliable and general AI systems to power economically valuable automation. Our mission is to usher in a new era of Transformative Artificial Intelligence (TAI): technology with the power to drive a societal shift on the scale of the agricultur…

Posted: Aug 18, 2026 0 views
Apply Now

Site Reliability Engineer

nebius Remote - United States Remote

About Nebius:

Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-hous…

Posted: Aug 19, 2026 0 views
Apply Now

Principal Engineer, Inference Cloud

cerebras Headquarters/Sunnyvale Office full_time

Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.

This order of magnitude increase in speed is transformi…

Posted: Aug 18, 2026 0 views
Apply Now
Previous Page 28 of 85 Next