Search Results for "inference-infrastructure-engineer-serving"
Found 854 jobs
Senior Site Reliability Engineer
Software Engineer - Senior Backend
About the Job
We believe using large language and multimodal models should be as simple as calling an API. To achieve this in production, we need to serve enterprises across clouds, with authentication, billing, multi-tenant isolation, and zero tolerance for downtime.
We are looking for a Senior Backe…
Software Engineer – Senior Backend
About the Job
We believe using large language and multimodal models should be as simple as calling an API. To achieve this in production, we need to serve enterprises across clouds, with authentication, billing, multi-tenant isolation, and zero tolerance for downtime.
We are looking for a Senior Backe…
Staff + Sr. Software Engineer, Cloud Inference Launch Engineering
About Anthropic
Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working toge…
Sr. Engineering Manager, Inference
AI DevOps Engineer
Seeking a Lead AI DevOps Engineer to oversee design and delivery of advanced AI/ML/GenAI solutions. The role combines cloud engineering and automation with hands-on leadership in deploying and integrating LLM/SLM models into enterprise applications, ensuring security, scalability, and operational ex…
AI Software Engineer (Back End)
About the role
Maincode is training Matilda, a large language model built and trained from scratch in Australia. Our new compute cluster is now live, and we are scaling the next version and deploying it publicly.
This role sits inside the production system that serves Matilda. You will build and maint…
Sr. AI Inference Platform Engineer
We are looking for senior engineer to build tooling, automation, and analysis capabilities that strengthen our AI inference platform. This role will focus on developing sophisticated performance benchmarking systems, capacity projection models, and data analysis pipelines that directly inform our AI…
Platform Engineer — Core Infrastructure
About Us
100ms operates two product lines at scale: a real-time Live Video platform powering latency-sensitive, high-concurrency video experiences, and an AI Agents platform that automates complex patient access workflows in U.S. healthcare.
Both products run on a shared, robust infrastructure foundation.…
Senior Site Reliability Engineer
Senior Site Reliability Engineer
Location: Global Remote / San Francisco · Full-Time
About Andromeda
Andromeda gives AI companies access to the kind of scaled compute once reserved for hyperscalers. Our platform connects 100+ AI customers to 50+ global providers, with billions of GPU-hours supported, a…