Search Results for "engineer-inference-optimizations"

Found 1032 jobs

Staff / Senior Software Engineer, Inference

anthropic San Francisco, CA | New York City, NY | Seattle, WA Hybrid

About Anthropic

Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working...

Posted: Apr 7, 2026 38 views
staff software engineer senior software engineer inference distributed systems machine learning LLM Kubernetes cloud infrastructure Python Rust AI Anthropic Claude
Apply Now

AI Infrastructure Engineer

intercom Dublin, Ireland Hybrid

Intercom is the AI Customer Service company on a mission to help businesses provide incredible customer experiences.

Our AI agent Fin, the most advanced customer service AI agent on the market, lets businesses deliver always-on, impeccable customer service and ultimately transform their customer e...

Posted: Apr 17, 2026 29 views
ai infrastructure engineer machine learning ml ai gpu cuda triton model training model inference python kubernetes aws llm transformer senior engineer hybrid dublin ireland
Apply Now

AI Infrastructure Engineer

intercom Berlin, Germany Hybrid

Intercom is the AI Customer Service company on a mission to help businesses provide incredible customer experiences.

Our AI agent Fin, the most advanced customer service AI agent on the market, lets businesses deliver always-on, impeccable customer service and ultimately transform their customer ex...

Posted: Apr 17, 2026 13 views
AI Infrastructure Engineer model training model inference CUDA Triton transformers LLMs Python GPU Kubernetes AWS distributed systems machine learning infrastructure
Apply Now

Staff Software Engineer, Infrastructure

decagon San Francisco Full time

About Decagon

Decagon is the leading conversational AI platform empowering every brand to deliver concierge customer experiences.

Our technology enables industry-defining enterprises like Avis Budget Group, Block’s Cash App and Square, Chime, Oura Health, and Hunter Douglas to deploy AI agents tha...

Posted: Apr 9, 2026 11 views
staff software engineer infrastructure production infrastructure SLOs low latency observability OpenTelemetry Prometheus Grafana Datadog incident response PagerDuty Kubernetes GKE EKS AKS GCP AWS Azure Terraform GitOps CI/CD networking compute storage security infrastructure-as-code data platforms ML infrastructure GPU model-serving LLM inference hybrid full time engineering
Apply Now

Backend Engineer- Inference Services

deepgram USA | Remote Remote

Company Overview

Deepgram is the leading platform underpinning the emerging trillion-dollar Voice AI economy, providing real-time APIs for speech-to-text (STT), text-to-speech (TTS), and building production-grade voice agents at scale. More than 200,000 developers and 1,300+ organizations build v...

Posted: Apr 7, 2026 9 views
backend engineer inference services rust python c++ c git unix machine learning torch audio processing speech processing networking distributed compute high performance computing
Apply Now

AI Infrastructure Engineer

intercom London, England Hybrid

Intercom is the AI Customer Service company on a mission to help businesses provide incredible customer experiences.

Our AI agent Fin, the most advanced customer service AI agent on the market, lets businesses deliver always-on, impeccable customer service and ultimately transform their customer ex...

Posted: Apr 17, 2026 9 views
AI Infrastructure Engineer model training model inference CUDA Triton transformers LLMs Python Kubernetes AWS GPU machine learning AI
Apply Now

DevOps/MLOps Engineer (ML / LLM Infrastructure)

kyivstar All Remote

We are looking for a DevOps Engineer to design, build, and operate the infrastructure behind our LLM platform. You will be responsible for keeping our ML infrastructure reliable, scalable, and efficient - from data pipelines to training and inference.

In this role, you will develop and maintain CI...

Posted: Apr 12, 2026 12 views
devops mlops engineer ml llm infrastructure gcp kubernetes gke docker ci/cd terraform ansible python bash airflow prometheus grafana cloud remote full-time
Apply Now

Staff + Sr. Software Engineer, Cloud Inference Launch Engineering

anthropic San Francisco, CA | Seattle, WA Hybrid

About Anthropic

Anthropic's mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working...

Posted: May 9, 2026 8 views
staff software engineer senior software engineer cloud inference launch engineering anthropic AI LLM inference cloud AWS GCP Azure Kubernetes Python Rust
Apply Now

Staff + Sr. Software Engineer, Cloud Inference Launch Engineering

anthropic San Francisco, CA Full-time

About Anthropic

Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working...

Posted: Jun 3, 2026 8 views
staff software engineer senior software engineer cloud inference launch engineering anthropic AI LLM inference kubernetes python rust
Apply Now

Staff+ Software Engineer, Inference Runtime

anthropic Remote-Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NY Remote-Friendly (Travel-Required)

About Anthropic

Anthropic's mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working...

Posted: Jun 13, 2026 7 views
Staff+ Software Engineer Inference Runtime Anthropic Remote San Francisco Seattle New York Rust Python GPU TPU Trainium ML Infrastructure
Apply Now
Page 1 of 104 Next