Search Results for "software-engineer-gpu-inference"
Found 861 jobs
Infrastructure Engineer (Storage)
Who We Are
Lightning AI is the company behind PyTorch Lightning. Founded in 2019, we build an end-to-end platform for developing, training, and deploying AI systems—designed to take ideas from research to production with less friction.
Through our merger with Voltage Park, a neocloud and AI Factory, Li…
AI Infrastructure Engineer
Fin is the AI Customer Agent company on a mission to help businesses provide perfect customer experiences.
Our AI Agent Fin is the highest-performing AI Customer Agent on the market today, enabling businesses to deliver impeccable, always-on customer support across the customer journey – from service,…
AI Infrastructure Engineer
Fin is the AI Customer Agent company on a mission to help businesses provide perfect customer experiences.
Our AI Agent Fin is the highest-performing AI Customer Agent on the market today, enabling businesses to deliver impeccable, always-on customer support across the customer journey – from service,…
AI Infrastructure Engineer
Fin is the AI Customer Agent company on a mission to help businesses provide perfect customer experiences.
Our AI Agent Fin is the highest-performing AI Customer Agent on the market today, enabling businesses to deliver impeccable, always-on customer support across the customer journey – from service,…
Cluster Operations Software Engineer
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.
This order of magnitude increase in speed is transformi…
Software Engineer, Hardware Health
About the Team
The Hardware Health and Observability team owns the end-to-end health lifecycle of OpenAI’s global compute fleet.
Our mission is to maximize healthy, usable compute across accelerator vendors, generations, cloud providers, and regions through reliable health signals, automated remediati…
Staff AI Infrastructure Engineer
About Luma AI
Engineering Manager, Kernel Reliability
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.
This order of magnitude increase in speed is transformi…
Software Engineer, Infrastructure
Scaled Cognition is the world’s only model lab dedicated exclusively to customer experience and pioneering agentic models purpose-built for reliable action-taking enterprise applications. Backed by Khosla Ventures, the company’s flagship Agentic Pretrained Transformer (APT) eliminates hallucinations…
Senior QA engineer (AI Platform)
About Gruve
Gruve is an innovative software services startup dedicated to transforming enterprises to AI powerhouses. We specialize in cybersecurity, customer experience, cloud infrastructure, and advanced technologies such as Large Language Models (LLMs). Our mission is to assist our customers in th…