Search Results for "software-engineer-gpu-inference"
Found 861 jobs
Machine Learning Engineer - Inference Maintainer & Developer Experience
Our mission is to make the world programmable. Sight is one of the key ways we understand the world, and soon this will be true for the software we use, too.
We’re building the tools, community, and resources needed to make the world programmable with artificial intelligence. Roboflow simplifies buil…
Senior DevOps Engineer
We are hiring a Senior DevOps Engineer to join our US team to lead the deployment and scaling of AI solutions for financial workflows. As our company delivers full-stack AI applications to some of the world’s leading financial institutions, your work will be critical in ensuring these solutions run…
Senior AI Infrastructure Engineer (Zürich, 100%)
We are hiring a AI Infrastructure Engineer
⏰ Start date: ASAP
📍Zürich, Switzerland (on-site, remote not possible)
🦾 Full-time (100%)
Your role
As an AI infrastructure SWE you will build the systems that underpin our robot learning. You will work across data pipelines, internal tooling, and model deploy…
Software Engineer, SystemML - Scaling / Performance
In this role, you will be a member of the Network.AI Software team and part of the bigger DC networking organization. The team develops and owns the software stack around NCCL (NVIDIA Collective Communications Library), which enables multi-GPU and multi-node data communication through HPC-style coll…
Software Engineer, SystemML - AI Networking
In this role, you will be a member of the AI Networking Software team and part of the bigger DC networking organization. The team develops and owns the software stack around NCCL (NVIDIA Collective Communications Library), which enables multi-GPU and multi-node data communication through HPC-style c…
Member of Engineering (Pre-training / Data Engineering)
ABOUT POOLSIDE
In this decade, the world will create Artificial General Intelligence. There will only be a small number of companies who will achieve this. Their ability to stack advantages and pull ahead will define the winners. These companies will move faster than anyone else. They will attract th…
Principal Engineer, Inference Cloud
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.
This order of magnitude increase in speed is transformi…
Principal SRE - AI Inference
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.
This order of magnitude increase in speed is transformi…
Software Engineer: ML Infra
About the Role
Generalist trains very large robot foundation models. This requires utilizing very large numbers of the latest generation GPU hardware and infrastructure (currently Nvidia) to run distributed training jobs and researcher experiments. We have extreme requirements on storage and data loa…
Infrastructure Engineer (Storage)
Who We Are
Lightning AI is the company behind PyTorch Lightning. Founded in 2019, we build an end-to-end platform for developing, training, and deploying AI systems—designed to take ideas from research to production with less friction.
Through our merger with Voltage Park, a neocloud and AI Factory, Li…