Search Results for "software-engineer-inference-runtime"
Found 142 jobs
Principal SRE - AI Inference
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.
This order of magnitude increase in s...
Machine Learning Engineer - Model Optimization
We are seeking experienced ML engineers to optimize and deploy machine learning models on heterogeneous embedded computing platforms. You will work at the intersection of machine learning, compilers, runtime systems, and computer architecture, helping bridge the gap between model...
Senior AI Engineer - Services Special Projects
At Apple, great ideas turn into phenomenal products, services, and customer experiences at a pace few companies can match.
We are seeking a highly experienced ML Engineer to build, deploy, optimize and operationalize Small and Large Language Model (LLM)-based applications, with a strong...
Staff/Principal Machine Learning Engineer
About Us
Epia Neuro is a neural technology company developing intent-driven systems that restore function and independence for people living with neurological conditions. Our platform integrates implantable neural interfaces, adaptive algorithms, and assistive devices to tr...
AI Infrastructure Engineer
About Meshy
Headquartered in Silicon Valley, Meshy is the leading 3D generative AI company on a mission to Unleash 3D Creativity by transforming the content creation pipeline. Meshy makes it effortless for both professional artists and hobbyists to create unique 3D assets...
Embedded AI Engineer, On-Device Models
Company Overview
Deepgram is the leading platform underpinning the emerging trillion-dollar Voice AI economy, providing real-time APIs for speech-to-text (STT), text-to-speech (TTS), and building production-grade voice agents at scale. More than 200,000 developers and 1,3...
Embedded AI Engineer, On-Device Models
Company Overview
Deepgram is the leading platform underpinning the emerging trillion-dollar Voice AI economy, providing real-time APIs for speech-to-text (STT), text-to-speech (TTS), and building production-grade voice agents at scale. More than 200,000 developers and 1,...
Senior Engineering Manager, NKP
Hungry, Humble, Honest, with Heart
The Opportunity At Nutanix, we are expanding the capabilities of the Nutanix Kubernetes Platform (NKP) to power the next generation of enterprise AI. This platform will serve as the foundation for AI/ML workloads, GPU infrastructure, and...
Engineering Manager, Continuous Deployment and Change Safety
About Anthropic
Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working...
Senior Application Security Engineer, AI and Machine Learning
Who We Are
Lightning AI is the company behind PyTorch Lightning. Founded in 2019, we build an end-to-end platform for developing, training, and deploying AI systems—designed to take ideas from research to production with less friction.
Through our merger with Volt...