Search Results for "software-engineer-inference-runtime"
Found 142 jobs
Senior Software Engineer, Middleware - Special Projects
Generative intelligence is redefining what’s possible in Apple products. We’re looking for engineers who understand the entire system — from hardware to UI — and are inspired by building the connective layers that make it all work seamlessly. You’ll help shape how Apple’s intelligent systems perf...
AI Software Engineer (Back End)
About the role
Maincode is training Matilda, a large language model built and trained from scratch in Australia. Our new compute cluster is now live, and we are scaling the next version and deploying it publicly.
This role sits inside the production system that serv...
Senior Software Engineer
About Varick. Varick builds AI agents that take over real operational workflows inside the world's largest enterprises. Our forward-deployed engineers and strategists embed with client teams, map how work actually runs, then our engineers build the agents into production. We&...
Software Engineer - Forecasting & Scheduling
About Assembled
Great customer support requires human agents and AI in perfect balance, and Assembled is the only unified platform that orchestrates both at scale. Companies like Canva, Etsy, and Robinhood use Assembled to c...
Software Engineer - Forecasting & Scheduling
About Assembled
Great customer support requires human agents and AI in perfect balance, and Assembled is the only unified platform that orchestrates both at scale. Companies like Canva, Etsy, and Robinhood use Assembled to...
Sr. Forward Deployed Software Engineer - Dubbing Platform
About Sarvam
Sarvam is building the bedrock of Sovereign AI for India. The company is developing India’s full-stack sovereign AI platform, building across research, models, infrastructure and applications with a singular focus on making AI genuinely work for India. Sarvam w...
AI Inference Core - SDET Technical Lead, Release Integration Testing
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.
This order of magnitude increase in s...
AI Inference Core - Junior SDET, Release Integration Testing
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.
This order of magnitude increase in s...
Systems Engineer - Simulation Correctness
The Mission
At Vinci, we are building the operator intelligence infrastructure that modern hardware programs rely on daily. We have already proven that a single foundation model works out of the box across physics on realistic production workloads.
Trained o...
Member of Technical Staff, Full-Stack
Highlight is building a shared intelligence layer for the modern workforce. Highlight unifies context across every person and tool on your team, seamlessly bridging information silos. It evolves with your organization to proactively route knowledge and reliably automate workflows.