Search Results for "engineer-inference-optimizations"

Found 1032 jobs

Member of Technical Staff (Software Engineer)

cerebras Headquarters/Sunnyvale Office full_time

Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.

This order of magnitude increase in s...

Posted: Aug 18, 2026 0 views
Apply Now

Senior Machine Learning Engineer

claritypay New York City full_time

About Us:

We give businesses and their customers peace of mind by solving complex credit challenges with precision, speed, and intelligence, combining deep expertise with advanced technology, to simplify the experience and deliver better outcomes, every time.

We'r...

Posted: Aug 18, 2026 0 views
Apply Now

AI Infrastructure Engineer

meshy Bay Area Office full_time

About Meshy

Headquartered in Silicon Valley, Meshy is the leading 3D generative AI company on a mission to Unleash 3D Creativity by transforming the content creation pipeline. Meshy makes it effortless for both professional artists and hobbyists to create unique 3D assets...

Posted: Aug 18, 2026 0 views
Apply Now

Software Engineer – GPU Kernel

friendliai San Francisco full_time

About the job

FriendliAI is looking for a GPU Kernel Engineer to design, build, and optimize the low-level compute kernels that power our large-scale, GPU-accelerated AI inference platform. You will be delivering world-class inference speed across NVIDIA and AMD GPUs. Wit...

Posted: Aug 18, 2026 0 views
Apply Now

Software Engineer – GPU Kernel

friendliai Seoul full_time

About the job

FriendliAI is looking for a GPU Kernel Engineer to design, build, and optimize the low-level compute kernels that power our large-scale, GPU-accelerated AI inference platform. You will be delivering world-class inference speed across NVIDIA and AMD GPUs. Wit...

Posted: Aug 18, 2026 0 views
Apply Now

AI/ML Infrastructure Engineer

zensors San Francisco full_time

The AI Infrastructure team at Zensors builds the engine that powers our visual sensing platform. We provide the tools to automate the lifecycle of our AI workflow, including model development, evaluation, optimization, deployment, and monitoring across thousands of video streams.

As a

Posted: Aug 18, 2026 0 views

Senior Solutions Architect

fal-ai San Francisco full_time

fal is the generative media ecosystem powering the next generation of AI products. We build the infrastructure, tools, and model access that teams need to move from idea to production, and do it at scale without compromise. For developers and enterprises, fal is the foundation that makes generati...

Posted: Aug 18, 2026 0 views

Software Engineer - GPU Kernels

baseten San Francisco full_time

ABOUT BASETEN

Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enab...

Posted: Aug 18, 2026 0 views

Software Engineer, GenAI Platform

deliveroo London - The River Building HQ full_time

Software Engineer, GenAI Platform

About the Team

Deliveroo's GenAI Platform team sits within Machine Learning Platform and builds the shared infrastructure that helps DoorDash, Wolt, and Deliveroo teams safely bring GenAI-powered produc...

Posted: Aug 18, 2026 0 views

Software Engineer, Machine Learning Infrastructure

deliveroo London - The River Building HQ full_time

Software Engineer, Machine Learning Infrastructure - Generative AI

About the Team

Deliveroo's GenAI Platform team sits within Machine Learning Platform and builds the shared infrastructure that helps DoorDash, Wolt, and Deliveroo teams...

Posted: Aug 18, 2026 0 views
Previous Page 14 of 104 Next