AI Specialist (AI Engineering)
✨ AI Summary
hyphen-connect-limited is hiring an AI Specialist Engineer in China to optimize large language and vision models for on-device inference. The role focuses on model compression, distillation, and hardware-specific deployment across NPU/GPU architectures. Required tech stack includes TensorRT, ONNX Runtime, C++, and Python, with expertise in quantization (4-bit/8-bit), pruning, and edge deployment. No specific years of experience or seniority level are stated.
We are looking for an AI Specialist Engineer to enhance the performance of large language and vision models for on-device inference. Your expertise will be crucial in developing and deploying cutting-edge AI solutions, ensuring optimal efficiency across diverse hardware architectures.
Responsibilities:
- Compress and optimize large language and vision models for on-device inference.
- Develop pipelines for model distillation and hardware-specific compilation.
- Benchmark performance across various NPU/GPU architectures.
Qualifications:
- Expertise in model distillation, pruning, and 4-bit/8-bit quantization techniques.
- Hands-on experience with TensorRT, ONNX Runtime, and edge deployment.
- Strong C++ and Python skills.
About the Company
No detailed information available about this company.
More jobs at Hyphen Connect Limited
-
Web3 Frontend Developer (UX)
APAC · Contract · Aug 19, 2026
-
Smart Contract Developer (DEX) - Remote/ Mandarin Speaking
APAC · Contract · Aug 19, 2026
-
Founding Mobile Development Lead (Game/ React Native/ Mandarin speaking)
APAC · Contract · Aug 19, 2026
-
MLOps & Agentic Platform Engineer (AI Infrastructure)
San Francisco Bay Area, USA · · Aug 19, 2026
-
MLOps & Agentic Platform Engineer (AI Infrastructure)
Oregon, USA · · Aug 19, 2026