AI Specialist (AI Engineering)
✨ AI Summary
hyphen-connect-limited is hiring an AI Specialist Engineer in Singapore to optimize large language and vision models for on-device inference. The role focuses on model compression, distillation, and hardware-specific deployment across NPU/GPU architectures. Tech stack includes TensorRT, ONNX Runtime, C++, and Python. Requires hands-on expertise in quantization (4-bit/8-bit), pruning, and edge deployment.
We are looking for an AI Specialist Engineer to enhance the performance of large language and vision models for on-device inference. Your expertise will be crucial in developing and deploying cutting-edge AI solutions, ensuring optimal efficiency across diverse hardware architectures.
Responsibilities:
- Compress and optimize large language and vision models for on-device inference.
- Develop pipelines for model distillation and hardware-specific compilation.
- Benchmark performance across various NPU/GPU architectures.
Qualifications:
- Expertise in model distillation, pruning, and 4-bit/8-bit quantization techniques.
- Hands-on experience with TensorRT, ONNX Runtime, and edge deployment.
- Strong C++ and Python skills.
About the Company
No detailed information available about this company.
More jobs at Hyphen Connect Limited
-
Web3 Frontend Developer (UX)
APAC · Contract · Aug 19, 2026
-
Smart Contract Developer (DEX) - Remote/ Mandarin Speaking
APAC · Contract · Aug 19, 2026
-
Founding Mobile Development Lead (Game/ React Native/ Mandarin speaking)
APAC · Contract · Aug 19, 2026
-
MLOps & Agentic Platform Engineer (AI Infrastructure)
San Francisco Bay Area, USA · · Aug 19, 2026
-
MLOps & Agentic Platform Engineer (AI Infrastructure)
Oregon, USA · · Aug 19, 2026