AI Specialist (AI Engineering)
Hyphen Connect Limited
Apply to this job Oregon, US Until 8/21/2026 First posted April 25, 2026 Last posted April 25, 2026
Job description
We are looking for an AI Specialist Engineer to enhance the performance of large language and vision models for on-device inference. Your expertise will be crucial in developing and deploying cutting-edge AI solutions, ensuring optimal efficiency across diverse hardware architectures.
Responsibilities:
- Compress and optimize large language and vision models for on-device inference.
- Develop pipelines for model distillation and hardware-specific compilation.
- Benchmark performance across various NPU/GPU architectures.
Qualifications:
- Expertise in model distillation, pruning, and 4-bit/8-bit quantization techniques.
- Hands-on experience with TensorRT, ONNX Runtime, and edge deployment.
- Strong C++ and Python skills.
About this role
Summary
Optimize AI models for on-device deployment; develop pipelines; benchmark across hardware architectures.
Job title
AI Specialist (AI Engineering)
Experience level
not specified
Industry
software
Location requirements
Oregon, USA, on-site; remote not specified
Salary
Not specified
Management role
No
Skills & keywords
Required skills
C++PythonTensorRTONNX Runtimemodel distillationquantization
Preferred skills
NPUGPUmodel compressionhardware optimization
Specializations
AImachine learningedge deployment
Locations
Structured locations inferred from the posting.
Unknown location
Work arrangement unknown
Related searches