AI Specialist (AI Engineering)

Hyphen Connect Limited

Apply to this job
Seattle, US Until 8/21/2026 First posted April 25, 2026 Last posted April 25, 2026
Job description

We are looking for an AI Specialist Engineer to enhance the performance of large language and vision models for on-device inference. Your expertise will be crucial in developing and deploying cutting-edge AI solutions, ensuring optimal efficiency across diverse hardware architectures.

Responsibilities:

  • Compress and optimize large language and vision models for on-device inference.
  • Develop pipelines for model distillation and hardware-specific compilation.
  • Benchmark performance across various NPU/GPU architectures.

Qualifications:

  • Expertise in model distillation, pruning, and 4-bit/8-bit quantization techniques.
  • Hands-on experience with TensorRT, ONNX Runtime, and edge deployment.
  • Strong C++ and Python skills.

 

 

About this role

Summary

Optimize large language and vision models for on-device inference across hardware architectures

Job title

AI Specialist (AI Engineering)

Experience level

not specified

Industry

software

Location requirements

Seattle, USA, on-site; remote not specified

Salary

Not specified

Management role

No

Skills & keywords

Required skills

model distillationpruningquantizationTensorRTONNX RuntimeC++Python

Preferred skills

edge deploymentperformance benchmarkinghardware architectures

Specializations

AImachine learningedge deployment
Locations

Structured locations inferred from the posting.

Seattle, WA, USA

On-site City