ML Engineer, Audio

New York City on site Until 8/23/2026 First posted March 18, 2026 Last posted March 18, 2026
Job description

About Sandbar

Sandbar is an interface company in New York City. We aim to augment individuals so we can each think, act, and move more freely. Our team has built SW, ML, and HW products across Meta, CTRL-labs, Google, Apple, Fitbit, Peloton, and Equinox.

Our first product, Stream, is a self extension—a private voice ring and conversational interface. Stream has been featured in WSJ, Bloomberg, & Wired, and begins shipping in Summer '26.

Join us in creating technology that extends human thinking.

About

We’re looking for a machine learning engineer to help build Stream, a new conversational computer. As a machine learning engineer, you will develop a system which spans voice, memory, and agentic control. This role is perfect for someone cares deeply about real-world ML deployment and human-in-the-loop agentic interactions.

Responsibilities

  • Develop, evaluate, and deploy audio models spanning cloud models to resource-constrained on-device settings

  • Optimize inference pipelines for latency, reliability, and concurrency

  • Work closely with other ML/AI engineers, infrastructure engineers, designers, and cofounders on existing and future products

Qualifications

  • 4+ years in machine learning experience

  • Experience shipping ML-based products is required

  • Experience developing audio / voice models is required

  • Passion for human-computer interaction

FTE Benefits

  • Health, vision, and dental benefits

  • Company-sponsored 401(k)

  • Unlimited PTO and sick time

  • Early stage equity

Compensation (from employer):
$180K – $250K • Offers Equity

About this role

Summary

Build and optimize audio voice models for real-world ML deployment in a conversational interface.

Job title

ML Engineer, Audio

Experience level

4+ years

Industry

software

Location requirements

On-site in New York City, remote work not specified.

Salary

$180K – $250K

Management role

No

Skills & keywords

Required skills

machine learningaudio modelsvoice models

Preferred skills

human-computer interaction

Specializations

audio modelsvoice modelshuman-computer interaction
Locations

Structured locations inferred from the posting.

New York, NY, USA

On-site City