About the role
Adapt Deepgram speech models to edge devices and non-NVIDIA accelerators. The work covers operator changes, quantisation, accuracy and latency benchmarking, and repeatable deployment onto constrained hardware. You will collaborate with embedded engineers on custom kernels; production Python and PyTorch experience matters.
Remote: United States.
Employer posting date: 2026-09-14.
Skills
How to apply
Apply directly on the employer's application page. Your application goes straight to them.
More from Deepgram
Jobs like this
Republished listing
This opportunity was discovered on Deepgram's public careers page and is republished here for discovery purposes. Applications are handled by the employer.
If this is your role and you want it off the board, contact us.