Skip to main content
D

Embedded AI Engineer: On-Device Models & Edge Kernels

Deepgram
2 hours ago
Full-time
On-site
San Francisco, California, United States
Deepgram is seeking an Embedded AI Engineer to work on the Partner Platform Engineering team, tackling the lowest layer of the edge stack. You will write and optimize custom kernels and operators for diverse hardware, including embedded SoCs, DSPs, and NPUs, enabling Deepgram models to run on non-NVIDIA accelerators. Your work will involve quantization, operator fusion, and architecture-specific compilation, with collaboration to fit models to constrained devices and delivery of reusable runtime

#J-18808-Ljbffr