Join a stealth, seed-stage startup building an AI-powered wearable platform for industrial field technicians. You'll own the AI core: an agentic vision-language model (VLM) system doing multimodal visual reasoning, evals, and model orchestration, running on real hardware against real industrial workflows β not benchmark demos.
What you'll do
Build and ship agentic VLM systems that reliably do visual reasoning in production (detection, segmentation, image/video understanding)
Own model orchestration and build real evals discipline (ground-truth, trajectory, or regression harnesses)
Work directly with the founding team, shaping the AI roadmap from day one
Location & culture Fully on-site in the SF Bay Area, initial team-house period, intense/high-ownership in-person culture (broadly 9-9-6).