Founding ML Researcher
- Salary
- $180K–$230KUSD per year
- Moves you to
- United States
- Support
- Visa sponsorship
- Posted
- Oct 2, 2026
About the Role
Join the founding engineering team at an industrial wearable AI startup and build the multimodal AI systems behind smart-glasses copilots for field technicians. You will own key parts of the applied ML stack, bringing real-time visual and voice assistance into demanding industrial workflows.
What You'll Do
Build and ship agentic vision-language model pipelines that support multi-step visual reasoning and tool use in inspection and standard operating procedure workflows.
Develop evaluation frameworks, capture failure modes, and create data and fine-tuning loops that improve model quality with each release.
Adapt inference to changing connectivity and latency constraints through model orchestration, runtime optimization, and graceful fallbacks.
Create real-time voice and video interfaces, including conversational and proactive alert experiences for different user needs.
Train and fine-tune multimodal models for enterprise deployments, including on-premise use.
What We're Looking For
0 to 3 years of relevant experience, with a track record of building and shipping agentic or multimodal systems to real users.
Experience with Python, PyTorch, multimodal LLMs, vision-language models, and tools such as Hugging Face Transformers, LangChain, and RAG.
Familiarity with model serving and optimization, including vLLM, Triton, Ray Serve, ONNX, TensorRT, Docker, and edge AI.
Knowledge of model training techniques such as SFT, RLHF, and quantization, including GPTQ or AWQ.
Interest or experience in computer vision, wearable devices, real-time AI, or applied AI for industrial workflows.
Compensation & Benefits
Salary range: $180,000 to $230,000 annually. Visa sponsorship is available.
Location
On-site in San Francisco, United States.