True Talent Solutions Partners logo
Salary
$180K–$230K
Moves you to
United States
Support
Visa sponsorship
Posted
Is this job info correct?
Show job description

About the Role

Join an early-stage industrial AI team as a founding researcher building the multimodal AI systems behind wearable assistants for field technicians. You will help develop and deploy real-time vision-language and voice capabilities for operational workflows in critical infrastructure.

What You'll Do

  • Build and ship production agentic vision-language pipelines that support industrial procedures and inspections.

  • Develop evaluation tools, capture model failure modes, and use field data to guide ongoing model improvements.

  • Adapt model orchestration and edge inference to changing connectivity, latency, and runtime constraints.

  • Create real-time voice and video interfaces, including conversational and proactive alert experiences.

  • Fine-tune and optimize multimodal models for enterprise deployments, including on-premise use.

What We're Looking For

  • 0 to 3 years of relevant experience and strong machine learning and computer science fundamentals.

  • Experience building or shipping agentic or multimodal AI systems for real users.

  • Familiarity with Python, PyTorch, Hugging Face Transformers, vision-language models, and multimodal LLMs.

  • Working knowledge of model serving and optimization tools such as vLLM, Triton, Ray Serve, ONNX, TensorRT, and Docker.

  • Exposure to fine-tuning methods such as SFT and RLHF, quantization, RAG, and edge AI; interest in computer vision, wearables, or industrial applications is valuable.

Compensation & Benefits

Salary range: $180,000 to $230,000 USD annually. Visa sponsorship is available.

Location

On-site in San Francisco, United States.

Similar jobs

Apply for this job