Member of Technical Staff – World Models About Us Veeda AI is building the next generation of multimodal foundation world models for Physical AI. We're a small, fast-moving team of engineers and researchers from leading AI labs, tackling some of the most challenging problems at the intersection of AI, robotics, and embodied intelligence. If you're excited about pushing the boundaries of what's possible with Physical AI, you'll have the opportunity to make an outsized impact from day one. Responsibilities Generative World Model Architecture: Design, train, and scale action-conditioned video and latent dynamics models: diffusion transformers with rectified flow, block-causal autoregressive hybrids, causal video tokenizers whose compression ratio sets how far a rollout survives. Action Conditioning & Latent Actions: Condition rollouts on robot action chunks and camera trajectories, and recover pseudo-actions from unlabeled video with inverse dynamics and latent action models, so training is not capped by what teleoperation produced. Long-Horizon Stability & Memory: Attack compounding drift by training on the model's own rollouts (diffusion forcing, self-forcing) and carrying persistent scene state as 3D Gaussians or point maps, so minute-long rollouts stay coherent. Post-Training & Distillation: Fine-tune against physics-grounded reward models and verifiers built from simulator ground truth and robot trajectories, then distill to few-step samplers so an agent can act inside the model at interactive rates. Evaluation of Learned Simulators: Build evaluation for action-following, physical plausibility, and long-horizon drift, scored against simulator ground truth and by whether a policy trained inside the model transfers to a robot, not by FVD. Requirements Bachelor's degree or equivalent hands-on experience in Computer Science, Engineering, or a related technical field Experience training generative or predictive sequence models end to end at multi-node scale (video, 3D, or latent dynamics) and fluency in PyTorch from prototype to production-scale training Ability to independently design, execute, and analyze machine learning experiments, from hypothesis to ablation to conclusion Working command of the fundamentals (diffusion and flow matching, long-context sequence modeling, and understanding of their practical limits) Serious approach to evaluation and experience building metrics that influenced modeling decisions Nice to Have Publications or contributions to research on generative models for image, video, or 3D content Experience designing video tokenizers or VAEs, with understanding of trade-offs between compression and rollout fidelity Built action-conditioned or interactive world models for games, driving, or embodied agents Experience with model-based reinforcement learning or planning in a learned latent space Experience building streaming or causal video generation with KV caching and rolling context at interactive rates Worked with real robot trajectory data (LeRobot, Open X-Embodiment) and understanding of noisy action labels Contributions to open-source generative model projects or related infrastructure
AI Research Scientist- World Model
Bosch Group
Research Scientist - World Model
Luma
Senior Machine Learning Engineer (Reinforcement Learning/World Model)
Path Robotics
Researcher, World Models
Menlo
Applied Research Intern, Proactive Intelligence & Customer World Models (PhD / Graduate Co-op)
Block
Human Interactive Driving Intern – World Models
Toyota Research Institute