Back to search
Pika Builtin · Indexed 2026-09-14

Research Scientist, Post-Training — Video Generation

Palo Alto, CA, USA 185K-400K Annually

Builtin
Continue to application Add your email once, then Caio opens the original posting.

Indexed description

Pika Research Scientist, Post-Training — Video Generation 22 Minutes AgoSaved In-Office Palo Alto, CA, USA 185K-400K Annually Junior 185K-400K Annually JuniorInformation TechnologyResearch Scientist responsible for RL-based post-training of large-scale video generation models. The role includes preference optimization, online RL, video reward model development, human and automated evaluation, reward-hacking safeguards, and secondary distillation of aligned models. Candidates need experience with generative modeling or post-training, reinforcement learning, diffusion or flow-matching models, PyTorch, and distributed training. Preferred experience includes visual-generation reward models, VLM-based judging, preference data collection, and video-specific temporal and physics failure modes.Top Skills: Diffusion ModelsFlow-Matching ModelsModel DistillationMulti-Node Distributed TrainingPreference OptimizationPyTorchReinforcement LearningVideo Reward ModelsVlm-As-Judge

Free. 20 seconds. No password. See every match in this search.

Create a free Caio profile to unlock more results and save your role and location preferences.

Unlock free search
Want help applying to roles like this? Search Caio for free. If repetitive applications get heavy, Managed Job Search adds supervised execution for $99/month.
View Managed Job Search