Back to search
Scale AI Builtin · Indexed 2026-08-27

Machine Learning Research Scientist, Evaluations

Remote / flexible 181K-226K Annually

Entry level Builtin
Continue to application Add your email once, then Caio opens the original posting.

Indexed description

Scale AI Machine Learning Research Scientist, Evaluations 4 Hours AgoSaved In-Office 3 Locations 181K-226K Annually Entry level 181K-226K Annually Entry levelArtificial Intelligence • Big Data • Machine LearningConduct research on evaluating frontier large language models and AI agents. Analyze model behavior, diagnose capability, reasoning, robustness, and alignment failures, and develop benchmarks for text and multimodal systems. Apply post-training techniques such as supervised fine-tuning, RLHF, and reward modeling to connect failures with training interventions. Collaborate with AI labs, define evaluation best practices, translate findings into technical strategy, and publish research at major conferences.Top Skills: Artificial IntelligenceDeep LearningGenerative AiInstruction TuningLarge Language ModelsMachine LearningMultimodal AiPreference ModelingReinforcement LearningReward ModelingRlhfSupervised Fine-Tuning

Free. 20 seconds. No password. See every match in this search.

Create a free Caio profile to unlock more results and save your role and location preferences.

Unlock free search
Want help applying to roles like this? Search Caio for free. If repetitive applications get heavy, Managed Job Search adds supervised execution for $99/month.
View Managed Job Search