Research Scientist, Gemini Horizon, DeepMind
Job Fit Check
Base Career helps you apply smarter for this job.
Key skills for this role
Key Skills for This Role
Full Job Posting
Responsibilities
- Improve Gemini with new Reinforcement Learning (RL) environments, evaluations, changes to the RL recipe, or ideas.
- Explore domains where Gemini should be superhuman (e.g., security, hardware, performance engineering, scientific computing, or something we have not thought of) and build the evaluations that show the gap is real.
- Invent new ways of manufacturing hard problems and the graders that make them verifiable.
- Evaluate and train Gemini on your environments, read the trajectories to see what it is actually learning, and land what transfers into production.
- Contribute to identifying and delivering breakthroughs for Gemini.
- - Improve Gemini with new Reinforcement Learning (RL) environments, evaluations, changes to the RL recipe, or ideas. - Explore domains where Gemini should be superhuman (e.g., security, hardware, performance engineering, scientific computing, or something we have not thought of) and build the evaluations that show the gap is real. - Invent new ways of manufacturing hard problems and the graders that make them verifiable. - Evaluate and train Gemini on your environments, read the trajectories to see what it is actually learning, and land what transfers into production. - Contribute to identifying and delivering breakthroughs for Gemini.
Minimum qualifications:
PhD degree in Computer Science, Artificial Intelligence, Machine Learning, a related technical field, or equivalent practical experience.
1 year experience with Generative AI, Large Language Models, natural language processing, or Agent-based systems.
1 year of experience working on modern large language model post-training (e.g., SFT, RLHF, DPO, PPO), model alignment, or core generative model development in an industry AI lab, research institute, or frontier AI organization.
Preferred qualifications:
Experience with reinforcement learning for LLM post-training.
Qualifications
- Minimum qualifications: - PhD degree in Computer Science, Artificial Intelligence, Machine Learning, a related technical field, or equivalent practical experience. - 1 year experience with Generative AI, Large Language Models, natural language processing, or Agent-based systems. - 1 year of experience working on modern large language model post-training (e.g., SFT, RLHF, DPO, PPO), model alignment, or core generative model development in an industry AI lab, research institute, or frontier AI organization. Preferred qualifications: - Experience with reinforcement learning for LLM post-training.
About Google
Technology company specializing in search, advertising, cloud computing, and AI
Visit company websiteJobs and hiring trendsApply for this job in 1 click
Skip the repetitive application forms
Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.
Trusted by over 500,000 job seekers on Base Career
More from this employer
More jobs at Google
Principal Engineer, Resource Optimization and Fleet Logic
Thornton, USA
Software Engineer III, Infrastructure, Google Cloud Performance
New York City, USA
Staff Software Engineer, Google Cloud Compute
Sunnyvale, USA
Software Engineer III, Infrastructure, Telemetry Router
Pittsburgh, USA
Staff Hardware/Software Systems Architect, Silicon
Mountain View, USA
Software Engineer III, Infrastructure, Google Workspace
Sunnyvale, USA
Strategic Negotiator, Data Center Leasing and Investments
San Francisco, USA
Senior CPU RTL Design Engineer
Mountain View, USA
