Remote Jobs RockRemote Jobs Rock

Research Scientist, Simulation Agents

📅 Mar 25
Imitation LearningReinforcement LearningGenerative ModelsPlanning

📜 Description

  • Own and pursue a research agenda to develop realistic and controllable simulation agents.
  • Advance the state-of-the-art in imitation learning, reinforcement learning, generative models, and planning for simulation agents.
  • Collaborate with simulation, autonomy, and safety teams to define high-impact research problems and ship solutions into production.
  • Mentor junior scientists and interns, fostering a culture of scientific rigor and rapid experimentation.
  • Publish high-impact research at top-tier conferences in machine learning or robotics.

🛠️ Requirements

  • Masters/PhD in machine learning, computer science, engineering, or a related field.
  • Strong background in imitation learning and/or reinforcement learning.
  • Publications in top-tier conferences or journals related to machine learning or robotics.
  • Proficiency with modern ML frameworks such as PyTorch, TensorFlow, or Jax.
  • Experience in self-driving, traffic simulation, or a related field is a plus.

Benefits

  • Competitive compensation and equity awards.
  • Health and Wellness benefits encompassing Medical, Dental and Vision coverage.
  • Unlimited Vacation.
  • Flexible hours and Work from Home support.
  • Daily drinks, snacks and catered meals when in office.
  • Regularly scheduled team building activities and social events.
Full job description
Waabi, founded by AI visionary Raquel Urtasun, is the leader in Physical AI. With a world-class team, we're unlocking the next era of autonomous transportation with technology that's powering commercial autonomous trucks and robotaxis. Waabi is backed by and partners with world leaders in AI, automotive, logistics, and deep tech.
With offices in Toronto, San Francisco, Dallas, and Pittsburgh, Waabi is growing quickly and looking for diverse, innovative and collaborative candidates who want to impact the world in a positive way. To learn more visit: www.waabi.ai

The Behaviors team at Waabi develops cutting-edge simulation agents and scenario generation algorithms for Waabi World, our simulation platform. As a Research Scientist on the Behaviors team, you will work closely with our multidisciplinary team of research scientists and engineers to invent the next generation of models and algorithms that power Waabi World. Your work will define the scenarios that push our self-driving system to its limits, generate the training signal that makes them better, and form a core pillar of our scientific safety case to put them on the road.
You will...
  • Own and pursue a research agenda to develop realistic and controllable simulation agents.
  • Advance the state-of-the-art in imitation learning, reinforcement learning, generative models, foundation models, planning and search, and other related areas for simulation agents.
  • Collaborate with our simulation, autonomy, and safety teams to define high-impact research problems, ship solutions into production, and drive progress towards Waabi’s milestones.
  • Mentor junior scientists and interns; foster a culture of scientific rigor and rapid experimentation.
  • Publish high-impact research at top-tier conferences in machine learning or robotics.
Qualifications:
  • Masters/PhD in machine learning, computer science, engineering, or a related field.
  • Strong background in imitation learning and/or reinforcement learning.
  • Publications in top-tier conferences or journals related to machine learning or robotics.
  • Proficiency with modern ML frameworks such as PyTorch, TensorFlow, or Jax.
Bonus:
  • Experience in self-driving, traffic simulation, or a related field.
  • Proven ability to take research from prototype to production systems.
  • Strong software engineering skills, including experience with large-scale training or simulation.
Waabi, founded by AI visionary Raquel Urtasun, is the leader in Physical AI. With a world-class team, we're unlocking the next era of autonomous transportation with technology that's powering commercial autonomous trucks and robotaxis. Waabi is backed by and partners with world leaders in AI, automotive, logistics, and deep tech.   With offices in Toronto, San Francisco, Dallas, and Pittsburgh, Waabi is growing quickly and looking for diverse, innovative and collaborative candidates who want to impact the world in a positive way. To learn more visit: www.waabi.ai The US yearly salary range for this role is: $158,000 - $269,000 USD in addition to competitive perks & benefits. Waabi US Inc.’s yearly salary ranges are determined based on several factors in accordance with the Company’s compensation practices. The salary base range is reflective of the minimum and maximum target for new hire salaries for the position across all US locations. Note: The Company provides additional compensation for employees in this role, including equity incentive awards and an annual performance bonus. Perks/Benefits: Competitive compensation and equity awards. Health and Wellness benefits encompassing Medical, Dental and Vision coverage (for full-time employees only). Unlimited Vacation. Flexible hours and Work from Home support. Daily drinks, snacks and catered meals (when in office). Regularly scheduled team building activities and social events both on-site, off-site & virtually. As we grow, this list continues to evolve!    Waabi is a technology start-up building technologies to transform the way the world moves. Join our talented team to be a part of the future and to make an impact!   Waabi is an equal opportunity employer. We celebrate diversity and are committed to creating a supportive, inclusive, and accessible workplace for all our employees. We seek applicants of all backgrounds and identities, across race, color, ethnicity, national origin or ancestry, age, citizenship, religion, sex, sexual orientation, gender identity or expression, military or veteran status, marital status, pregnancy or parental status, caregiver status, disability, or any other characteristic protected by law. We make workplace accommodations for qualified individuals with disabilities as required by applicable law. If reasonable accommodation is needed to participate in the job application or interview process please let our recruiting team know.
Thinkingmachines

Research, RL Scaling

📅 Aug 21
Thinkingmachines👥 51 - 200 employees🏢 Technology

This role focuses on scaling reinforcement learning for frontier models, requiring expertise in asynchronous RL algorithms and the integration of training and inference systems.

PythonDeep Learning FrameworksReinforcement LearningAsynchronous RL Algorithms
Wayve

Research Scientist, Robot Foundation Model

📅 Aug 6
Wayve👥 501 - 1000 employees🏢 Computer Software

Join the MEGA team as a Research Scientist to develop foundation models for general-purpose robots, focusing on intelligent agents that can perceive and manipulate the physical world.

Machine LearningVision-language ModelsVideo ModelsRobot Policies
OpenAI

Researcher, Multimodal Safety

📅 Aug 3
OpenAI👥 10,000+ employees🏢 Research Services

As a Researcher on the Chat and Multimodal Safety team, you will advance multimodal safety research, ensuring models behave safely across text, vision, and audio interactions.

Multimodal ModelsVision-language ModelsVideo UnderstandingAudio Systems

Trusted by Remote Workers