Bachelor’s, Master’s, or PhD degree in Computer Science, Machine Learning, Robotics, or a related technical field
4+ years of experience in machine learning, or a PhD with 2+ years of relevant industry experience, with a strong focus on model training
Strong programming skills in Python and hands on experience with modern ML frameworks such as PyTorch or JAX
Strong understanding of LLM post-training techniques, including supervised fine-tuning, reinforcement learning, and on-policy distillation
Experience developing, evaluating, and deploying machine learning models in production environments
Strong research and problem solving skills, with the ability to work effectively in ambiguous, fast moving environments
Publications, open source contributions, or demonstrated research impact in LLM post-training or agent learning
Experience with distributed training, GPU acceleration, and large scale model training systems
Experience leading technically complex projects or mentoring other engineers
We strive to use the best tool for the job when building and deploying our production services. Sometimes that means writing our own custom code, and often it means leaning on the work of others. As part of building Serverless RL, we depend on the following libraries and frameworks (among many others)
We work hard, have fun, and move fast! We’re in an exciting stage of hypergrowth that you will not want to miss out on. We’re not afraid of a little chaos, and we’re constantly learning. Our team cares deeply about how we build our product and how we work together, which is represented through our core values
We support and encourage an entrepreneurial outlook and independent thinking. We foster an environment that encourages collaboration and provides the opportunity to develop innovative solutions to complex problems. As we get set for takeoff, the growth opportunities within the organization are constantly expanding. You will be surrounded by some of the best talent in the industry, who will want to learn from you, too. Come join us!