Description
This research engineer position focuses on aligning artificial intelligence models using Reinforcement Learning from Human Feedback (RLHF), Constitutional AI, and red-teaming strategies. The ideal candidate will have prior experience with PyTorch and large-scale training.
