Train transformer language models with reinforcement learning.
Reading the pull requests.
0 / 613 merged PRs · 90 days
No AI-agent authors in this read.