about
Stories from August 4, 2023 (UTC)
Go back a day, month, or year. Go forward a day.
1. Training a Helpful and Harmless Assistant with Reinforcement Learning from Human (arxiv.org)
2 points by awyugan on Aug 4, 2023 | hide | past | pdf | 1 comment
2. L-Eval: Instituting Standardized Evaluation for Long Context Language Models (arxiv.org)
2 points by awyugan on Aug 4, 2023 | hide | past | pdf | discuss