500AI
Search

Eric Mitchell

  • Direct Preference Optimization: Your Language Model is Secretly a Reward Model
  • Meta-Learning Online Adaptation of Language Models
  • Enhancing Self-Consistency and Performance of Pre-Trained Language Models through Natural Language Inference

All names