500AI
Search

Tim Dettmers

  • QLoRA: Efficient Finetuning of Quantized LLMs
  • LLM.int8(): 8-bit Matrix Multiplication for Transformers at Scale

All names