RLHF
concepts · 3 notes linked
Related: Openai · Scale AI · Large Language Models · Chatgpt · Contextual AI · Training Data · Synthetic Data · Hugging Face
Notes
- 'If journalism is going up in smoke, I might as well get high off the fumes': confessions of a chatbot helper — Human annotators writing gold-standard training data for LLMs
- Hugging Face Introduces StackLLaMA: A 7B Parameter Language Model Based on LLaMA and Trained on Data from Stack Exchange Using RLHF — Hugging Face RLHF fine-tuning of LLaMA 7B on Stack Exchange Q&A data
- Inside the lucrative, disturbing world of humans training AI chatbots — Data annotators' pay, conditions, and experiences training AI models at Scale AI and others