Starling-7B: UC Berkeley’s New Open-Source LLM
How do you copy GPT-4 without actually copying the model weights? In this article, you’ll learn how! π‘ Researchers from UC Berkeley have unveiled Starling-7B, an innovative large language model (LLM) trained using Reinforcement Learning from AI Feedback (RLAIF), as opposed to the Reinforcement Learning from Human Feedback (RLHF) approach used by many competitors. Starling-7B … Read more