flâneur

How Developers Steer Language Model Outputs: Large Language Models Explained, Part 2 | Center for Security and Emerging Technology

cset.georgetown.edu · saved by 2 readers

Large language models (LLMs), the technology that powers generative artificial intelligence (AI) products like ChatGPT or Google Gemini, are often thought of as chatbots that predict the next word. But that isn't the full story of what LLMs are and how they work. This is the second blog post in a three-part series explaining some key elements of how LLMs function. This blog post explores fine-tuning—a set of techniques used to change the types of output that pre-trained models produce.

saved by