CENIA “Large Language Models Usage and Evaluation Patterns”
ABSTRACT
Large Language Models (LLMs) have changed the way computers understand and use human language. They are used in many different areas, and in this talk we will look at how people use and evaluate them. We will start by looking at the different ways people use LLMs. First, when they are used as general assistants for tasks like writing, summarizing, coding, etc. Then, when they are adapted to address more domain-specific tasks using two approaches: 1) retrieval-assisted generation, and 2) fine-tuning. We will also see how LLMs are integrated into software applications, such as when they are invoked by computer code (API calls) or used by autonomous agents to make decisions on their own. On the evaluation side, we will talk about a method called MTBench, a multi-turn question set, and Chatbot Arena, a crowdsourced battle platform between LLMs.
SPEAKER
Felipe Bravo, Associate Researcher Cenia. Assistant Professor DCC, Universidad de Chile.
WHEN AND WHERE
- In person, it will be at the Auditorio Ramón Picarte, DCC, Universidad de Chile. 3rd floor, Edificio Norte, Beauchef 851.
- Virtually, via ZOOM, the link will be shared by email and Slack that day.
REGISTRATION:
https://docs.google.com/forms/d/e/1FAIpQLSev6QO4JteuxkLduyxLP6slYYIJ92G55DKaFEFfZ-XD5aENYw/viewform

