Lost in the Middle: How Language Models Use Long Contexts

FromDeep Papers

Start listening View podcast show

Lost in the Middle: How Language Models Use Long Contexts

FromDeep Papers

ratings:

Length:

42 minutes

Released:

Jul 26, 2023

Format:

Podcast episode

Description

Deep Papers is a podcast series featuring deep dives on today’s seminal AI papers and research. Each episode profiles the people and techniques behind cutting-edge breakthroughs in machine learning. This episode is led by Sally-Ann DeLucia and Amber Roberts, as they discuss the paper "Lost in the Middle: How Language Models Use Long Contexts." This paper examines how well language models utilize longer input contexts. The study focuses on multi-document question answering and key-value retrieval tasks. The researchers find that performance is highest when relevant information is at the beginning or end of the context. Accessing information in the middle of long contexts leads to significant performance degradation. Even explicitly long-context models experience decreased performance as the context length increases. The analysis enhances our understanding and offers new evaluation protocols for future long-context models. Full transcript and more here: https://arize.com/blog/lost-in-the-middle-how-language-models-use-long-contexts-paper-reading/Follow AI__Pub on Twitter. To learn more about ML observability, join the Arize AI Slack community or get the latest on our LinkedIn and Twitter.

Released:

Jul 26, 2023

Format:

Podcast episode

Titles in the series (22)

Deep Papers is a podcast series featuring deep dives on today’s seminal AI papers and research. Hosted by AI Pub creator Brian Burns and Arize AI founders Jason Lopatecki and Aparna Dhinakaran, each episode profiles the people and techniques behind cutting-edge breakthroughs in machine learning.

Skip carousel

More Episodes from Deep Papers

Skip carousel

Related podcast episodes

Skip carousel

Discover this podcast and so much more

Lost in the Middle: How Language Models Use Long Contexts

Lost in the Middle: How Language Models Use Long Contexts

Description

Titles in the series (22)

More Episodes from Deep Papers

Related podcast episodes