29 subscribers
با برنامه Player FM !
Training Large Language Models to Reason in Continuous Latent Space
Manage episode 461129295 series 3448051
LLMs have typically been restricted to reason in the "language space," where chain-of-thought (CoT) is used to solve complex reasoning problems. But a new paper argues that language space may not always be the best for reasoning. In this paper read, we cover an exciting new technique from a team at Meta called Chain of Continuous Thought—also known as "Coconut." In the paper, "Training Large Language Models to Reason in a Continuous Latent Space" explores the potential of allowing LLMs to reason in an unrestricted latent space instead of being constrained by natural language tokens.
Read a full breakdown of Coconut on our blog, or join us live for the next paper reading.
Learn more about AI observability and evaluation, join the Arize AI Slack community or get the latest on LinkedIn and X.
51 قسمت
Manage episode 461129295 series 3448051
LLMs have typically been restricted to reason in the "language space," where chain-of-thought (CoT) is used to solve complex reasoning problems. But a new paper argues that language space may not always be the best for reasoning. In this paper read, we cover an exciting new technique from a team at Meta called Chain of Continuous Thought—also known as "Coconut." In the paper, "Training Large Language Models to Reason in a Continuous Latent Space" explores the potential of allowing LLMs to reason in an unrestricted latent space instead of being constrained by natural language tokens.
Read a full breakdown of Coconut on our blog, or join us live for the next paper reading.
Learn more about AI observability and evaluation, join the Arize AI Slack community or get the latest on LinkedIn and X.
51 قسمت
همه قسمت ها
×








1 How to Prompt LLMs for Text-to-SQL: A Study in Zero-shot, Single-domain, and Cross-domain Settings 44:59

1 The Geometry of Truth: Emergent Linear Structure in LLM Representation of True/False Datasets 41:02


1 Large Content And Behavior Models To Understand, Simulate, And Optimize Content And Behavior 42:14
به Player FM خوش آمدید!
Player FM در سراسر وب را برای یافتن پادکست های با کیفیت اسکن می کند تا همین الان لذت ببرید. این بهترین برنامه ی پادکست است که در اندروید، آیفون و وب کار می کند. ثبت نام کنید تا اشتراک های شما در بین دستگاه های مختلف همگام سازی شود.