با برنامه Player FM !
Synthetic continued pretraining
Manage episode 439457774 series 3524393
The paper proposes synthetic continued pretraining using EntiGraph to enhance language models' learning efficiency from small, domain-specific corpora by generating diverse text from salient entities.
https://arxiv.org/abs//2409.07431
YouTube: https://www.youtube.com/@ArxivPapers
TikTok: https://www.tiktok.com/@arxiv_papers
Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016
Spotify: https://podcasters.spotify.com/pod/show/arxiv-papers
--- Support this podcast: https://podcasters.spotify.com/pod/show/arxiv-papers/support
1653 قسمت
Manage episode 439457774 series 3524393
The paper proposes synthetic continued pretraining using EntiGraph to enhance language models' learning efficiency from small, domain-specific corpora by generating diverse text from salient entities.
https://arxiv.org/abs//2409.07431
YouTube: https://www.youtube.com/@ArxivPapers
TikTok: https://www.tiktok.com/@arxiv_papers
Apple Podcasts: https://podcasts.apple.com/us/podcast/arxiv-papers/id1692476016
Spotify: https://podcasters.spotify.com/pod/show/arxiv-papers
--- Support this podcast: https://podcasters.spotify.com/pod/show/arxiv-papers/support
1653 قسمت
All episodes
×به Player FM خوش آمدید!
Player FM در سراسر وب را برای یافتن پادکست های با کیفیت اسکن می کند تا همین الان لذت ببرید. این بهترین برنامه ی پادکست است که در اندروید، آیفون و وب کار می کند. ثبت نام کنید تا اشتراک های شما در بین دستگاه های مختلف همگام سازی شود.