Introduction to Mechanistic Interpretability
بایگانی مجموعه ها ("فیدهای غیر فعال" status)
When? This feed was archived on February 21, 2025 21:08 (
Why? فیدهای غیر فعال status. سرورهای ما، برای یک دوره پایدار، قادر به بازیابی یک فید پادکست معتبر نبوده اند.
What now? You might be able to find a more up-to-date version using the search function. This series will no longer be checked for updates. If you believe this to be in error, please check if the publisher's feed link below is valid and contact support to request the feed be restored or if you have any other concerns about this.
Manage episode 458945499 series 3498845
Our introduction introduces common mech interp concepts, to prepare you for the rest of this session's resources.
Original text: https://aisafetyfundamentals.com/blog/introduction-to-mechanistic-interpretability/
Author(s): Sarah Hastings-Woodhouse
A podcast by BlueDot Impact.
Learn more on the AI Safety Fundamentals website.
فصل ها
1. Introduction to Mechanistic Interpretability (00:00:00)
2. Why might mechanistic interpretability be useful? (00:01:16)
3. Looking inside neural networks (00:03:34)
4. What makes mechanistic interpretability hard? (00:06:33)
5. Addressing polysemanticity (00:08:34)
85 قسمت