ycliper

Популярное

Музыка Кино и Анимация Автомобили Животные Спорт Путешествия Игры Юмор

Интересные видео

2025 Сериалы Трейлеры Новости Как сделать Видеоуроки Diy своими руками

Топ запросов

смотреть а4 schoolboy runaway турецкий сериал смотреть мультфильмы эдисон
Скачать

Sparse Autoencoders Unlearn Knowledge in LLMs | A Paper-Based Walkthrough

#Sparse Autencoders

#SAEs

#Interpretability

#AI

#ML

#Artificial Intelligence

#Machine Learning

#Alignment

#Unlearning

#Machine Unlearning

Автор: Papers Are Wonderful

Загружено: 2025-05-24

Просмотров: 4520

Описание: I made a video about one of my favorite papers! I hope you enjoy :)

===Summary===
"Applying Sparse Autoencoders to Unlearn Knowledge in Language Models" investigates using SAEs—tools that peer into the inside of LLMs—to remove undesirable capabilities from language models. In this video, I walk through the motivation of this work, the methods used, and the interesting results the authors found.

I highly recommend you read it for yourself here: https://arxiv.org/pdf/2410.19278#page...

===My other videos on Sparse Autoencoders===
Matroshkya SAEs:    • Matryoshka (Nested) Sparse Autoencoders Ex...  
SAEs from the Ground Up:    • A Window  Into LLMs | Sparse Autoencoders ...  

===Video Chapters===
0:00 Intro
0:14 Context/Motivation
0:46 SAE Negative Clamping
1:01 Feature Identification
1:35 Experimental Setup
1:49 Single-Feature Steering
2:24 Multi-Feature Steering
3:29 Investigating RMU Hypothesis

Не удается загрузить Youtube-плеер. Проверьте блокировку Youtube в вашей сети.
Повторяем попытку...
Sparse Autoencoders Unlearn Knowledge in LLMs | A Paper-Based Walkthrough

Поделиться в:

Доступные форматы для скачивания:

Скачать видео

  • Информация по загрузке:

Скачать аудио

Похожие видео

© 2025 ycliper. Все права защищены.



  • Контакты
  • О нас
  • Политика конфиденциальности



Контакты для правообладателей: [email protected]