Login
Download
Skill UI
Browse and discover
15558+
curated skills
All
Development
Artificial Intelligence
Design & Creative
Product & Business
Data Science
Marketing
Soft Skills
Productivity
Engineering
Languages
Search
Superposition
, found
1
results
Default
Newest
Most Downloaded
Sparse Autoencoders for Model Interpretability
sparse-autoencoder-training
Orchestra-Research/AI-Research-SKILLs
251
SAELens provides a framework for training and analyzing Sparse Autoencoders (SAEs). SAEs decompose the dense, often polysemantic activations of large language models into sparse, monosemantic features. Use this when you need to discover the discrete, interpretable concepts a model has learned, study feature superposition, or analyze specific safety-relevant behaviors (like bias or deception) within deep neural networks.
View Details
1
Language
简体中文
English