|
You are here |
www.lesswrong.com | ||
| | | | |
www.alignmentforum.org
|
|
| | | | | A new Anthropic interpretability paper-"Toy Models of Superpostion"-came out last week that I think is quite exciting and hasn't been discussed here... | |
| | | | |
goodfire.ai
|
|
| | | | | Goodfire is an AI research company building practical interpretability tools for safe and reliable generative models. | |
| | | | |
deepmind.google
|
|
| | | | | Announcing a comprehensive, open suite of sparse autoencoders for language model interpretability. | |
| | | | |
www.nicktasios.nl
|
|
| | | In the Latent Diffusion Series of blog posts, I'm going through all components needed to train a latent diffusion model to generate random digits from the MNIST dataset. In the second post, we will bu | ||