/explore

Click through on any links that interest you or select the planets on the right to continue exploring the Outer Web.
You are here

lambdalabs.com
| | blog.moonglow.ai

Parameters and data. These are the two ingredients of training ML models. The total amount of computation ("compute") you need to do to train a model is proportional to the number of parameters multiplied by the amount of data (measured in "tokens"). Four years ago, it was well-known that if
2.2 parsecs

Travel
| |
| | lacker.io

I've been playing around with OpenAI's new GPT-3 language model. When I got beta access, the first thing I wondered was, how human is GPT-3? How close is it ...
1.7 parsecs

Travel
| |
| | gwern.net

On GPT-3: meta-learning, scaling, implications, and deep theory. The scaling hypothesis: neural nets absorb data & compute, generalizing and becoming more Bayesian as problems get harder, manifesting new abilities even at trivial-by-global-standards-scale. The deep learning revolution has begun as foretold.
2.2 parsecs

Travel
| |
| | www.soundsofthe60sblog.net

[AI summary] A blog post highlights a new 5-CD box set containing the complete Pye Recordings (1963-1967) by the band The Searchers.
57.8 parsecs

Travel
|