logoalt Hacker News

pfdietzyesterday at 4:28 PM1 replyview on HN

> It’s far larger than the resulting model.

Is it? How many different books are we talking about, and how much information is that, after conversion to text and lossless compression? Images, maybe, but text?


Replies

dparkyesterday at 5:48 PM

These models are trained on way more than just books. GPT-3 was trained on about half a terabyte of filtered plaintext and the training corpuses have grown significantly by then by all accounts.

show 1 reply