logoalt Hacker News

elzbardicotoday at 2:08 AM0 repliesview on HN

LLMs are an imprecise, more of a marketing term, to define Transformer models based on the self-attention mechanism, trained with massives amounts of data.

And this implements a transformer. Actually it is a very cool didactic example.