Tuesday 17 January 2023

Let's build GPT: from scratch, in code, spelled out by Andrej Karpathy [video]

Let's build GPT: from scratch, in code, spelled out by Andrej Karpathy [video]
495 by georgehill | 46 comments


No comments:

Post a Comment

New exponent functions that make SiLU and SoftMax 2x faster, at full accuracy

New exponent functions that make SiLU and SoftMax 2x faster, at full accuracy 379 by weinzierl | 72 comments