Tuesday 6 June 2023

GPT Best Practices

GPT Best Practices
571 by yla92 | 174 comments


No comments:

Post a Comment

New exponent functions that make SiLU and SoftMax 2x faster, at full accuracy

New exponent functions that make SiLU and SoftMax 2x faster, at full accuracy 379 by weinzierl | 72 comments